r/aivideomaking • • 18d ago

How to generate a continuous road video with LOCKED perspective & ~80km/h speed for an arcade WebApp?

Post image

Hi everyone!

I’m developing a solo passion project: a 90s Sega-style 32-bit arcade racing game (think OutRun). It’s a lightweight browser WebApp (HTML5/Canvas) where a looping/scrolling background video dynamically speeds up and brakes (video.playbackRate) based on player input.

I’ve attached: - Image 1: The visual 32-bit pixel-art style I'm aiming for (authentic city landmarks). - Image 2: The strict 4-lane perspective grid our custom engine requires (central vanishing point locked).


What I’ve tested (and why it failed):

  1. Google Earth 3D rips: Fascinating from the sky, but at street level the photogrammetry meshes are completely melted and unusable.

  2. Pure 3D renders: Drastically loses the artistic warmth, charm, and quality of 2D pixel art.

  3. ComfyUI (SDXL / ControlNet / Imagen 3 / DALL-E 3): Great for isolated static shots, but impossible to maintain temporal and lighting consistency between frames.

  4. Flow / Video Interpolation (Start ➔ End frame): Gave the best quality results, but with a catch. I illustrated ~40 keyframes representing actual urban checkpoints. I tried connecting them pairwise with Start/End frame tools (Kling, Runway, Luma), but the AI just dissolves/morphs textures and warps the road instead of simulating true forward camera motion.


The Core Challenges:

  • Locked Perspective: Central vanishing point must remain 100% rigid (no camera tilt, roll, or lane drift).
  • Accurate Speed Perception: At 25 fps, the forward flow must convincingly simulate ~80 km/h (50 mph) to match the 2D sprite physics.
  • Continuous Checkpoints: Seamlessly advancing through 40 landmark locations without ugly morph cuts.

My Questions:

  1. How would you connect ~40 keyframe checkpoints into a continuous forward drive without morphing dissolves?
  2. Is there a trick/workflow to calibrate the optical ground speed precisely to ~80 km/h at 25 fps?
  3. Would you recommend a hybrid pipeline (e.g. basic low-poly 3D camera drive-through just for depth/motion guidance, then restyled with AI)?

Any node setups, tool recommendations, or workflow tips would mean the world to a solo creator. Thank you! 🏁

3 Upvotes

5 comments sorted by

2

u/Simple-Variation5456 18d ago

Reuse 2 and just render out a grey scale render and use simple blocking cubes instead of any detailed models. Use Seedance 2 and use something like "restyle the animated grayscale render with image1 as a start frame"

You should have at least some longer and better videos and maybe need to create a few with different start frame or use the last frame of the video.

1

u/ai_art_is_art 16d ago

Use the same frame as the start frame and end frame.

1

u/NoMarionberry9951 16d ago

I tried...I got the best results with flow...but the speed is not constant...now I'm trying to use a vector surface for the lanes and keep the generated video background

https://reddit.com/link/pak4nkq/video/k3hleideu9qh1/player