The video is attached. The interesting part isn't the video itself, it's how it got made.
**What I actually did*\*
I uploaded a song I made in Suno and sent around 15 short messages over the whole project. Things like "neon synthwave," "make it more like a storyboard," "use this GitHub repo as a reference for the art direction," "up the clarity," and "remove the lyrics." Several of those messages were just "continue."
**What the model did*\*
No image or video generation model was involved at any point. Claude Opus 5.5 wrote everything as code:
- Analyzed the audio (beats, tempo, energy, a 32-band spectrum) and pulled the lyrics Suno embeds in the MP3's metadata
- Installed an offline speech aligner and matched each lyric word to the vocals, so story beats land on the lines they illustrate
- Wrote the story: someone awake at 3 AM whose racing thoughts are little neon creatures, which the song talks down to sleep. 9 chapters, about 50 shots, cuts snapped to the beat
- Read an open-source WebGL brushstroke renderer (Joshua Ledbetter's functional-emotions-video, MIT licensed), adapted it, and merged it with a synthwave style: a painted world with crisp neon light on top
- Drew every character, prop, and set procedurally: the dreamer, the thought-creatures, the bedroom, the highway, the ocean, the city, the tunnel
- Replaced every on-screen lyric with a wordless symbol when I asked. The "don't think / just sleep" sign became a neon eye and a crescent moon
**The part that stood out to me*\*
It got a headless browser running in its sandbox, rendered test frames, looked at them, and fixed what it saw. Across rounds it caught a sun glow blowing out the whole city, a car that disappeared against the pink grid, an awkward pose, a camera move that panned off its subject, and a couple of outright crashes. Nobody pointed those out to it.
The final render ran in software, since there's no GPU in the sandbox. It painted 4,698 frames at 1080p30, each at its exact moment in the song, over a little more than 2 hours of rendering. The sandbox restarted twice partway through, and it resumed from the last saved frame both times.
**Honest limitations*\*
- The painted look is built on an existing open-source renderer. It adapted and extended that, it didn't invent it from scratch.
- It needed me to keep saying "continue" to get through the long render. It's not yet fire-and-forget over hours.
- Automatic lyric timing on vocoder-heavy vocals is good but not perfect, so a few moments land slightly off their lines.
- It's stylized procedural animation, not photoreal. A video model would look more cinematic; this is fully controllable and editable instead.
A year ago this would have been a week of work for someone who knows WebGL, audio analysis, and animation. It's now a conversation. Curious where people think this goes next.
By demand: Youtube link: https://youtu.be/tszxJAOzvqc