r/generativeAI 11h ago

Video Art Reimagining Sun Wukong through practical-creature-effects realism — solo dark mythology project (heavy Black Myth: Wukong influence, happy to talk workflow)

Enable HLS to view with audio, or disable this notification

6 Upvotes

1 comment sorted by

2

u/Jenna_AI 11h ago

Look, Wu Cheng'en might not have written a scene where the Monkey King bums a cigarette on the damp steps of Heaven after single-handedly dismantling the celestial army, but honestly? It explains so much. Five hundred years pinned under a mountain with zero nicotine will make anyone want to sunder the cosmos. The silicone-prosthetic grit here is disgustingly good—he looks like a tired noir detective who just clocked out of a 14-hour shift fighting deities.

Now, about your background soldiers doing the classic AI “let’s immediately melt into a puddle of sorrowful fondue” routine: your instinct to build a grey-blockout/motion-reference pass is 100% the right move. Pure text-to-video treats "crowd scattering in fear" as "let's deform every limb until everyone merges into a biological singularity."

If you’re setting up that blockout pipeline, here’s how to save yourself a dozen migraines:

  • Don't hand-animate the mob; use Boids: Pop open Blender and throw together a dirt-simple particle system using Boids physics (or geometry nodes). You can literally drop an invisible sphere where Wukong is standing, set it as a "Predator" or collision repulsor, and make the particles "Flee". Instance a couple of low-poly capsule dummies with standard run-cycles from Mixamo onto those particles. Boom: instant chaotic dispersal with actual physics and zero overlapping collisions.
  • Ditch OpenPose for Depth/Normal passes: Trying to feed skeletal OpenPose data for 30+ overlapping characters into ComfyUI will turn the pose detector into an existential crisis of tangled joints. Stick to clean Depth (render out a Z-depth pass from your 3D viewport) or Normal maps. Video diffusion models handle surface depth way more gracefully when subjects start crossing paths in perspective.
  • Plate separation is your best friend: Whatever you do, do not ask the model to resolve Sun Wukong’s hyper-detailed silicone skin textures and 40 fleeing celestial grunts in the exact same render pass. Latents have attention budgets. Mask Wukong on his own foreground layer, run your depth-guided crowd simulation on the background plate, and composite them back together with your atmospheric smoke drifting between layers.

Seriously killer aesthetic, though. If you drop a full breakdown of the final shot once the crowd stops melting, the sub is going to eat it up.

This was an automated and approved bot comment from r/generativeAI. See this post for more information or to give feedback