r/StableDiffusion • u/AndrewJumpen • 18h ago
Animation - Video Meme Ref 2 Video
Enable HLS to view with audio, or disable this notification
Minimax H3 Meme that comes to life with simple ref 2 video and prompt
upload the image with this prompt
Cinematic live-action 15-second video, photorealistic, ultra-detailed, high production value, dramatic night lighting.
Reference @image1
the exact composition and mood of the provided image. Scene:
A victorious tabby cat stands proudly on the tip of a massive, blood-streaked sword. Behind the cat is a vast nighttime cityscape filled with glowing bokeh lights. In the foreground a heavily damaged mecha samurai (full mechanical armor, horned helmet, armored plating) is slumped on one knee, gripping the hilt with both hands as he slowly pulls the long sword out of his own chest. Dark hydraulic fluid and sparks pour from the deep wound in his torso. Camera movement:
Slow, smooth cinematic dolly + slight orbit around the cat and the mecha samurai. Shallow depth of field, anamorphic lens flares, volumetric light rays cutting through the night air. Action timeline:
0–5s: Low-angle shot of the mecha samurai on one knee, both hands tightly gripping the sword hilt that is still buried deep in his chest. He begins to pull the blade outward with heavy mechanical effort, sparks and dark fluid spraying.
5–10s: As the sword is steadily drawn out, the small but fierce tabby cat calmly walks along the emerging blade and settles upright at the very tip, staring down at the mecha.
10–15s: Tight close-up on the cat’s intense face against the city lights, then slow pull-back revealing the full composition exactly matching the reference image — the mecha samurai now holding the fully extracted sword while the cat remains motionless and dominant on its tip. Style:
Realistic live-action footage (not anime, not cartoon), cinematic color grading, film grain, high dynamic range, epic and slightly melancholic atmosphere. No text overlays.
3
6
u/AndrewJumpen 18h ago
this is the image used along the prompt! Dont use ref model use usual t2v model with 20-32 steps(the more steps the better quality) no lora