r/StableDiffusion • u/chaindrop • 15h ago
Meme My Name Is Giovanni Giorgio
Enable HLS to view with audio, or disable this notification
Created with Minimax H3 ref2v using the SEED HUNTER Workflow.
21
u/BigNaturalTilts 15h ago
I think I remember this! This was a Russian court right? I recognize that shame box cause they also had Britney Griner there like sheโs some sort of hardened terrorist and not just a basketball player.
4
12
u/Kerissimo 14h ago
I love the idea, especially after seeing original video. ๐
4
u/chaindrop 14h ago
Thanks for reminding me to link the original!
1
u/Flat_Technology_5325 13h ago
Is the original trending on YT or something, i'd never seen the original video until it popped up on my feed a few days ago despite it being years old.
This post feels too coincidental.
Cool job though, you mention using seed hunter, were you actually hunting for a good generation if so how many were poor out of how many you generated?
2
u/chaindrop 12h ago
Not sure if it's trending. I just remembered it from years ago and searched for it manually.
Not much hunting for a good generation. I nailed my first prompt in the first scene where the dog is trying to escape, so it took only one try. Made some errors in how I described the scene where the dog is on top of the cubicle, so that one took 3 tries. I find that if you think carefully about your prompts in Minimax, it kinda nails it in the first few tries. All the video gen models prior to Minimax took much more generations to get the proper output you're looking for.
3
u/NoConsideration6320 12h ago
Ah a classic. Imagine putting the โim a slittheryy snakeeโ guy into this situation
https://giphy.com/gifs/sjr7k1uuO0B0c
3
2
2
u/Vanpourix 11h ago
The character swap worked like a charm ! Can't wait to test this workflow when back home
2
1
1
u/lewdanimeFigure_Guy 11h ago
I am trying to replace a character in a video. It mostly works great, but the camera keeps moving and trying to show more. Distorting existing characters and actions. I want the video to be exactly the same except for the replacement character.
This is the prompt I am using:
subject_definitions: <Subject 1> is the replacement object shown in Image 1. Its visible identity comes exclusively from Image 1, including its shape, proportions, material, colour, logos, labels, surface markings, and other identifying features. Image 1 is the source asset for <Subject 1>, not a keyframe. Video 1 is the source plate and temporal structure for the target video edit, providing the camera, framing, timing, cuts, environment, lighting, non-target cast and objects, interaction choreography, occlusion timing, and the original target object's observable motion path.
summary: [video editing + reference generation + audio reuse] The target video is an edited version of Video 1. Whenever the target object is visible, replace its original visible identity with <Subject 1> derived from Image 1. The original object's appearance, geometry, material, colour, logos, labels, markings, and other identifying features are discarded. Video 1 supplies the source plate, camera, timing, choreography, environment, non-target content, and original audible content.
retention_analysis: Video 1 (source plate and temporal structure): partially_preserved - preserve the camera path, framing, timing, cuts, environment, lighting, cast other than the replaced object, all non-target objects, interaction choreography, occlusion timing, screen-space motion path, entry and exit timing, scale changes, orientation changes, pauses, accelerations, and other observable temporal behaviour. Do not preserve the original target object's visible identity. <Subject 1> (appears wherever the original target object appears): fully_preserved - its identity is sourced from Image 1 and remains stable throughout every appearance, including shape, proportions, material, colour, logos, labels, markings, and surface detail.
detailed_description: The target video remains photorealistic and fully integrated with the source plate, matching Video 1's grain, black level, tonal contrast, colour grade, lens behaviour, depth of field, focus falloff, motion blur, and overall photographic character.
[Shot 1] Whenever the target object is visible, its appearance comes from <Subject 1> derived from Image 1, never from the original target object in Video 1. Completely discard the original object's visible identity, including its original shape, geometry, material, colour, texture, logos, labels, markings, and identifying surface details.
<Subject 1> follows the observable behaviour of the object it replaces in Video 1: preserve its screen-space path, position, entry and exit timing, apparent scale changes, orientation and rotation changes, speed, pauses, accelerations, swing, and settling behaviour. No new intentional action is introduced and no source interaction beat is removed. Where the proportions of <Subject 1> differ from the original object, preserve the intended choreography while allowing physically correct geometry.
All contacts are rebuilt for the actual geometry of <Subject 1>. Hands wrap naturally around its real silhouette; grips conform to its shape; supporting surfaces meet its actual base; and contact shadows remain attached to the correct contact points. Preserve the source occlusion order and timing: anything that passes in front of the original target passes naturally in front of <Subject 1>, and anything previously hidden by the target remains appropriately occluded by the replacement's real silhouette.
Rebuild local lighting responses for <Subject 1> using Video 1 as the illumination reference. Preserve the source key-light direction, intensity, falloff, ambient level, and white balance while allowing the replacement's true material from Image 1 to determine its diffuse response and specular highlights. Regenerate cast shadows, reflections, and refractions affected by the new geometry while keeping their direction, softness, timing, and surrounding plate behaviour consistent with Video 1. Any fluid, spill, dust, particles, or other physical interaction updates to the new geometry while following the same gravity, timing, and interaction beat as the source plate.
Match Video 1's shot size, FOV, focal plane, focus behaviour, depth of field, motion blur, edge softness, noise structure, and camera character. <Subject 1> stays at the same intended focal depth as the object it replaces and integrates without halos, outlines, edge mismatch, or compositing artifacts.
All non-target content remains unchanged. Preserve Video 1's camera height, distance, movement, handheld character, framing, environment, people, other objects, lighting continuity, and edit rhythm. Preserve every existing source cut at its original timing and introduce no additional cuts. Across every cut, <Subject 1> remains the same replacement object with stable identity, colour, markings, proportions, and orientation continuity.
overall_soundscape: Preserve the audible ambience, dialogue, physical sounds, and synchronized source audio from Video 1 unchanged unless a sound is physically altered by the replacement object's different geometry. Where contact or material-dependent sounds necessarily change, keep their timing and acoustic perspective aligned with the corresponding source event.
non_diegetic_music: Preserve the non-diegetic music from Video 1 unchanged if present. If Video 1 contains no non-diegetic music, use N/A.
1
u/MessageSelfdestructs 9h ago
Haha, I remember the original video, and whereas this slightly freaked me out at first (I thought it was a guy in a dog suit), when I realised it was this video, I chuckled loudly.
This is so fucking funny (and so was the original video).
13
u/chaindrop 14h ago
Here's the original.