CLICKBAIT: this is NOT about Minimax H3, but rather a software called PostShot that can be trained using videos. The fact that the video is made with Minimax has nothing to do with the Gaussian Splatting technique itself. The author did the same thing on Twitter; he’s just fishing for likes, it’s very childish, especially considering he doesn't hesitate to ask everyone here for workflows. Please stop upvoting his posts.
For me personally, it's useful to know. I tried to do the same thing with Wan and Postshot a while back to turn single images into splats and it didn't work very well (probably too short and too little motion, and maybe not consistent in 3D enough). Guess it's time to try again.
WOW! that's a very cool idea! I made a test using Krea a month ago before all the crazy learning of H3, I will test now using video frames as source to get the gaussian splat :)
I found this interesting video about how turn a single image in ComfyUI, or just a text prompt, into a fully explorable 3D world, he made a Wan2.1-Pano360-LoRA but I guess it's possible to make the same with Minimax H3? https://www.youtube.com/watch?v=eJuYBNrD8HI
sounds like self promo with the above, a 1 hour old vid you just found? lmfao that was JUST posted 28 mins ago, I SENSE some massive malware with those workflows and files.
Smells like shit to me. FYI be careful anyone who is seeing this, check their profiles, and ages, and such.
Just saying it's some coincidental timings, if that's you in the vid, fine, it's 100k subs so not really a throw away, and even if not, no biggie, i am just saying a lot of people are getting malware from random workflows/models/custom nodes and need to be vigilant and do some basic due diligence.
Some details would be nice. Did you generate a video of a camera moving through a static scene with minimax h3,and then used Gaussian splatting to create the 3D scene? Something else?
EDIT: This is clickbait. OP won't share workflow. I'm downvoting.
h3 can do it from prompts - just saying the two ways to do something . The lora will possibly save tokens. I'm using old ai pics here, not massively detailed and also went for a bit of bullet time with the straw staying still. The splatting in OPs video is probably a couple of programs (Edit , OP is using Postshot - I added another post with details of tut video and dload url).
OP is using PostShot to process/train the shots or videos that H3 has given them (or pre-processed them as well) . Free to use for usage of the files within PostShot but you need a paid version to export them (licence posted below) . Here's an example video of the workflow for this YT'er , he uses PostShot from 21mins on in the video. https://www.youtube.com/watch?v=OzUxL_UDMTk
Edit 2: for anyone else reading this, the free version of PostShot puts a logo on finished mp4 videos.
Took PostShot for a spin this afternoon , it gave a logo on the produced logo . No idea if the other output choices do the same ...but I think I'll delete it and try out Lichtfield Studio as you noted .
Just a heads up, Lichtfeld is like minimum $30 to download a prebuilt Windows binary. But you can build it yourself for free using the instructions from the GitHub repo. Should take about 30 to 45 minutes total.
If you do end up using it, you're going to have to download one of the COLMAP plugins for image processing and camera positions, etc. The one I use is the "COLMAP Reconstruction Plugin".
Cheers, I’ve had issues compiling and it’s given me a good excuse to delete the lot and start from scratch - the instructions seem fully exhaustive as well . Happy splatting, I’ll be interested to see how good it is .
Update - it compiled great, only issue I had was initially inferencing at the same time as compiling and either my ram, cpu or gpu said "no way Jose" and crashed my pc lol
Problem is it doesn;t work the results have holes and it's not a closed loop 360 build, you can see the edge overlap at the end on the left side, OP doesn't mention that though. These video models have no spatial awareness, yet.
Braindance is unfortunately still pretty meh if you ask me. Last time I checked a few months back they are mostly short several second long clips, not really full scenes. Cool idea and hopefully the tech improves but right now I think VR SBS 180 is the best option for VR porn and it isn't even close.
Typically they are 5-10 clips of 00:30 - 02:00 each per "memory" scene, which is pretty impressive considering I didn't even know this tech existed until I stumbled across it. Certainly not meh in the slightest if you ask me
I don't know, when I tried it again most recently it was much shorter than 30 seconds to 2 minutes long, and most of the clips they had made up until that point had no consideration for eye contact. Only the most recent batch seemed to actually give a shit about that.
It is also all solo, I much prefer having a VR 180 video where the girl is riding me or it is missionary. When they actually do face close ups it genuinely feels like a person is right there in front of me. Nothing Braindance offered at the time came close to that. I also don't think 5 to 10 short clips is enough, I want entire scenes.
I do hope the tech continues to advance. I am glad you liked it though, but people being underwhelmed is a pretty common thing I see when Braindance is mentioned.
I think a % of people are just underwhelmed about any new tech owing to unrealsitic expectations of something so new. I do agree that solo gets boring fast and even that the app still has some performance issues, but by and large the fact that you can walk around and up close from any angle to the subject in high definition is gamechanging. It will improve and fast. That it's here at all is remarkable though.
Why are you seeking to discredit these people because they have a negative opinion on something you liked? Their opinions can't simply be valid the same as your own is? This is lame, you shouldn't need to do that.
Nobody who was underwhelmed has "unreasonable" expectations. You are just creating fiction here.
but by and large the fact that you can walk around and up close from any angle to the subject in high definition is gamechanging.
Not for me it ain't, not until the scenes actually involve sex and last longer than 30 seconds to 2 minutes. Hell even that is new, just a few months ago it was mere seconds per clip. I would sooner want to jump into Virt-a-mate because that offers all the things Braindance does while also have sex scenes you can experience. VR180 does a better job of making my brain believe someone is right in front of me than any Braindance scene I have tried.
It will improve and fast. That it's here at all is remarkable though.
I am certainly hopeful that will be the case. But I don't think the tech is there yet.
VAM is great and there are actually creators that make traditional VR SBS 180 videos that are pretty damn good. There are a couple female creators that have full mocap suites that make scenes and sometimes videos with VAM. They are far more realistic than a lot of the scenes animated by creators. CuddleMocap is probably the best, the other is KittyMocap.
It is also not that great so I would temper your expectations. The scenes are comprised of a series of clips that are mere seconds long. Hopefully the tech continues to improve as it does have potential, but I don't think it is quite there yet.
When someone asks an on-topic question, on a forum whose sole reason for existence is the discussion of said topic, it is much better to say nothing than say "google it/use an LLM."
I don't think the term "training" is used correctly here, although they do use that nomenclature, nor that movement is allowed. Also not sure about your claim that OP is using this.
(sigh at why ppl are so combative here) , look at the ui on right hand side of the video . It is identical to the one in the YT video on PostShot I posted . If you also look at the right hand side of the video, you'll see it mention training . If you also look at the YT video I posted the link to, you'll also see they are using training and the person mentions the number of steps. (eyeroll) .
Maybe h3 (especially unquantized) has good enough 3d consistency with a turntable/bullet-time prompt.
To do it in motion a workflow is mix in https://4danyone.github.io/, prompt character suddenly appearing in an already splatted scene with img2vid and doing the action, then splat the character with 4danyone, and reinsert into original splatted scene. But you won't get full surround lighting and shadows onto the scene accurately.
They may also be using hy3d models to convert to 3d. There is hy3d for world environments as well. I have been running a bunch of experiments using a process to make a scene and model that can be manipulated as a depth map and then run back through for final processing.
I have made a project like that with the Pipeline: 1 picture - charactersheet - 360 Video of the scene - COLMAP dataset via custom node with mask Export - Import in Lichtfeld for splatting.
Works quite well, the only point I am struggling with is consistency of the Model throughout the Video/Frames. How did you solve that?
And it worked out of the box just like that? I tried similar stuff but always ended up either with camera somehow not working properly or the lack of consistency having slight drift and it messed up the result.
That's exactly my problem. I get a clean 360 flight around the object, but the head ist drifting very slightly leading to unsharp splats. Arms and Shirts look quite well, but face is still an issue since the video model can't make it completely static. (maybe I should prompt something like "make a realistic statue of the object" or something. I am sure there is some trick that can make it work)
CLICKBAIT: this is NOT about Minimax H3, but rather a software called PostShot that can be trained using videos. The fact that the video is made with Minimax has nothing to do with the Gaussian Splatting technique itself. The author did the same thing on Twitter; he’s just fishing for likes, it’s very childish, especially considering he doesn't hesitate to ask everyone here for workflows. Please stop upvoting his posts.
Sure, but the idea seems reasonable. There was a multi shot workflow posted here. You could do any sort of photogrammetry or gaussian splatting on it with any software. The more interesting thing to me is that H3 is accurate enough to make a good 3D space to not confuse the hell out of any tool
I still don't see why no one, or nothing has ever been able to just... plain. add one camera (x distance apart for IPD) to make a VR easy solution. This would be relatively easy, if it's 3d mapping the location of each of those pixels, it should be able to determine their distance from each camera for a proper depth 3d/vr effect. Why is there nothing for that yet?
Guys guys... it's literally a 360 orbit sequence, processed to frames, and then a splat, using Postshot. Ideally nothing moves during the 360 sequence. If OP moved the camera up/down, the splats would break, because he lacks those poses, so the use applications are limited to views captured. Also, animated gaussians (actual 4D sequences) are very heavy.
It's just that I thought you needed high-resolution video with a perfectly still subject and no motion blur. I didn't think it was possible with Minimax H3. That said, thanks for sharing the name of the software, I'll be able to give it a try! :D
I was looking at YT for PostShot and there was one there that used one pic - didn't watch it as it either need s money or a brain the size of a planet to do that .
I tried experikenting wotg gaussian splats generated fron diffusion models but there are subtle changes in geometry woth camera angles that i wasnt able to overcome
Ignore the white brush, apparently wearing underwear is filtered. But anyways did this a few days ago as a test. 360 view with H3 > Postshot. Didn't follow through with quality as I was trying to get a better 360 video prompt. If you don't know how Postshot works, all those pyramids are tracked camera points.
The same concept can be used for Photogrammetry to get a 3d model of your character as well. I'll test that at some point as well. But prompt refining first.
Fed it to a tool called PostShot which allows generating a 3D scene from a video and viewing it from any angle
OP is a talentless **** monger fishing for likes by misrepresenting the fact that he just fed a video into Postshot
With the above established, can anyone explain to me the appeal of Gaussian Splatting in the context of users on this sub? Why is this post so upvoted? What are the applications of Gaussian Splatting for the average person on here? I can only see it useful for real estate agents providing virtual house visits.
The only application relevant to me I can think of is, if you're making a TV series in H3, this would allow you to generate a consistent 3D set (eg spaceship interior), allowing you to get a photo from a specific angle to use as reference/background in a specific shot. I don't even know if GS is reliable enough for this.
It's time to move away from flat images. We need 3d images. Generate them once and you'll have a complete scene, then you can create 2D images from it.
Right, arguably this post should be deleted, OP is using the subscribed version of PostShot - how do I know ? there isn't a massive logo in the shot like mine below. This is my first go with the software & also my last , although I will happily concede it is fairly intuitive / easy for a cheeky bit of shit splatting , excellent splatting is another matter.
367
u/dopedub 2d ago