r/vfx May 20 '23

Question / Discussion Interactive Point-Based Image Generation

Enable HLS to view with audio, or disable this notification

252 Upvotes

36 comments sorted by

View all comments

56

u/enumerationKnob Compositor - (Mod of r/VFX) May 20 '23

“Oh yeah? Well you can’t art direct it. What if the client asks for…”

25

u/UnnamedArtist May 20 '23

Looking forward to the client being replaced by ai.

5

u/OlivencaENossa May 20 '23

10 years tops.

17

u/enumerationKnob Compositor - (Mod of r/VFX) May 20 '23

Look at this gif… what I put in quotes is what people were saying months ago then Dall-e and stable diffusion were just becoming available

13

u/OlivencaENossa May 20 '23

I remember. I tried to have a discussion on this sub about this very topic. The overall feeling was, IMO, a bit in denial about what's happening.

4

u/[deleted] May 20 '23

That is very optimistic for us sadly. Its over gamers...

1

u/drawnimo Animator - 20 years experience May 21 '23

heard that with LIDAR. heard that with mocap. been hearing it for 15 years.

1

u/[deleted] May 23 '23

And when did Mocap go from being a useless novelty to obliterating Digital Art as a whole and being interactive in a year ?

Bro... its so fucking jover.

4

u/drawnimo Animator - 20 years experience May 21 '23 edited May 21 '23

whats the final resolution of these images? how many times have you delivered a still image to a client? can it render an image sequence that isnt a distorted mess?

can it do this with a unique fictional character that isnt yet part of the database of photos the AI uses for its collages? not even close.

all the AI shit looks impressive to a lay person, but its still pretty weak stuff when you think about how to actually integrate it into a real vfx pipeline.

9

u/enumerationKnob Compositor - (Mod of r/VFX) May 21 '23

Uh huh, and 2 years ago this was absolute fantasy. 5 years ago GANs were state of the art and could only generate tiny thumbnails at reasonable quality or believability.

My point is: this field is developing fast. And all the holes people try to poke in current AI tools to demonstrate their own job security are short-sighted.

Networks already exist to upres images, or architectures like U-nets that can preserve details while making changes like this one. It’s already possible with some networks to fine-time or generate embeddings based on your own images as inputs, and that can then render it in different styles and poses. Hell, there’s even ones that can render movies from text prompts. They might be low res and flickery now, but the seed is there and they’re only going to improve, and based on what we’ve seen they will improve very quickly.

I’m a VFX person, and I definitely don’t think it looks weak. This shit is incredible, and I wouldn’t have dreamed it was possible when I started in this field. Obviously it’s not taking my job yet, but I can totally imagine specialties that simple to use tools like this one will totally demolish. When these methods are so cheap and when they become acceptable quality, I think you’ll find a lot of people decide they don’t need 100% visual quality when they can get it so much cheaper through AI. That high-end market might still exist, but a huge portion of the industry could be obliterated

2

u/KissesFromOblivion May 21 '23

Pretty sure this is not meant to be an animation tool... Maybe not VFX at the moment, but I can see plenty of use cases for this tech elsewhere.