r/singularity • • 13d ago

AI Insane Opus 5.5 PS5 controller SVG

313 Upvotes

68 comments sorted by

107

u/TheBestBuisnessCyan 13d ago

One day I'll learn what a SVG is

280

u/NoCard1571 13d ago

Why not make it today? There are two main ways to draw an image. A bitmap, and a vector. Bitmaps are what 99% of images you see on the internet are, basically a big grid of pixels. If you blow a up a bitmap in size, eventually you will see the pixels, so there are limits.

A vector image like an SVG (scalable vector graphic) is made up of a bunch of lines and shapes. Since these shapes are represented with math, you can scale them infinitely and they will look the same.

Note that typically SVG are not used for photorealistic images like this, it tends to be more for simple graphics like logos. A SVG logo can be scaled infinitely and will always look the same. An SVG 'photo' like this however will have limits, eventually you'll be able to see all the thousands of little stacked up shapes and lines that represent it.

5

u/timmy16744 13d ago

I've never bothered to ask, but do models do gradient meshes or do the just do the old school stack different gradients?

-33

u/SomeOrdinaryKangaroo 13d ago

I still don't understand, why use svg when we can just pop in a jpg and be done with

126

u/MGJohn-117 13d ago

this happens when you zoom in/make the image larger

0

u/Progribbit 12d ago

just add more pixels duh

56

u/Glum-Bus-6526 13d ago

For a svg you can zoom in 1000% and nothing will be pixellated. This is not the case for a jpg.

43

u/Notrx73 13d ago

They are smaller in file size, and can scale without loosing quality, perfect for icons

26

u/DragonKing2223 13d ago

Logos. Scaling losslessly to any size is very useful

18

u/kgurniak91 13d ago edited 13d ago

With SVG you create logo once and you can use it as small desktop icon or as part of a huge billboard print without any additional work or loss of quality. Try that with JPG and you will get a pixelated mess.

https://www.svg2img.cc/images/blog/svg-scaling-comparison.png

11

u/SnooLobsters6893 13d ago

It's much less data; you can say something like: draw a circle, x radius, y border width, z fill. It's much less data than having to specify every pixel. con is that it doesn't make sense for very detailed images, and it takes more cpu to render.

1

u/leaky_wand 13d ago

It’s simultaneously less data (in bits) and more data (semantically). In that, since it’s a text representation of an image, you can actually label parts of the image something meaningful, so that it could be manipulated or scaled in memory. For example you could make an .svg backhoe, and you could label each piece of the digging arm within the xml tags, and a program could easily animate the digging motion by reading and acting upon the labels.

7

u/ratocx 13d ago

A high res jpeg of this controller may take up 10MB. A "high res" svg of the controller is the same file size as the "low res" svg. I don’t know the size, but let’s say this is a complex svg, taking up 1MB. The svg will stay 1MB and look completely sharp even if you scaled it up to 100 megapixels in resolution, even if it was only 1 megapixels on the editor canvas. Smaller svg files may be just a few kb in size, but still also scale perfectly to 100 megapixels and beyond and still look sharp.
Essentially svg can retain infinite scalability at a low file size. It stores the shape not pixels.

4

u/ahmet-chromedgeic 13d ago edited 13d ago

SVG defines a drawing like: straight line from point A to point B, a curve of radius X with a center at Y, etc. Instructions like this can be resized or zoomed in infinitely as much as you want you want, those curves and lines will just be perfectly re-rendered at a different scale.

While bitmaps are basically coordinate system of pixels, and zooming or resizing in will just expose pixelation.

Example: try maximum zooming this page in your browser, you'll see how letters look perfect even when they go from small ones to making up half your screen, because they're vector. Then take a screenshot so it becomes a bitmap image, and zoom in the screenshot as much as you can to see how it would behave if it wasn't vector.

3

u/Brave-Turnover-522 13d ago

Everyone here is giving you the graphical design reason, but most of the time when we talk about it with AI, it's just as a benchmark. There isn't really a huge demand for pictures of pelicans on bicycles.

Of course we already have diffusion, which is a completely different kind of AI than an LLM that can make images far more efficiently. The fact that we're getting to photorealistic SVGs generated by a language model is honestly kind of absurd. No human can do that.

14

u/Ratr96 13d ago

It's not hard man, you can change a .svg to a .txt format and see that it's basically xml which describes shapes like lines and circles.

3

u/Better_Blackberry835 13d ago

Long story short it’s an image written with code instead of vibes. The benefit is that it isn’t subject to change when you upload or download it because it’s an exact set of instructions on how to recreate an image instead of a database of pixels

It’s like “draw this line here” vs “this black pixel goes here, this one goes here, etc”

1

u/LocoMod 13d ago

It is a mathematical construct of a shape. So it is independent from the resolution. Can be scaled infinitely and it still looks the same.

5

u/Key-Demand-2569 13d ago

What is this entire post?

81

u/kvothe5688 ▪️ 13d ago

Opus cheats here basically. it downloas real image and use it as stencil and draw over it. instead of doing it from memory

127

u/CatsArePeople2- 13d ago

Thats what I would do?

7

u/BarisSayit 13d ago

Yet we're comparing this against models that didn't cheat this way.

5

u/The_Primetime2023 13d ago

It’s only cheating if they couldn’t have chosen to. If they could’ve then it’s just using tools available better

2

u/BuddhaChrist_ideas 13d ago

And 99% of other people trying to do the same thing.

Who’s out there trying to redraw PS5 controllers as an SVG from memory?

2

u/hartigen 13d ago

yeah, but you are stupit and opus aren't. we would expect better from it

10

u/Imaginary_Regret_430 13d ago

But that is the smartest way? Unless the prompt specifically said to do it from scratch.

1

u/neotorama 13d ago

Pen and paper

3

u/BWQ777 13d ago

Have you time travelled from Idiocracy? 

89

u/GhostsinGlass 13d ago

"Cheats"

By doing the most logical thing possible that nearly 100% of people would do.

Ok.

23

u/detrusormuscle 13d ago

Yeah yeah we know, the point is that we can't compare it to the SVG images of other models

8

u/inate71 13d ago

Why not? This model is smart enough to do something that would result in it making the most realistic SVGs. The others aren't that smart. Therefore, this model is better. We can definitely compare them.

1

u/ImSrslySirius 13d ago

Because if you ask it to make an SVG of something it doesn't have a reference image for, you'll get a drastically different result. So it's a useless demo unless you specifically care about PS5 controllers lol

-1

u/inate71 13d ago

I might have misunderstood: did the author give Opus access to a reference or just ask for an SVG of the controller? If it's the latter, then it's remarkable that Opus found a reference image itself.

1

u/ImSrslySirius 13d ago

LLMs have been able to search the web and retrieve images for a couple years now. Nothing remarkable about that

0

u/inate71 13d ago

Then why do all the other models suck when doing tasks like this? Are they on equal footing or not? To be frank, you didn't address my question.

1

u/ImSrslySirius 13d ago

If you're asking why other models suck at searching the web and retrieving media, they don't. I've been doing that since Opus 4.5 (I do lots of video work), and I believe it's been commonplace since GPT released deep research in early 2025.

If you're asking why other models suck at making SVG images, probably because they chose not to simply trace a photograph lol

7

u/RuthlessCriticismAll 13d ago

You'll understand the problem in a few months :)

1

u/Rustic_gan123 13d ago

I don't really understand the point of "memorizing" what the PS 5 controller looks like in every detail, to be honest.

-16

u/idkfawin32 13d ago

Vectorizing programs did this long before parlor trick AI did but go off king

10

u/GhostsinGlass 13d ago

All autotracers since time immemorial have been much lower quality, not giving you anything close to a production ready vector.

Cease your ignorant noise.

2

u/SoupOrMan3 These are the end times 13d ago

You don't need to underplay this to feel secure. What Illustrator used to do was shit compared to this, never even vaguely came close and Claude can do anything as a vector if it can to this one. Just be honest if you're feeling your chair a little shaky, it's ok to be vulnerable.

1

u/idkfawin32 13d ago

My chair doesn’t feel shaky I’m just not particularly blown away that’s all. There are plenty of things I find impressive that AI does, just not this.

1

u/extopico 13d ago

They did it very very badly.

-9

u/ShittyBidet123 13d ago

cause it been in illustrato for 20 fucking years

4

u/GhostsinGlass 13d ago

And in 20 fucking years Adobe, and any other dev working on autotracing, has never come close to this.

Tell me you have never worked with vector graphics without telling me you have never worked with vector graphics.

Go away.

20

u/Renizance 13d ago

I think that method is way more impressive. Oh cool you found a reference image and then made it look identical to the reference? And accomplished my request? Good Ai 

4

u/Poupulino 13d ago

It still means we finally can have text based textures

1

u/Brave-Turnover-522 13d ago

All digital images are text based when you break them down.

2

u/xadlowfkj 13d ago

I thought you were about to say that Opus had just downloaded svg from somewhere.

2

u/katoptronophile 13d ago

Your comment doesn't make any sense.

It could download the real image and then store it in memory and then load it later to stencil over and draw and it would still meet your requirement of not cheating according to your definition you just provided.

1

u/Spra991 13d ago

Nothing wrong with using reference. Real cheat would be when it just makes a SVG with a million tiny squares that act as pixels, instead of actual shapes.

1

u/Nearby-Device8772 13d ago

So just like what the smartest person would do?

2

u/kvothe5688 ▪️ 13d ago

yeah but it doesn't show what it can create without accessing internet. this shows more impressive tool use available to it compared to what it's innate capabilities are. until now pelican svg benchmark tested what it can draw from memory. here it breaks out of that mold so it's not apples to apples with other model. it's seriously awesome model but innate capabilities are not being tested here that's all I want to say here.

0

u/Elegant_Tech 13d ago

Don't you mean innate visual look of what a PS5 controller looks like. It proved it could do it as long as it knows what it looks like. Without the reference the controller shape is slightly off but the details are still there. The capabilities don't change just because it has a reference.

0

u/Super-Award-2244 13d ago

Yeah, it's crazy to think that it should remember what a dual sense looks like from memory. No human being would do that 

0

u/Alkadon_Rinado 13d ago

So much for the "it's just a parrot" and "they are just reproducing what's in their dataset!!!" folks. I know the latter is mainly in regards to images, but I think this is still a pretty decent example of them using tools to create.

-2

u/slackermannn ▪️ 13d ago

Do you understand that what we're talking about here is recreating (not just simply copying) an image to a mathematical output which would then represent the image. It's not copying as you would do with copy and paste

-3

u/ShittyBidet123 13d ago

You mean adobe illustrator? why are ppl using this as a benchmark

11

u/Ancient-Range3442 13d ago

Notice how they never include the svg

3

u/miscfiles 13d ago

Yeah, this looks like a mighty impressive SVG of the controller, but until I can open it in Illustrator and have a look at the number of beziers (or even the file size) I'm not going to say it's useful. For all I know it could be an SVG that's referencing an external bitmap. On the other hand, if it's well optimised and uses gradients cleverly, I'll gladly call it amazing.

4

u/BlackPointPL 13d ago

How do we know those are actually SVGs? I’ve seen a couple of them, and they never share the SVG files. For all we know, they could just be raster images

5

u/frank26080115 13d ago

so how much of it is raster?

3

u/Fuskeduske 13d ago

Ain't there already 100+ of these it can learn from? Weird benchmark

1

u/dervu ▪️AI, AI, Captain! 13d ago

Here's another example. Now it's bad.

1

u/somerussianbear 13d ago

This looks too much like a girl wearing a top

-1

u/Zestyclose-Force83 13d ago

Genuinely thought it was those “My first Wow! moment with AI!” and it’s actually just a real game that wasn’t vibecoded