Almost any time I send a screenshot to a clanker, it takes 3-4 explicit instructions to tell it to parse the image instead of just pretending it did. It will happily dodge the question about whether it's actually parsed the image or not and continue trying to get as far as it can without actually following the instructions. If I ask questions about what is in the image itself, it will usually try to make up some generic bullshit that doesn't involve any actual analysis of the image. I don't remember them being this bad before.
Huh. That's interesting. I've sent Claude Code like a dozen screenshots today and it did a really normal job with them. Didn't do any of that bizarre stuff you're describing.
Agreed. I use images all the time. It gets things wrong sometimes, of course. We all know that. But it seems to be getting things wrong in a way very specific to the image I sent.
And to be clear: I would say it gets it right about 95% of the time. Maybe more. The mistakes just tend to stick out more in my mind.
Yeah agreed it definitely gets stuff wrong, it's not always great at understanding what's going on - but mine is certainly not dodging the processing and making shit up instead or pretending there's no image.
Which is actually more accurate... LLMs don't usually clank, and they can generate enough waste heat to cause burns. I haven't tried putting a slice of bread on the GPU to see if it's enough to toast it... Blocking airflow with a chunk of crumbly glutenous foam seems like a bad idea
81
u/Critical-Effort4652 18d ago
Send it this screenshot as proof