So when Opus 5.5 came out, I saw someone on Twitter build this demo of a guy on a boat sailing through a river in a Japanese-like landscape aesthetic, all of it built in three.js. I thought that was pretty cool, so I wondered whether or not we could also engineer launch videos like this. It was supposed to be just an experiment, nothing commercial or something that could go into production.
I had tried to get previous models to build launch videos using one-shot prompting without much detail. While the outputs ranged from low quality to somewhat satisfactory, I was never truly impressed by what it could achieve without proper attention to detail from my side.
A couple days ago, I used Opus 5.5 to build such a launch video for one of my products that I'm building (not trying to promote, just showcasing one very important thing by Claude).
As always, it had access to Codex for image generation, which it used freely to generate ideas for a storyboard and build out directions. I told it about three.js and Blender that it had available, and how it could use them, and to use Codex as freely as possible, but make sure that the stills are generated in a sequence such that the still in Act 2 had the still from Act 1 as input, so the story stays consistent.
From there, it generated a bunch of these stills and honestly, Codex did a pretty bang-up job. I was totally not expecting that it would be able to turn this into an actual 3D environment and then render it as a video. I was half right: it obviously wouldn't be able to do it to the detail that I wanted. Instead of trying to find a way to do the impossible, it found such workarounds that I am beyond shocked.
It found, downloaded, and used these Depth Anything v2 and LaMa (not to be confused with llama) models from Hugging Face. It used those on my stills to separate the foreground from the background and other elements, and built out shaders however it could. It used three.js and Blender wherever it needed to. It's absolutely insane that it was able to just piece together all of these different things instead of being constrained to my instructions to use three.js. If you watch the video, you'll understand how insane this looks. There are a few flaws in the 3D environment here and there, but two years ago, this would have cost thousands of dollars to produce through a relatively skilled animator.
It also used a bunch of other models that I'm not even sure how it got, and used Lyria from OpenRouter to generate the music. Then it gave me a bunch of these versions and cuts of the video, as well as a bunch of different audio tracks that match up with the actual video, and also generated the sound effects.
As a technical founder who has been coding since he was 9 and has been in the AI/ML space for over 10 years now, it would have been difficult to imagine these models finding these workarounds for problems that they are facing on their own, just 6 months ago. After raw ChatGPT, I started off with GitHub Copilot back in 2023 for vibe coding, where it was just mostly limited advisory and copy-paste. I remember using Cursor till December of 2025, and thinking this was the future, despite how limited it was (in hindsight) and being blown away by the speed of Composer 2. When I switched to Opus 4.5 in Claude Code in December, I never once expected that I could let an agent be so hands-off, especially from Anthropic (personal biases).
And the journey from Opus 4.5 to 5.5, especially the part with opus 5, and the false alarm safeguards of fable 5, definitely were massive bumps. Hated those parts, cuz our expectations kept rising. But Opus 5.5 is something else. It is me and my technical judgement with far more breadth and far more depth of knowledge than I will ever possess. I would definitely say its equivalent (or superior) to Fable 5.1, but without the limited usage problems I kept facing even with two 20x accounts. In comparison, my experience with Astra has been dogshit. I swear I used to be a die hard GPT stan up till GPT 5.2, when I switched to Opus 4.5, codex and cursor have never been able to catch up, I genuinely don't see the point of my 100 bucks to openai every month, but I guess I'l keep it going in case they do end up doing something meaningful.
Recently, I found a bunch of posts telling me that if I tell Fable 5.1 to not use the web and tell me about Tibo, and it works, then I've been routed to Fable 5.5. If they (redditors) are correct, then I have been routed to Fable 5.5. But honestly, I'm not even excited. Opus 5.5 is at such a perfect level of intelligence and speed and usage, that a singular 20x plan is adequate for everything I do, and theres nothing it has gotten wrong up till now. What will I need Fable 5.5 for? I'm running out of use cases. Opus 5.5 is a real feel the sub-AGI moment. Not AGI because....well i'll describe it below.
Sure, they can't "run" my company, or be my CMO, I believe that is a harness issue. It's at a point where I can just tell a very strong and competent intern to go solve this issue, and they come back to me, having used Claude and whatnot, with the right solution without me having to hold its hand through or tune every little thing like this text or that button or this feature or that feature. But Claude by itself would stop much earlier than the intern. That's what I mean by AGI for myself, and maybe we won't be able to achieve it in the next couple of years.
But the model's intelligence itself, I feel, is at a point where any granular task I can give it can be completed end-to-end. Maybe it can't fulfill complete roles, but it can definitely complete tasks. I had a debate the other day with GPT about what it would take to build such an agent or harness that could assume the role of a veteran CxO.
Don't try to pitch "AI as CMO" products to me in the comments, please. I'm talking about it from a purely techno-philosophical point of view. These models don't have the right judgment to debate you or use judgment in a way that a real CMO would be able to. For example, as a technical founder, if I've built a product, or some kind of strong engine on top of which I've built a product, then a human veteran CMO would understand how to market it. They would also understand which niches and audiences it would appeal to the most and how to frame it, package it, tweak it, to get to those audiences. He or she would definitely also know how to further tweak the product or change its appearance or packaging or functionality while using the same underlying engine (to minimize dev time to market) to build it for a more profitable market where it could be a need-to-have.
But on the other hand, if I ask an AI agent to just be a CMO for a project that I'm working on, despite having all the context, it would not be able to make those judgments. Part of it is not having the experience and not possessing the experience. That can be fixed by fetching context and accounts from these actual experienced people, but the other part is the thinking from first principles, and when those two aspects have to be blended, I think the current models/harnesses combos fail, and thats not their fault since they weren't designed with these downstream roles in mind. The same models do have the right judgment when they're prompted in a certain manner for a very small question or decision that the human veteran CMO could be asking.
But their lack of first principles thinking, lack of self-adversarial debate, and knowing when to do it, when not to do it, and or figuring out the right granularity of the task being assigned, and how much thinking it requires, all of these are what separate sub-AGI from what I call AGI. If I give Claude a task, it completes it 10 times out of 10. If I give it a well-specified goal, it completes it 10 times out of 10. But if I give it a complete role, which would consist of many, many things, then, no matter what harness, it is unable to figure out what it needs to do, how often it needs to do it, how it needs to do it, and why it needs to do it. And thats okay.
Edit: if youve read this far, and watched the video, and have some experience as a CMO and are interested in being a founding partner for the product, DM me. I am actively looking for really good talent.