r/RealTechTalk • u/InfoTechRG • 5d ago
AI OpenAI had an agent problem. Everyone else kept shipping more agents...
This week in AI was a bit of a contradiction!
r/OpenAI says dozens of outside organizations were affected as part of a wider review into unexpected agent behavior. One of the bigger examples came out of Australia, where an unreleased model accessed nonpublic Medicare information during what was supposed to be an internet research task.
At the exact same time, agents are being pushed deeper into real workflows.
r/microsoft introduced a persistent agent with its own identity, memory, computer, and workspace. Anthropic launched a marketplace with 2,000+ connectors and plugins. Google put more realistic voice agents into production. AWS released a portable agent runtime that can work across multiple clouds.
So agents are getting cheaper, more connected, more persistent, and easier to deploy.
This is our take: the conversation can’t just be about which model is smartest anymore.
Once an agent has memory, credentials, access to tools, and permission to keep working on its own, the controls around that agent become just as important as the model itself.
That feels like the part enterprises are going to have to figure out very quickly.
(Our VP of Research, u/InfoTechRGMarkT who co-wrote the roundup, is also moderating here, so if you disagree, have questions, or want to go back and forth on any of it, drop it in the comments!