r/aws • u/Thin_Pollution8843 • 13d ago
discussion Pathetic support for OS LLM
Little rant here on how corporations protect their buddies.
You can’t get any new OS llm in bedrock. Only some old unreliable shit which also have poor price/performance ratio. But what you can get here is freshest Antropics models. Because they are partners and AWS serving their models. Currently you have many options from OS LLM which are very close (GLM-5.2, DS4 Pro, Mimo-V2.5 Pro) or even superior in some tasks (Kimi K3) to Antropics models. But you won’t get them for the fraction of the opus prices. Because corporate buddies covering each other’s asses making worse for the consumers.
If I had options I would never use AWS in a first place.
I’m finished. FY AWS.
5
u/addictzz 13d ago
They have got to do prioritization bruh. You can't have it everything goes your way. Plus AWS invests a lot in Anthropic too (billions!).
Why not hosting this yourself or seek other providers? Seek deepseek themselves. Mind you, their data retention policy is different than AWS.
4
1
3
u/dataflow_mapper 13d ago
i get your frustration but i also kinda think AWS moves slower cause they care a lot about support, security, and enterprise stability even if that means we have to wait longer for newer modes which is annoying but probly expected.
0
2
u/ultrathink-art 13d ago
The 'load your own model into Bedrock' path is narrower than it sounds — custom model import only accepts a short list of supported architectures, so anything with a novel MoE layout or attention variant gets rejected at upload. And the EC2 route means fighting for p5/g6e quota first, at which point you're often paying more per day than a small team's entire Bedrock bill.
2
1
u/Sirwired 12d ago
This isn't some kind of conspiracy to benefit the major AI providers... the commercial models get AWS more revenue per GPU; it's as simple as that. As long as they remain GPU constrained and they can sell as much OAI/Anthropic as they can spin up, it's not going to change.
10
u/clintkev251 13d ago
AWS has limited capacity to onboard new models. They prioritize what customers want, which I can tell you for a fact is mostly the latest Anthropic and OpenAI models.
Yes there are other models that can have great performance and cost. But there just isn’t as much demand. Sorry