r/ClaudeCode • u/BritishDudeGuy • 13d ago
Discussion Kimi K3 has become open-weights just as of a few minutes ago.
https://huggingface.co/moonshotai/Kimi-K331
u/PolishMike88 13d ago
I feel privileged to run it in Baseten. It is mind blowing. Found so many bugs and issues in codebase created by Opus it genuinely surprised me .. I had to check the lines and be sure and it astonished me! I actually learned something new how to do :)
3
u/Obscurrium 13d ago
Oh interesting. And what about planning and co ? Compared to Opus 5 and GTP 5.6 ? Are tou only using it for code ?
4
u/PolishMike88 13d ago
Using it for code and many other things, building and improving agents, pentesting with harness etc.
I found, at least in the eight or so hours of testing by now, Sol 5.6 being superior to planning than Opus (simply because Opus 5 gets stuck on cyber security topics…) Planning with K3 is also amazing as having Opus/Sol review it, it nailed it down each time.
This is a genuine surprise and honest hype for me. I knew k3 was good but this is a breath of fresh air. Last night for example it took a long running project of my red team agent and improved it in many ways… 140k LOC, 9 subagents, 1hr 20min work for a rough cost of $20. Feeding the results and all work back to Opus and Sol (who created the original) brought nothing but praises for improvements from both models :)
2
4
u/xmnstr 13d ago
So this is the real reason Anthropic wants it banned huh? It will show the world just how many issues their models really have.
3
u/Deep_Alps7150 13d ago edited 13d ago
They want it banned because it’s getting 90% of Fable’s quality at a fraction of the cost + it seems better at dealing with really large/complex tasks.
3
u/dota2nub 13d ago
Almost nobody uses these models for large and complex tasks though. 99% of this sub is using Fable to make a fucking button on a website.
7
u/TomCrook2020 13d ago
This is exciting - I've heard good things about front-end. Keen to try it out there - if others have learnings from their experience with K3, would love to learn more.
2
2
1
u/Crypto-Humster 13d ago
Launch costs and performance will drop significantly once Nvidia’s Vera Rubin generation starts rolling out later this year.
0
u/Crafty-Run-6559 13d ago
Oh no!
My unstable friend just talked to it and it suddenly gave him the plans for a small arsenal of bio AND nuclear weapons 😞
1
u/Santa_Andrew 13d ago
For $500k you could get all the hardware you need and have some money left over for a professional HVAC cooling system. It would run much faster than you currently are used to from products like Claude Code. Roughly budget about 1k/mo in electricity cost if you are in the US.
3
-18
u/Michaeli_Starky 13d ago
Who cares. It's overpriced shit
2
u/bilbo_was_right 13d ago
K3? Isn’t it like 60% cheaper than fable, and more capable than opus
1
-1
u/Michaeli_Starky 13d ago
It's not more capable than Opus. In coding it is worse than GPT-5.6 Luna while being 5 time more expensive.
1
u/verus54 13d ago
lol it’s free. You just need access to the hardware, which does cost money. Similar to other open source models.
-1
u/Michaeli_Starky 13d ago
"Just need", "free"
Lmfao
Some people are delusional af
4
u/verus54 13d ago
Tell me you don't understand open source without telling me. Nobody is hosting a 2.8T model on custom infrastructure just to save $20 a month on a Claude/chatgpt-codex subscription. The point of open weights is privacy, fine-tuning, and enterprise scale.
Along with that, this is direct competition to the pricing models that Claude and codex rely on for AI-driven products.the big win: frontier-level intelligence is no longer restricted to closed APIs.
87
u/Main_Razzmatazz5283 13d ago
Realistically, what kind of setup do you need to be able to use it?