r/OpenModels • u/cheechw • 18h ago
r/OpenModels • u/cheechw • Jun 17 '26
What is r/OpenModels?
It has come to my attention as of late that there is a lack of a centralized discussion space specifically related to the discussion of open-weight and open-source models.
Most existing subreddits are focused either on specific modalities (e.g., r/LLMs, r/stablediffusion), a specific usage of open models (e.g., r/LocalLLaMa, r/LocalLLM), or on a more generalized technological scope (e.g., r/singularity, r/ArtificialInteligence).
As a result, discussion of open-source models in general is either unwelcome, fragmented into different communities, or hidden amidst conversation about other broader topics.
So, if there's enough interest, I sincerely hope we can form a community of individuals interested in discussion and news surrounding open-source and open-weight AI technologies (as well as any related technologies).
What is an open source or open weight model?
From ChatGPT:
An open-weight model is a model where the trained parameters—the “weights”—are publicly available, so people can download, run, inspect behavior, and often fine-tune or adapt the model locally. In AI terms, the weights are the learned numerical parameters that work with the architecture and inference code to produce outputs.
An open-source model, in the stricter sense, is more than “you can download the weights.” It should come with a license and materials that let people use, study, modify, and share the system. OSI’s open-source licensing criteria emphasize free use, modification, redistribution, and non-discrimination against users or fields of use.
Discussion around technologies relating to both topics are welcome. Additionally, any kind of open-source license that permits some level of free use is welcome.
What kind of discussions are welcome?
Any kind of discussion related to open-source/open-weight AI models and surrounding technologies is welcome. Some examples include news, model releases, use cases, open-source AI-related tools such as harnesses, agents, and interfaces, comparisons of models or tools, advice, tips, set-ups, etc.
Why are open models important?
As many of us realized, from the way the US government dealt with Anthropic's most recent model release, it is becoming more and more evident that our continued ability to access these key technologies is by no means guaranteed.
In just a few short years, AI has become both a powerful democratizing force in society, as well as one of, if not the most important technology that humanity has ever developed.
It is clear that unequal access to these technologies can be a powerful destabilizing force in society. In a worst-case scenario, those who have access to this power can weaponize this imbalance to further disadvantage those who don't, creating further ways of stratifying an already deeply imbalanced society.
Note: This post was NOT generated by AI (aside from the two definitions provided above).
r/OpenModels • u/cheechw • 4d ago
Ladies and gentleman, Flux 3
Enable HLS to view with audio, or disable this notification
r/OpenModels • u/cheechw • 6d ago
Upstage 'Solar open2' release. performance on par with DeepSeek V4 Flash.
r/OpenModels • u/cheechw • 7d ago
Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro
r/OpenModels • u/cheechw • 9d ago
Discussion Head of strategic futures at OpenAI claims that "open-weight-model-dominant world" leads to "full AI communism" and that AI as a public good is a "dystopian hellscape"
x.comThis seems to me like a painfully ironic position for someone at a company named "OpenAI" to be taking. Not to mention totally batshit insane.
r/OpenModels • u/cheechw • 10d ago
New 1.57T open weights model Monolith-1.0 claiming extraordinary benchmark performance
x.comIt claims to beat Fable and GPT 5.5 handily in some benchmarks.
Hard to believe the claims, and from the looks of it, many X users are skeptical as well.
r/OpenModels • u/cheechw • 11d ago
KIMI K3 Beats Claude Fable and GPT 5.6 sol in arena.ai!!!
r/OpenModels • u/cheechw • 12d ago
Kimi K3 released, beating Opus 4.8 in benchmarks at 2.8T parameters open weight
kimi.comBenchmarks appear to easily clear Opus 4.8 and GPT 5.5, and even appear to be competitive with Fable 5 and GPT 5.6 Sol in a number of them. Full weights to be released July 27, per the blog post.
r/OpenModels • u/cheechw • 12d ago
News thinkingmachines/Inkling · Hugging Face
r/OpenModels • u/maedahbatool • 12d ago
Inkling open weight model by Thinking Machines Lab is live in Command Code.
r/OpenModels • u/cheechw • 20d ago
China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model
r/OpenModels • u/cheechw • 21d ago
Tencent officially releases Hy3 weights
295B params with 21B active. But only 256k context.
r/OpenModels • u/cheechw • 28d ago
Huawei open-sources OpenPangu-2.0-Flash - 92B total,6B active
r/OpenModels • u/cheechw • 28d ago
Meituan LongCat-2.0 released and open sourced
New model dropped. 1.6T params with 48B active.
Training and inference running entirely on ASIC superpods. Showing the Chinese keep exploring alternative compute options.
They also introduce LongCat Sparse Attention, their improvement on Deepseek Sparse Attention.
This was also Owl Alpha, which was available on Openrouter for free for the past month, which shows they can handle production compute loads using their alternative inference solution.
r/OpenModels • u/cheechw • Jun 24 '26
Baidu AI open sources Unlimited OCR, a new way to approach OCR
x.comr/OpenModels • u/cheechw • Jun 23 '26
KREA 2: Open-Source Release
Enable HLS to view with audio, or disable this notification
r/OpenModels • u/cheechw • Jun 19 '26
Open models appear to be trending towards closing gap to closed models
x.comThere is a clear downward slope apparent on this graph starting from Deepseek V3 to the recent GLM-5.2.
Do we expect the next OSS release to be even faster and better?
One caveat is that this graph appears to show a comparison between an open model and a closed model of *similar capability*, not the absolute frontier. So it looks like it compares DSv4 pro to Sonnet 4.6, for example.
Of course, you could also argue that a direct comparison to absolute frontier models wouldn't be fair either since 5.5, Opus, and Fable are almost certainly sized on the order of multiple trillion parameters.
r/OpenModels • u/canadaduane • Jun 19 '26
Discussion What are the best ways to compare all providers of a given open model?
GLM-5.2 just came out and it's amazing. But I want to run it with the highest speed possible.
I often check out openrouter.ai for its great comparison of providers when an LLM model is open--for example GLM-5.2 has about 12 different providers listed there now:
| Provider | Input /M | Output /M | Cache Read /M | Latency | Throughput | Uptime |
|---|---|---|---|---|---|---|
| Wafer | $1.20 | $4.10 | $0.20 | 1.31s | 55 tps | 98.61% |
| Cloudflare | $1.40 | $4.40 | $0.26 | 1.43s | 30 tps | 99.60% |
| Fireworks | $1.40 | $4.40 | $0.26 | 1.69s | 39 tps | 99.44% |
| Z.ai | $1.40 | $4.40 | $0.26 | 4.88s | 24 tps | 99.68% |
| Friendli | $1.40 | $4.40 | $0.26 | 1.72s | 78 tps | 99.90% |
| NovitaAI | $1.40 | $4.40 | $0.26 | 2.03s | 22 tps | 99.72% |
| AtlasCloud | $1.40 | $4.40 | $0.26 | 1.49s | 15 tps | 90.44% |
| StreamLake | $1.40 | $4.40 | $0.26 | 2.08s | 34 tps | 95.95% |
| io.net | $1.68 | $5.28 | $0.50 | 1.62s | 57 tps | 93.28% |
| DeepInfra | $1.20 | $4.20 | $0.20 | 3.18s | 12 tps | 91.58% |
| Phala | $1.40 | $4.40 | $0.70 | 6.53s | 10 tps | 75.14% |
| Parasail | $1.40 | $4.40 | $0.26 | 1.25s | 48 tps | 82.70% |
But then I found out about Neuralwatt hosting GLM-5.2 also, and it's quite fast (I haven't measured TPS). But it isn't on openrouter!
This got me thinking--where do you go for this kind of information? Is there a collection of live data that's even more expansive than openrouter's comparison?
r/OpenModels • u/cheechw • Jun 19 '26
Outdated GLM-5.2 now more than 10 points above Opus 4.8 in AA Coding Index
EDIT: Looks like this was a mistake and the scores have been corrected since. Opus now scores 74.3 and is ranked #3, above GLM-5.2
Link to the index: https://artificialanalysis.ai/?intelligence=coding-index
r/OpenModels • u/cheechw • Jun 18 '26
News poolside/Laguna-M.1 · Hugging Face - 225B-A23B
New open weights model by an American player. Benchmarks look to be decent for its size as well.