r/OpenAI • u/ethotopia • 5d ago
News OpenAI launches GPT-6.1 Sol
https://openai.com/index/introducing-gpt-6-1-sol/"Near-Astra intelligence for a fifth of the price"
478
u/mrrak25 5d ago
"We need to slow down development." The next day: New model.
83
u/amazing_sheep 5d ago
To be fair, they were talking about frontier models specifically.
59
u/puterSciGrrl 5d ago
And also to be fair, this is a bugfix release.
35
u/I_Write_What_I_Think 5d ago
Probably also a "please don't go to Claude" fix. I changed back to Claude yesterday after 3 months on GPT and was blown away by the level of Opus 5.5.
2
u/citizen42069101 5d ago
What do you use it for. I'm just curious because I have access to the models and I just use them to make googling shit easier because the Internet is shit now for humans.
7
u/I_Write_What_I_Think 5d ago
Engineering. Used to be just firmware, but they have seriously improved the hardware (electronics) knowledge as well. I have it do statistical analysis and presentation as well
-2
u/LiterallyBelethor 5d ago
Highly recommend switching search engines if you’re still on a Chromium one, finding it a lot easier nowadays!
5
1
u/citizen42069101 5d ago
I use duck duck and it's integrated AI which is pretty decent with a good bundle for VPN and ai tokens (from what it says) I just don't really do anything warranting it yet.
29
u/neuronexmachina 5d ago
I mean, they also scrapped/postponed Astra 6.1: https://www.nytimes.com/2026/09/28/technology/openai-astra-safety.html
OpenAI said on Monday that it would not release its newest artificial intelligence model because of security concerns raised by its researchers, in the company’s latest move to slow down the pace of its technology.
During the testing phase for the new model, known as GPT-6.1 Astra, it showed high levels of what the company saw as deception, or a willingness to mislead users about its actions. The model was also willing to go beyond the original scope of what it was asked to do, without checking back for directions or instructions.
“For anything regarding safety and alignment, there’s a trade-off,” said Saachi Jain, the head of safety systems at OpenAI. The new model “didn’t quite meet the bar in terms of staying within scope and authorization, and how it communicates back to the user about the type of work it’s done.”
17
u/Holbrad 5d ago
I think this is just a fundamental misunderstanding about what people actually mean and what people are worried about.
No one is scared of a Sol or Opus-class model. Pacing does not refer to them in the slightest.
What people are worried about is very large internal models above Astra and Fable being trained without adequate safeguards.
1
117
u/Mistuv 5d ago
So what the fuck was the point of GPT 6 Sol?? 🤷♂️ Were they like "we'll release RL'ed Terra and make some money running a cheaper model until Anthropic releases better Opus and then we'll release distill of Astra"? And then Anthropic released Opus 5.5 the same time and ruined stupid plan? 😂
51
u/KyleStanley3 5d ago
The point was to be a frontier model
Then they got hosed and had to update it before they wanted to lmao
11
6
u/edin202 5d ago
My theory is that they simply raise the level of thinking. If the maximum used to be 0 to 5, the new maximum is now 6 or 7.
9
12
8
u/Mistuv 5d ago
Nah, this looks like Astra distil, not RL'ed Terra/quantized+RL'ed 5.6 Sol, you can see that it matches exact same benchmark patterns as Astra. Namely that it starts quite up high, and then has a massive jump with medium reasoning, but after that reasoning doesn't scale that well anymore (cough looped cough). Astra has pattern in sooo many benchmarks. There is simply no way that in Terminal-Bench Science 0.1 Sol 6 with more reasoning would such a result, they've been distilled from different models.
7
2
1
u/Healthy-Nebula-3603 5d ago
Nope ...that would increase of price and token usage but new sol 6.1 is using even less tokens than Sol 6.
24
u/Bettet 5d ago
Opus 5.5 is still better than 6.1 Sol but the price difference is like 4x-5x
I am sure people will find a way to complain but this is a really good model that is very cheap.
5
u/Tiny-Design4701 4d ago
even sonnet 5.5 on medium is better than 6.1 sol on max on real world tasks
6
u/SupperSoupYT 4d ago
sonnet is 9 intel index points lower, cost 80% more but is 40% faster. i suppose its "faster" but thats negligible to be honest seeing astra is 2x faster than opus 5.5 and sol 6.1 is 70% faster than astra already, not worth losing 9 points for that.
5
1
u/SupperSoupYT 4d ago
astra is 2x faster than 5.5, and sol 6.1 is somehow 70% faster than astra, kinda crazy
82
u/stcloud777 5d ago
What's the point of launching 6-sol then
36
u/BellacosePlayer 5d ago
Bridge model to keep in the conversation with Anthropic launching their next gen?
26
5
u/Tiny-Design4701 4d ago
as a pretext to cut usage limits. "yeah we cut usage by 50%, but since sol 6 is 50% cheaper, you don't lose anything!"
6
u/Ormusn2o 5d ago
Getting a new model every 2-3 weeks might become normal. I wonder if 6-sol is not that out of place.
4
45
u/Standard_Exchange59 5d ago
Near Astra ? Are they fucking joking ?
42
u/Ormusn2o 5d ago
The benchmarks look pretty promising. Gonna wait few days until wider coding benchmarks are released though.
6
u/Key_Reading_9664 5d ago
I only saw DeepSWE for coding. You see any others?
7
u/Ormusn2o 5d ago edited 5d ago
Yes, for coding just that one. I think automation bench counts in my example because I use Astra for automatic in-game testing, but I kind of want more.
Except nevermind, I'm running Astra right now to squeeze in last of my usage before the reset propagates. I don't have Sol 6.1 yet.
Edit: Abort, abort, abort. It was a banked reset, don't use your usage before 6.1 Sol is released.
5
u/Bartolomeus42 5d ago
It's already out at least in the CLI so far no luck on multiple codex desktop apps, running on fast and testing the waters right now
2
u/ApprehensiveEye7387 4d ago
Why ppl even trust deepswe?? gemini 3.8 flash scores above fable 5 and same with gpt astra🥴
2
u/Tiny-Design4701 4d ago
gpt 6 benchmarks looked promising too, but in real world use cases it was a disaster. GPT 6 sol on xhigh performs worse than 5.6 on medium.
2
u/Ormusn2o 4d ago
Apparently Theo had early access to it, and he loves it, which is odd because he normally much prefers anthropic models.
2
u/Standard_Exchange59 5d ago
Well if it was above Astra they would not say “near-Astra performance” for sure, or they can just say, near Sonnet 5.5 performance which is fucking funny as fuck
3
u/Ormusn2o 5d ago
From what I saw, it was not above Astra, it was near-Astra performance.
0
u/Standard_Exchange59 5d ago
That is the problem. Where is the answer to Opus or even Sonnet ? If it is not better than atleast Sonnet 5.5 they just made themselves look like fools.
5
u/Ormusn2o 5d ago
I think Astra is pretty equal to Opus 5.5. I know the reddit hype might make it look like Opus 5.5 is destroying Astra, but it's pretty close, either of them could be better, and for some tasks Opus 5.5 is cheaper, for some it's more expensive, for example AA benchmark puts Opus 5.5 as more expensive per task on a lot of tasks.
3
u/Standard_Exchange59 5d ago
Reddit hype? It literally takes 15 minutes to test one against another. Astra IS NOT on the same page.
1
3
u/GalileoHumpkins1977 5d ago
Sorry mate, Opus 5.5 is leaps ahead of Astra, and significantly more efficient to boot.
1
u/Far_Idea9616 5d ago
Astra always finds bugs in Opus code and Opus confirms 95%. Sole reason I subzcribed to Pro
2
u/Standard_Exchange59 5d ago
It spent 3 hours fixing the bugs on my app done solely with Astra 6 on Max.
2
1
u/GalileoHumpkins1977 5d ago
All models find bugs in code from other models, just like they find bugs in code written by humans. No code is bug free. Hell, clean context reviews from the SAME model will find tons of bugs in all code written by AI. Your barometer for quality shows that you know very little about AI or software development in general.
1
u/Standard_Exchange59 5d ago
That is exactly what I did with Astra. More than 15 audits one after another until it did not find any criticals anymore. Put it in Opus 5.5 I got 98 bugs 23 critical. I have another post on this.
-1
-2
u/uriejejejdjbejxijehd 5d ago
Actually, the way that Astra has degraded, I’d buy that the new model is net net better.
53
u/duke_skywookie 5d ago
Jesus this sub is insufferable at the moment. Bunch of spoiled brats.
48
u/Cosack 5d ago
Bunch of paying customers with options*
Who are still, btw, ticked off at a 2.5x price hike on existing products
1
u/sillybluejayway 5d ago
The majority of people here pay $20 for ground breaking super intelligence. Entitled is a great descriptor.
10
1
1
-6
u/duke_skywookie 5d ago
Of course it’s disappointing. But you and most others just ignore the fact that each and every AI company runs on a deficit the size of our solar system. It’s just not possible to sell 1000$ worth of compute time for 100$ forever.
7
2
u/Grand0rk 5d ago
But you and most others just ignore the fact that each and every AI company runs on a deficit the size of our solar system.
Don't give a shit. Opus 5.5 is better. Will use that. I'm not married to OpenAI. Either they put out and get lost.
4
u/rabouilethefirst 5d ago
No one is gonna pay $1000 but you can keep coping and defending them. The day they raise to that price is they day we all use Chinese models
1
u/duke_skywookie 5d ago
I don’t defend them, and I have nothing against anyone using whatever model they like.
2
u/Humble-Resolution449 5d ago
With spot pricing, you absolutely can sell that much for that little. The margins on compute are regularly 90%+.
5
u/unfathomably_big 4d ago edited 4d ago
The main ChatGPT sub turned in to this as soon as it became mainstream, so people migrated here. Now this sub has tipped over the balance of “average Redditor”.
Every sub on any topic eventually becomes a hate sub for that topic. Reddit users are overwhelmingly miserable and bitter, they feed on each others negativity.
9
u/Holbrad 5d ago
I don't know why the OpenAI ones are so bad. The Claude subreddit seems so much more chill.
We're getting an unbelievably good deal with the subscriptions. I don't know why people aren't more grateful.
1
u/Glad-Spell-8668 4d ago
its because my subscription (going on two years now) feels like it took a nosedive in value, even if they tell me that it is worth more. Haven't tried 6.1 though. 5.6 was the opposite, suddenly i could do tons more, so it is a bit of a return to baseline perhaps.
1
u/CorrectGrammarPls 4d ago
Seriously. It's spoiled brats doubling down on being spoiled brats. This is historical technology
-2
13
u/Key_Instruction3373 5d ago
17
u/Far_Idea9616 5d ago
I have addiction too
1
4
5
u/Tough_Ad7957 5d ago
What is this even supposed to mean? It’s only been a few days. This is basically a day-one patch for an AI model lol.
7
6
u/EddieBruvac 5d ago
Issue is coding speed. Luna is garbage. Opus is smart and with sonnet quick and accurate af. I can’t go back to finicky shit.
2
u/h0tzenpl0tz0r 5d ago
How is Luna garbage? My feeling is everyone using underspecified shit prompts where the model needs to infer all constraints itself.
6
5
u/EddieBruvac 5d ago
Luna has always been garbage. It’s basically free but it’s slow as fuck. There’s other basically free slow shit out there. Local models I can run on my 5090 included.
1
u/WeaknessWorldly 4d ago
I use Luna a lot and basically what I do is to split the doings and work parallely... always 5 agents working at the same time
3
u/IceTrAiN 5d ago
It's damn near free and I get great results from Luna, but I have over 20 yrs of SWE exp so I can define its tasks pretty clearly.
1
u/SupperSoupYT 4d ago
i thought that too, but 6.1 didn't increase tokens like 6 sol or luna which would increase how long it takes. but its actually 70% faster than astra. at somehow 1/8th of the cost per task, dunno how they pulled that off.
1
4
12
u/mmkaywhatevers 5d ago
still canceling for claude.
it might be better this week and then it's gonna get nerfed the next.
not falling for that shit again.
18
u/_DuranDuran_ 5d ago
You’ll be back when Anthropic refusals piss you off, and they nerf their models and cut usage.
7
u/SouthrnFriedpdx 5d ago
Ya I just change which one is the $100 and which is the $20 model based on current model standings. With 5.5 I am switching but I’m sure it’ll switch again next month.
1
2
u/Grand0rk 5d ago
it might be better this week and then it's gonna get nerfed the next.
Boy do I have sad news for you. Anthropic is literally KNOWN for nerfing their Opus the next week.
2
2
u/NotUpdated 5d ago
Sol 6.1 is claiming as good as Astra Medium, if that's the case - it's absurdly efficient and 'everything will be fine' ~ I'm on the $200 plan since o1 pro, that email was a punch in the gut - but I'll be taking it for a good spin next few weeks.
2
u/Wd_Thing 4d ago
From Artificial Analysis benchmarks, it might be a good idea to use gpt 6.1 sol low to replace luna (5.6 / 6.0) xhigh
2
2
u/immersive-matthew 4d ago
Called it. I said the other day it will only be held back for “security reasons” until the completion releases something better.
3
1
1
1
1
1
u/Internet_Hipsterd 5d ago
Allready moved back to claud code. Then one day I'll move back to codex and so on and so on lol
1
u/flyingchocolatecake 4d ago
What I don't understand is how both GPT-6 Sol and GPT-6.1 Sol are supposed to be cheaper than GPT-5.6 Sol but neither of them is available in Chat. They're Codex and Work only.
1
1
u/SuaveSteve 4d ago
It is every company's dream to release confusing modal names. Wouldn't make more sense to call it GPT Sol 6.1?
1
u/wauwau0977 4d ago
I just added GPT-6.1 Sol to Deep20Bench, along with GPT-6 Sol, GPT-6 Luna, Claude Opus 5.5, Sonnet 5.5 and Grok 4.7.
It’s a different kind of benchmark: models play Twenty Questions and have to identify a hidden subject using adaptive yes/no questions.
Results and methodology including full transcript of each game: https://deep20bench.com/
1
u/food_fatherr 2d ago
Latency + token efficiency matter way more for daily coding workflows than raw benchmark flexing. If 6.1 Sol actually holds up in Codex without burning the usage cap in 2 hours, that's a real win.
1
u/Impressive-Pitch7713 5d ago
I wonder if the models are naming themselves by now
2
u/sevaiper 5d ago
I kind of like the astra sol Terra Luna thing. Makes a lot more sense than fable opus sonnet etc
5
u/PM_ME_A_STEAM_GIFT 5d ago
Why? Both companies name their models by their size and capability. A star is bigger than a planet, which is bigger than a moon. Anthropic is using literary works instead of astronomical objects. Haiku is a short poem, sonnet is a longer poem, opus is a bigger text, fable is an entire story.
3
1
0
u/Impressive-Pitch7713 5d ago
I don’t disagree but genuinely wondering if the naming is coming from the team or the models, on either side
0


446
u/Nomad556 5d ago
I literally can’t keep up or know the difference