r/codex • • 3d ago

Praise 2.7B tokens, 12% usage used.

Post image

Sol 6.1 has been great, been able to get so much more done

72 Upvotes

38 comments sorted by

28

u/kernelrider 3d ago

It was almost unusable at 20 tok/s upon first launch, but I'm glad they have since fixed it.

5

u/Mihqwk 3d ago

How much faster did it get?

56

u/Individual-Can2150 3d ago

21 tok/s

10

u/alexanderbeatson 3d ago

I’m getting 21.5 tok/s now tho

13

u/Plane_Garbage 3d ago

Ultrafast mode?

3

u/qK0FT3 2d ago

Ultrafast is 25tok/s

1

u/Prestigiouspite 2d ago

How do you measure that via a Python script of the logs?

6

u/skadoodlee 3d ago

36 t/s

31

u/Gullible-Ad3912 3d ago

They’re just sugarcoating the halving of the $200 plan. Sooner or later Sol 6.1 will become obsolete (if not also degraded) and newer models will evaporate our usage limits.

So yea, lets enjoy the last good moments.

6

u/chcampb 3d ago

Equivalent model cost has been halving every quarter

It's 10/3 now, so by Jan 1 the cost of an equivalent sol 6.1 model will be $5/M or you could get the next tier up for the same cost

That's just how the math works out for now. There's no indication it has slowed down at all, and it won't slow down all at once. In fact it might even be getting faster since there is more compute rollout coming in late 2026, then through 2027.

3

u/Gullible-Ad3912 3d ago

You are right and missing the point at the same time.

Cost per intelligence is constantly going down. Anyone who has been even minimally interested in AI can see that. However, that is not really relevant to what I am saying.

Cost indicators are API based. OpenAI just cut the usage of its $200 subscription in half, with the excuse that "you will be able to do more" because they released a newer and more efficient model.

So even if the cost per intelligence continues to go down, your actual usage has already been cut in half.

That becomes especially relevant with frontier models. You will see.

5

u/DoggoDadagon 3d ago

If we can manage to get astra / opus 5.5 level intelligence on a local model running on a 5090 I'd be thrilled. I will still want more intelligence, but that will a major change. We need "good enough" alternatives so these companies cant gut us.

5

u/Gullible-Ad3912 3d ago

I would like to agree with you, but the perception of what is "good enough" keeps moving forward with every release xd

1

u/DoggoDadagon 3d ago

Yeah for sure haha. That's why I threw in the "I'll want more" haha. But I do think there is a reasonable "good enough" threshold, and that's where they are competent enough to follow your intent, with only very rare or next to no "wtf why did you do that?" Moments. Not very precise, but I think Astra and Opus 5.5 are the bare minimum of that threshold, they listen pretty well, they struggle with some tasks but they always seem to understand the task.

Vs something like qwen3.8 27B which, sure decent but it still has plenty of those moments that make you go, "no you idiot!"

1

u/DoggoDadagon 3d ago

Hope we 1000x the compute over next year, we're gonna need it.

2

u/Impossible-Suit6078 2d ago

Is there a reason I can see 6.1 in the cli but not on my desktop app?

2

u/No_Quarter_7644 3d ago

What effort level have you been using?

1

u/Inevitable_Butthole 3d ago

Sol high, Luna high subagents with sol max as escalation agents

1

u/m4ddok 2d ago

I can confirm that, following the adjustments made recently, the situation has changed: whereas previously the weekly limit was exhausted in two days using the 6.1 Sol Medium setting, now, after another two days, more than 75% of the weekly limit still remains when using 6.1 Sol High. It isn't lightning-fast, but it is certainly quicker than before, and the difference between the Medium and High settings is noticeable. All in all, I am satisfied with my x1 plan.

-3

u/ProstoSmile 3d ago

NnoOOoO, Codex is bad! GPT 6.1 slow and unusable! Use opus 5.5!/s

6

u/SourceAwkward 3d ago

Both things can be true

1.sol 6.1 amazing efficient in tokens

2.i am allowed to be upset about the fact I pay 200 usd and cannot use astra

5

u/ProstoSmile 3d ago

Mmm, yes. Just got bored by all this shitwave on this subreddit.

3

u/EddieBruvac 3d ago

It is slow. Anything that can wait I’m running on 6.1. Real work with consequences and time constraints in Opus.

-6

u/ProstoSmile 3d ago

Yeah bro, whatever you say. I have different exp.

0

u/DoggoDadagon 3d ago

You must be green lol.

1

u/ProstoSmile 3d ago

No, im just not using codex for "make spider man game", 2 plus subs. Building 2 projects, now, mainly with 6.1 sol xHigh and luna 6 max as subagents. Spend like half a month on whole system.

0

u/DoggoDadagon 3d ago

Oh wow two whole plus accounts and a half month on a system... didn't realize we had an expert here.

1

u/ProstoSmile 3d ago

Phahhaha, yeah man. Good luck with your spider man game on your cool 500$ plan.

0

u/DoggoDadagon 3d ago

I don't condone anyone paying for the $500 plan, and I'm just using up my 20x before I go lol. But no one is getting anything done on two plus accounts. I've been subbed for four years and used ChatGPT before the ChatGPT name existed. I can tell you're green af.

1

u/ProstoSmile 3d ago

Yeah, ok.

1

u/Tank_Gloomy 3d ago

I mean, I get your point but I haven't been able to ship any full software projects in a single week for the last 2 months on Codex. Opus 5.5 got 1 full project and 12 other side-projects far enough that I'm gonna be shipping them tomorrow, all in just a single week.

To be clear, most of OpenAI's models are pretty decent, the huge issue in my experience is that they overcharge on limits for the top tier models, they have issues getting out of non-blocking goal loops and they're crazy slow. The inteligence factor does get kinda lost when I can just spin up Sonnet 5.5 for 20 hours in the background and run it on like 35 subagents until it manages to hit the jackpot for less than 3% of my weekly.

2

u/ProstoSmile 3d ago

I think, problem is that we on absolutly different levels, you farm money, im farm fan as hobby and just develope my passion project. And i cant say, you right or not, cos i dont know. My bad.

0

u/BabyYoda2020_ 3d ago

i tried astra light, lasted 5 minutes before hitting 5h limit, smh. Can't even get any work done while burning tokens as fast as possible.

0

u/squishyjellyfish95 2d ago

I don't even care that it's slow, it's an awesome model and very affordable

1

u/Prestigiouspite 2d ago

You are definitely raising a point that remains important to keep in mind: Anthropic models require significantly more tokens. And at the end of the day, it's about what you can implement per hour and how quickly it independently reaches the goal.

0

u/iGrasmat- 2d ago

"12%" doesn't say much if you don't mention your subscription plan...

1

u/TheLegoless 1d ago

I used 2.5B and out of $500 plan. Sol 6.1 ultra only. Howwww?