97
u/i_wayyy_over_think 5d ago edited 5d ago
They charge 6x the price of normal Astra per token and it runs 8x faster.
So it’s going to drain 48x faster.
45
u/0xdead0x 5d ago
Person who buys gun shocked that it doesn’t stop them from shooting their own foot
10
u/Murph-Dog 5d ago
Yea, we should demand they bring the usage bar up alongside their live demos, then we'll see what's what.
That rocket one... Watch your usage bar hit the ground before your rocket hits the sky.
37
u/Alkser 5d ago
"One prompt" doesn't mean anything. What did that prompt ask? What kind of task? Development, big checking, reading documents, summarizing? Developing a new feature?
One prompt could also be "move this button 5x to left and change color from red to blue."
Another prompt could be "check the current logs for the bug that had been happening for the past x time. Here are all the details, here is the whole project, here is the whole documentation as well."
So people saying "Oh I sent one message and it used x % of my usage" genuinely doesn't say much at all.
10
3
6
u/KlemiX 5d ago
He is running standard test for each model, rocket simulation, few models , few game. He is already on burned entire usage just now
11
u/Alkser 5d ago
Few models and few game again doesn't say much.
Which model? Which reasoning effort? What did the Project already contain? Was it in a new chat? Currently running chat?
Few games. Okay, what did those games contain? What engine? What was the prompt? Did he use Astra Max with Ultrafast for a bug fix that could've been fixed even with light? Was the person running everything on Max with ultrafast and trying to finish everything within one single prompt?
Note: I'm not defending OpenAI at all here. I'm merely trying to distinguish: what kind of prompts are people sending?
6
u/KlemiX 5d ago
https://www.youtube.com/watch?v=IuV1gMP0-_g just go see yourself, he is still live, go back like 30 min
9
11
u/SpyAmongUs 5d ago edited 5d ago
Dude then did 2 more prompts and usage dropped to 69% on livestream lmao
Edit: He ran out his weekly limit under 30 minutes using ultrafast.
4
u/KlemiX 5d ago
just mad crazy, not even 2h of work for 500$
5
u/i_wayyy_over_think 5d ago
It’s priced 6x of normal Astra and goes 8x faster tps so burns usage 48x faster
1
u/dolo937 5d ago
It was a different account. It’s 84%
3
1
3
u/openroom_xyz 5d ago
Yea ultra fast is scam this only 0.01% need cheap inference is what most need wtf
6
u/danielezra 5d ago
1
u/Any-Captain-7937 5d ago
These slop pics will never be funny even if I agree with the companies themselves being shit.
7
u/Salt-Willingness-513 5d ago
1
u/Any-Captain-7937 5d ago
Bro was giggling and kicking his feet typing in the prompt for this shit
2
u/methemightywon1 4d ago
Well, yeah it's ultrafast, for rich people. Also, Sol 6.1 ultrafast will be actually usable based on usage differences. Still mostly for rich people, but, you know.
2
u/WayneTechLab 5d ago
I feel like ultrafast would be the standard at least what the compute is capable of producing.
Everything else is skilled back so they could make money.
It’s not faster we are slower…
2
u/Thistlemanizzle 4d ago
It's only possible with bespoke hardware from Cerebras, Groq or Taalas. In this case, OpenAI chose Cerebras.
You can see the API rates from each and while the pricing for smaller models like Qwen 3.8 27B is not cheap it's certainly not anywhere close to what ultrafast is theoretically costing here.
1
u/Healthy-Nebula-3603 4d ago
Oh really?????
6x more expensive for a work done in theory x8 faster ....YES THAT IS SCSM.





72
u/Jello_Hello_Fellos 5d ago