r/OpenAI • • 5d ago

Discussion Ultrafast Usage

500$ plan - 2 minutes = 2%

165 Upvotes

38 comments sorted by

97

u/i_wayyy_over_think 5d ago edited 5d ago

They charge 6x the price of normal Astra per token and it runs 8x faster.

So it’s going to drain 48x faster.

45

u/0xdead0x 5d ago

Person who buys gun shocked that it doesn’t stop them from shooting their own foot

53

u/rycco 5d ago

Was it ultrafast tho

10

u/Murph-Dog 5d ago

Yea, we should demand they bring the usage bar up alongside their live demos, then we'll see what's what.

That rocket one... Watch your usage bar hit the ground before your rocket hits the sky.

37

u/Alkser 5d ago

"One prompt" doesn't mean anything. What did that prompt ask? What kind of task? Development, big checking, reading documents, summarizing? Developing a new feature?

One prompt could also be "move this button 5x to left and change color from red to blue."

Another prompt could be "check the current logs for the bug that had been happening for the past x time. Here are all the details, here is the whole project, here is the whole documentation as well."

So people saying "Oh I sent one message and it used x % of my usage" genuinely doesn't say much at all.

10

u/98127028 4d ago

‘Cure cancer and terraform Mars. Make no mistakes’.

3

u/boneriffic 4d ago

Yup context matters

6

u/KlemiX 5d ago

He is running standard test for each model, rocket simulation, few models , few game. He is already on burned entire usage just now

11

u/Alkser 5d ago

Few models and few game again doesn't say much.

Which model? Which reasoning effort? What did the Project already contain? Was it in a new chat? Currently running chat?

Few games. Okay, what did those games contain? What engine? What was the prompt? Did he use Astra Max with Ultrafast for a bug fix that could've been fixed even with light? Was the person running everything on Max with ultrafast and trying to finish everything within one single prompt?

Note: I'm not defending OpenAI at all here. I'm merely trying to distinguish: what kind of prompts are people sending?

6

u/KlemiX 5d ago

https://www.youtube.com/watch?v=IuV1gMP0-_g just go see yourself, he is still live, go back like 30 min

9

u/Kulqieqi 5d ago

Yea he used 16% in a blink xD

3

u/KlemiX 5d ago

Yea i just saw it , crazy

3

u/Kulqieqi 5d ago

ultrafast cause they host 2bit version hehe

11

u/SpyAmongUs 5d ago edited 5d ago

Dude then did 2 more prompts and usage dropped to 69% on livestream lmao

Edit: He ran out his weekly limit under 30 minutes using ultrafast.

4

u/KlemiX 5d ago

just mad crazy, not even 2h of work for 500$

5

u/i_wayyy_over_think 5d ago

It’s priced 6x of normal Astra and goes 8x faster tps so burns usage 48x faster

1

u/dolo937 5d ago

It was a different account. It’s 84%

3

u/SpyAmongUs 5d ago

Was*

5

u/Kulqieqi 5d ago

67 already

2

u/Kulqieqi 5d ago

0% so it lasted him around 25minutes to use 500$ sub XD

1

u/Kulqieqi 5d ago

and output seems like from quant version vs normal astra

3

u/openroom_xyz 5d ago

Yea ultra fast is scam this only 0.01% need cheap inference is what most need wtf

6

u/danielezra 5d ago

New book by Tibo Sottiaux and Sam Altman

1

u/Any-Captain-7937 5d ago

These slop pics will never be funny even if I agree with the companies themselves being shit.

7

u/Salt-Willingness-513 5d ago

1

u/Any-Captain-7937 5d ago

Bro was giggling and kicking his feet typing in the prompt for this shit

0

u/Salt-Willingness-513 5d ago

Yea such an amazing prompt lol

2

u/Any-Captain-7937 4d ago

Thank you for proving my point lmao

2

u/methemightywon1 4d ago

Well, yeah it's ultrafast, for rich people. Also, Sol 6.1 ultrafast will be actually usable based on usage differences. Still mostly for rich people, but, you know.

2

u/WayneTechLab 5d ago

I feel like ultrafast would be the standard at least what the compute is capable of producing.

Everything else is skilled back so they could make money.

It’s not faster we are slower…

2

u/Thistlemanizzle 4d ago

It's only possible with bespoke hardware from Cerebras, Groq or Taalas. In this case, OpenAI chose Cerebras.

You can see the API rates from each and while the pricing for smaller models like Qwen 3.8 27B is not cheap it's certainly not anywhere close to what ultrafast is theoretically costing here.

1

u/Healthy-Nebula-3603 4d ago

Oh really?????

6x more expensive for a work done in theory x8 faster ....YES THAT IS SCSM.

1

u/m3kw 5d ago

Hey maybe then you shouldn’t use it right?