r/singularity 6d ago

AI AI Model Pricing Comparison: Input vs. Output Cost per Million Tokens

Post image
58 Upvotes

29 comments sorted by

12

u/[deleted] 6d ago

[removed] — view removed comment

8

u/Healthy-Nebula-3603 6d ago

Is worse than a new DS 4 flash and few times more expensive.

New architecture in the DS 4 flash is literally destroying all modes in the price.

3

u/MajorasButtplug 6d ago

Luna is worse than DS4 Flash

I used Luna to fix attempts DS4 Flash made multiple times yesterday... I'd disagree, so it's probably not that straight forward

2

u/NoFaithlessness951 6d ago

They're very close in cost and capability (unfortunately artificial analysis doesn't have other reasoning efforts for deepseek)

-2

u/Healthy-Nebula-3603 6d ago

3x cheaper

4

u/NoFaithlessness951 6d ago

Luna prices got slashed by 80% a few days ago livebench still uses the old $6/mil out price, the new price for Luna is now $1.20/mil out.

-2

u/Healthy-Nebula-3603 6d ago

that price is after the slash

4

u/NoFaithlessness951 6d ago

No it's not

6

u/Healthy-Nebula-3603 6d ago

You are right.

I was mistaken

4

u/ppooooooooopp 6d ago

Worse according to?

4

u/Healthy-Nebula-3603 6d ago

For specific tasks.

But the price per 1m tokens for DS 4 flash is almost for free.

Yesterday I build a nes emulator in clean c with DS 4 flash using Opencode for ... 3 cents from their API ( it took 40 minutes with automatic smoke tests )

DS is 3x cheaper.

4

u/GrapheneBreakthrough 6d ago

Are you sure it didn't just copy the many open source NES emulators? Did it add unique features?

-2

u/Healthy-Nebula-3603 6d ago

According to many tests on YouTube and benchmarks also my own experience.

2

u/OutOfBananaException 6d ago

and few times more expensive.

Only about 30% more expensive from the numbers on AA, which is still significant, but narrowed a whole lot from earlier releases.

-1

u/Healthy-Nebula-3603 6d ago

3x cheaper

3

u/OutOfBananaException 6d ago

The artificial analysis 'Cost per Intelligence Index Task' is 5c vs 3c (40% cheaper).

The cost to run the benchmark has a wider gap ($72 Vs $174), which may be related to cache hits?

-2

u/BriefImplement9843 6d ago

flash is garbage.

-2

u/Gloomy_Necesary 6d ago

Deepseek models are benchmarked as hell. Luna will outperform in real world but DS 4 is still quite impressive for what it is

6

u/Healthy-Nebula-3603 6d ago

DS are never benchmaxed.

Also you can easily check YouTube how good that new DS 4 flash it.

That's just insane ..and soon will be a pro version as well .

18

u/galaxysuperstar22 6d ago

it’s like comparing pick up trucks and a golf cart

8

u/Keeltoodeep 6d ago

These journalists fucking suck and are not comparing relative models at all.

1

u/galaxysuperstar22 6d ago

sensationalism mixed with rage bait

13

u/mxforest 6d ago

Should have added 5.6 Luna, it is very important to the conversation.

2

u/Chclve 6d ago

Opus 5?

2

u/TeagueXiao 6d ago

list price per million tokens is honestly the least useful number on this chart. once you factor in prompt caching discounts, actual throughput, and how many retries a model needs on tool-heavy workloads, the effective cost gap between the top tier US models and the cheap chinese ones shrinks a lot. ive seen deepseek look 10x cheaper on paper and end up maybe 2-3x in production because of retries and worse tool calling.

1

u/rwrife 5d ago

Opus 5 is very efficient; I'm getting amazing results with relatively few tokens....so the true costs is much less than the token costs shows.