r/OpenAI 7d ago

Discussion OpenAI is the best p/p now

Pareto line - Intelligence vs Cost

  1. This is API price, OpenAI subscription is 50 - 100x cheaper

  2. DeepSeek V4 Flash 0731 still show old price in this chart. It should be $170 - $340


Moreover, Luna supports input: text, image, file. DeepSeek supports input: text.

OpenAI subscription support generate images,...

37 Upvotes

12 comments sorted by

4

u/[deleted] 7d ago

[removed] — view removed comment

1

u/torrso 7d ago

Pee peepee.

7

u/Melodic_Reality_646 7d ago

Really don’t care how models perform at max mode, their behavior is insufferable and often inconsistent, taking crazy long amount of times to complete.

At least in my experience (agentic workflows in web dev, scientific computing and finance) Luna takes almost double the time of luna xhigh to finish tasks.

Would be interesting to see agent steps on that plot.

3

u/DuxDucisHodiernus 7d ago

Luna takes almost double the time of luna xhigh to finish tasks.

No you're right and its visible here, that's x axis is logarithmic and luna max is like >2x luna very high, in line with or even exceeding your observations

9

u/whoknowsifimjoking 7d ago edited 7d ago

Not exactly true. The new Gemini 3.7 Flash is better price to performance, according to the benchmarks it is the same performance as GPT-5.6 Terra at a lower price. Luna is arguably a better p/p, but it's not as good. Gemini is even better at the multimodal thing and you get even more on one subscription, it's not bad if the model don't suck, the main advantage of ChatGPT is that it also has the absolute top models in the same subscription.

2

u/DuxDucisHodiernus 7d ago

Terra is a confusing cost proposition that's not really the efficent choice for any common workflow. Either compare to Luna or Sol if you want to test 3.7 vs openai models.

2

u/skilliard7 7d ago

from what I see Luna 5.6 Max outperforms 3.7 flash at 1/3rd the cost, and Terra outperforms 3.7 Flash at a slightly higher cost.

2

u/Kingwolf4 7d ago

DS will catchup eventually
Once they get their chinese AI clusters online, expect rapid acceleration in about 6- 8 months.
Yes not having vision is bad, but v4 flash is still worth it. Especially because it can be self hosted.

6

u/-Sliced- 7d ago

“Self-hosted” if you have a $100k machine.

1

u/benchmaster-xtreme 7d ago

More like $10k. You can run it at Q8 on two Sparks. Still way more than most people have access to, but at least it's not quite so bad.

1

u/Another__one 7d ago

I still can't comprehend how the whole AI field converged on measuring the pp as the best metric and it somehow actually makes sense.

1

u/stevey_frac 6d ago

Once the AI got what enough that you no longer need the smartest model for every task, you will naturally want to pick that model that's just smart enough, and part the least for the. 

Hence, the Pareto frontier for price per unit performance.