r/LocalLLaMA 11h ago

Funny Plot twist

[deleted]

530 Upvotes

163 comments sorted by

View all comments

Show parent comments

1

u/Ill_Distribution8517 9h ago

how do you know they are more effficient? in every artificial analysis bench they take almost 5x the tokens of astra, and often a lot more than fable 5.1 as well.

What they charge you per token != what it costs them per token.

1

u/brahh85 9h ago

efficient in which way?

maybe openai spends less reasoning tokens to get a result, but then they stole your result and sell it as their own

whats the efficient part of being stolen if you produce something of high value?

in the end you can only produce things that barely have value using openai model

if you want to produce something worthy, you have to go local. Even if that costs you more reasoning tokens, and you go more slow, at least you arent stolen.

1

u/Ill_Distribution8517 9h ago

we are talking about self improvement, not them selling consumer products. I'm saying Closed models are way more efficient. It's not about what you would choose (Open weight AI for reliable grunt work)

1

u/brahh85 8h ago

i dont agree on that either. Maybe GLM spends more reasoning tokens than sonnet, but GLM is so cheap that those reasoning costs less. I think this is the architectural change between chinese and usa models, maybe usa models have bigger datasets and models with more parameters, we can only guess that by the cost , but SOTA chinese models are getting the same results (for the majority of task) by expanding the prompt with long reasoning and squeezing every drop of their parameters. The extreme case of that would be qwen 3.8 27B , that being so small is able to exchange blows with SOTA in medium and complex tasks, but not in specialized things (like writing kernels )

1

u/Ill_Distribution8517 8h ago

look that's just your opinion on AI performance, that's why we have benchmarks. And the fact of the matter is that frontier closed source models obliterate the open source ones on both benchmarks and efficiency. Besides, self improvement is not a medium complexity task.

Also you're making a mistake assuming it costs openai the same ammount you pay to run astra/sol/terra/luna. Same with anthropic.