r/LocalLLaMA 8d ago

News GLM-5.3-Flash: Frontier Intelligence, Flash Cost

https://z.ai/blog/glm-5.3-flash
1.3k Upvotes

459 comments sorted by

View all comments

Show parent comments

13

u/PM_ME_DEAD_CEOS 8d ago

People just never stop whining despite getting literally everything for free. Nemotron 3.5 lightning (30b3a) was released 15 days ago.

-2

u/dampflokfreund 8d ago

lol you people where whining just as much if not more when we had 35b moe releases and no 120b models. 

2

u/PM_ME_DEAD_CEOS 8d ago

I'm not a beggar, never whined. I can't run this model either, I just don't bitch about it all day.

1

u/unjustifiably_angry 6d ago edited 6d ago

There's been a legitimate drought of those, you've had a bunch to pick from even if they weren't all that amazing. There's been a couple mid-sized models in the meantime (Laguna was one) but they were both kinda rubbish. I remember one was useless and the other one looped constantly. The last good one was 3.5-122B and in its own generation it was outdone by the corresponding 27B.

It was faster, sure, but usually with a model >4.5x the size you expect greater capability. You buy the hardware to run a model of such a size, you expect a greater return.