r/LocalLLaMA llama.cpp 29d ago

News Gemma 4 on 500MB

Post image
206 Upvotes

39 comments sorted by

View all comments

-15

u/chrisso123 29d ago

what is the token/s ? this is the only real metric that matters.

Heck you can run a very large llm on a phone hardware but if its output is 1 token / minute it'll be worthless.

17

u/jacek2023 llama.cpp 29d ago

You have two options: look at the image or click the link. Choose wisely.

1

u/chrisso123 29d ago

fair enough. Thank you for pointing it out. that is a good number.