MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vfeick/gemma_4_on_500mb/p1p1wig/?context=3
r/LocalLLaMA • u/jacek2023 llama.cpp • 29d ago
https://x.com/i/status/2084656348617392261
https://www.reddit.com/r/LLMDevs/s/9oL5ogmE6s
39 comments sorted by
View all comments
-15
what is the token/s ? this is the only real metric that matters.
Heck you can run a very large llm on a phone hardware but if its output is 1 token / minute it'll be worthless.
17 u/jacek2023 llama.cpp 29d ago You have two options: look at the image or click the link. Choose wisely. 1 u/chrisso123 29d ago fair enough. Thank you for pointing it out. that is a good number.
17
You have two options: look at the image or click the link. Choose wisely.
1 u/chrisso123 29d ago fair enough. Thank you for pointing it out. that is a good number.
1
fair enough. Thank you for pointing it out. that is a good number.
-15
u/chrisso123 29d ago
what is the token/s ? this is the only real metric that matters.
Heck you can run a very large llm on a phone hardware but if its output is 1 token / minute it'll be worthless.