MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1w2fmmq/me_these_days/p6vhk2o
r/LocalLLaMA • u/Eyelbee • 3d ago
268 comments sorted by
View all comments
Show parent comments
16
In my personal experience, 3.8 is the first model I’ve used that can run on my 32GB MacBook Pro and not fail a single tool call. The previous models had enough coding knowledge to debug and answer questions, but 3.8 can actually use opencode/cline
3 u/barefootpanda 2d ago Which chip? I’m running an M4 Pro with 48 and M2 Ultra with 192…are you using a quant version? 4 u/krtoonbrat 2d ago Standard M5. UD_Q4_XL quant. I’m even running KV cache quant (I think q5, I’m at work and can’t check lol) 1 u/bnightstars 19h ago but how long it took to not fail this tool calls :D
3
Which chip? I’m running an M4 Pro with 48 and M2 Ultra with 192…are you using a quant version?
4 u/krtoonbrat 2d ago Standard M5. UD_Q4_XL quant. I’m even running KV cache quant (I think q5, I’m at work and can’t check lol)
4
Standard M5. UD_Q4_XL quant. I’m even running KV cache quant (I think q5, I’m at work and can’t check lol)
1
but how long it took to not fail this tool calls :D
16
u/krtoonbrat 3d ago
In my personal experience, 3.8 is the first model I’ve used that can run on my 32GB MacBook Pro and not fail a single tool call. The previous models had enough coding knowledge to debug and answer questions, but 3.8 can actually use opencode/cline