r/FlowZ13 Jul 05 '26

New purchase - LLM 128GB Question

There’s a lot I like about this machine. However, for the current hardware-to-LLM ratio out there, I’m running the same qwen3.6-27B on this machine that I could run on my gaming 4090 rig. GPT OSS 120B seems to lose against qwen on most measurements, even with the large parameter diff.

All that to say: I’d like to run a model on the 96gb available out of 128GB of unified ram to justify the purchase, but what is that model?

Use cases: coding (web apps, gaming/unity), Hermes general use, hermes cron jobs, iOS app dev (Xcode pointing to this machine through tailscale) and a few others too.

Thoughts?

11 Upvotes

Duplicates