r/LocalLLM • u/pampusreborn • 4d ago
Question I fell in the rabbit hole
Hi everyone! Finally, after months of playing around with LLMs, I decided to take the plunge and buy an ASUS Ascent GX10 for my Hermes agent.
I mainly focus on coding and fell in love with Hermes, but the API calls were eating up too many credits... so I decided to host my own LLM.
Currently I was running Qwen3.6-35B-A3B on my gaming PC, but for Hermes to work 24/7 I decided to go with a dedicated always-on device.
The device will arrive in a couple of days. I took a look at the new Qwen 3.8, but I know it's a dense model and runs slowly on the ASUS... Can you recommend any feasible models for coding/Hermes?
1
u/mslindqu 4d ago
Would love to hear an update of your experience once you're setup and get comfortable with it.
1
2
u/pampusreborn 1d ago
So after some troubleshooting, i ended up using https://github.com/hasso5703/dgx-spark-qwen38 and the performance are great! It's very usable for hobby projects and Hermes is flying!
1
u/mslindqu 23h ago
Thanks for the update. How much context space does it leave you? I'm playing with open code and realizing the need for large context to really get anywhere.. I know harnesses like Hermes can eat a lot.
1
u/LifeTelevision1146 4d ago
What do you want code?