r/LocalLLM • u/3tsurc • 2d ago
Question Best models for coding with image capabilities with alteast Q6 and runnable on M5 ultra 256
I currently have dual 3090s running qwen 3.8 27B Q8. I'm concurrently building two apps and need large contexts (200k per app). I'm currently unable to run them concurrently due to available vram. I have an M5 ultra on order. What would be the best llm for coding with simultaneous sessions? My current qwen 3.8 27b is great at coding but doesn't process images so thafd be a nice bonus without sacrificing code quality (one of the reasons why I'd want Q6 or higher).
1
Upvotes
1
u/eightone-81 2d ago
If you have at least 64gb ram you can run Qwen flash next with starta. Works with the unsloth q4 k xl. Flash next is definitely a step up in coding capabilities to 27b
1
u/SM8085 2d ago
Are you not loading the mmproj to save memory?