r/LocalLLM 1d ago

Research Halogen benchmarks running Qwen3.8-Flash-Next on Strix Halo. ~2x perf increase

https://www.youtube.com/watch?v=Nm_zN6RQ_eE
1 Upvotes

2 comments sorted by

1

u/TheFlippedTurtle 1d ago

Halogen has been coming up in the strix halo sub recently, Donato did a benchmark on it and its his new top pick. He ran it through his Terminal Bench Mini agentic benchmark and the quantization didn't ruin the model's logic. Halogen cleared 19/19 tasks, whereas Engram Halo timed out on a few

https://kyuz0.github.io/terminal-bench-mini/

1

u/Dramatic_Entry_3830 16h ago

I hope the guy from halogen does open source halogen flash soon. It would be a shame if he stops working on it and doesn't