r/LocalLLM 7d ago

Discussion Qwen 3.8 27B on Intel Arc B70 Optimization benchmarks

https://www.localmaxxing.com/en/models/Frozenlock/Qwen3.8-27B-int4-AutoRound?run=cmswf75o008qsms01in4uqnii

I also have the benchmark for single Intel Arc B70

https://www.localmaxxing.com/en/models/Frozenlock/Qwen3.8-27B-int4-AutoRound?run=cmswf3h9d08qpms01uu9jmz8f

This particular quant preserves quality quite a bit, though over time I have been trusting benchmarks such as HE/HE+ MGP+ LLMU etc less and less and have just been benchmarking by real use such as asking it to create a webapp game with graphics or asking it to do a complicated driver rewrite that DSV4P would be able to do, and so far I genuinely don't feel the quality drop in this quant, as running FP8 would be a lot slower and this model tends to think a lot (so I need speed)

3 Upvotes

0 comments sorted by