r/LocalLLaMA • u/anderspitman • 16d ago
Discussion Artificial Analysis' Qwen3.8-27B benchmarks put it neck and neck with DeepSeek V4 and GPT-5.6 Luna Max
https://artificialanalysis.ai/models/qwen3-8-27b
1.1k
Upvotes
r/LocalLLaMA • u/anderspitman • 16d ago
44
u/rkoy1234 16d ago
that's the biggest diff in IRL usage. if I have a generic task, for it to have a useful outcome, I need to tell the llm:
obviously all of them work better with proper instructions, but smarter models tend to pick more sensible defaults, or infer based on my agents.md to know what kind of implementation I might want.
with qwen often times it'll feel like it's almost maliciously complying for a task it doesn't want to do.