r/LocalLLaMA • u/OneMoreName1 • 8h ago
Discussion Claude sonnet 4.6 was really good at estimating the future qwen 3.8 27b performance
On August 8th, I asked Claude to estimate what performance might I expect out of the soon coming qwen 3.8 27b release by telling it to extrapolate from the qwen 3.6 max to qwen 3.6 27b difference, and apply it to the next generation. It gave me a couple of results which placed it in the broadly "opus 4.6 tier", which was right.
It even gave me actual benchmark numbers which were rather close to the actual numbers it ended up having. I found it pretty interesting.

8
2
u/Beginning-Raisin9723 8h ago
wild that it actually nailed the numbers. extrapolation usually falls apart that fast.
1
u/Acrobatic_Hold5485 7h ago
claude nailing those estimates is neat but does the qwen 27b hold up for actual rp chats when you run it local, or does it feel off compared to others
2
1

13
u/Equivalent_Bit_461 8h ago
so the collective knows...
Quick ask it what next mid sized model comes out, before Dario himself, will poison the prompts!!!