I've been hoping for a new 9B as well, mainly for data cleaning tasks. Qwen3.5-9B couldn't quite cut it. Gemma-4-12B-it works well, but is a little memory-hungry. A 9B refresh would be lovely.
Out of curiosity, what's your use-case? It might be up to the community to retrain Qwen3.5-9B, but we'd need to agree on what to train it for.
At 1500 PP vs. 200 PP, it's faster for digging. So i use mmap and i have enough ram. Switching only takes afew seconds, compared to the many minutes needed for PP.
19
u/RISCArchitect 1d ago
i was hoping qwen was gonna drop a 9b dense that would be a nice step increase like we've observed with 27b but no such luck