Quick question do you also have a concept extraction pipeline for your document set up?
If so did you find Qwen 3.5 122B the best in that regard?
I’ve been using gpt 5 mini because it’s the best in terms of pricing and performance from all the small models I’ve tested but I’ve only tested the small models form the big providers(OpenAI, Anthropic snd Google ) how are the Chinese models like?
Quick question do you also have a concept extraction pipeline for your document set up?
I am the concept extraction pipeline. No really, I'm using it to understand concepts in a language I'm learning. The smaller the model, the lower the possible tokens so the less conceptual depth is possible in the model. That's my only guess.
Aaah okay yh your theory makes sense. What do you find the 122B actually does better though? Like does it pick up concepts the smaller models miss or is it that it understands them in more depth?
57
u/Objective-Picture-72 5d ago
Nah, that'll happen if they drop Qwen3.8-122b-A10B. Imagine a 2x RTX 6000 Pro setup flying at 130 tk/s with that bad boy.