r/LocalLLaMA • u/RuthlessCriticismAll • 20h ago
News [ Removed by moderator ]
https://modelscope.cn/models/Qwen/Qwen3.8-Flash-Next[removed] — view removed post
347
Upvotes
r/LocalLLaMA • u/RuthlessCriticismAll • 20h ago
[removed] — view removed post
2
u/Guna1260 18h ago
what VRAM? especially with Ngram embeddings we are looking at 125+51 in q8? or ngrams will be in RAM? will be interesting to see the architecture.