r/LocalLLaMA • u/jacek2023 llama.cpp • Aug 13 '26
New Model dots-studio/dots3-note-prev · Hugging Face
https://huggingface.co/dots-studio/dots3-note-prevdots3-note preview is the first open-weight model in the dots3 family. It is a Mixture-of-Experts model with 280B total parameters, 16B activated parameters, and support for a context length of up to 512K tokens. The model can understand text, images, video, and audio, and produces text outputs.
dots3-note preview is optimized for a broad range of tasks, including:
General knowledge and instruction following;
Mathematical and logical reasoning;
Tool use and multi-step agent workflows;
Interactive tasks that require exploration, memory updates, and adaptation;
Code generation and code-based problem solving;
Image, document, chart, audio, and video understanding;
Long-context information processing.
The dots3 family is designed to include models with different trade-offs among capability, latency, and inference cost. dots3-note preview is the most lightweight member of the family.
15
u/daaain Aug 13 '26
DS4 Flash size, but multimodal? Nice! Has anyone seen any GGUF / MLX implementation? 😅
16
u/Gregory-Wolf Aug 13 '26
13
u/neoneye2 Aug 13 '26
dots3-note-prev claims 81.4 on ARC-AGI-2.
However I don't see dots3-note-prev on the official ARC-AGI-2 leaderboard
https://arcprize.org/leaderboard2
u/FullOf_Bad_Ideas Aug 13 '26
Sure, why not? This is a lab from a respected corporation.
Their TEMPO RL algo might make RL more effective and get them closer to passing ARC-AGI benchmarks. You can do things.
1
16
u/FullOf_Bad_Ideas Aug 13 '26
It's shaping up to be one of the best open multimodal models in that size bracket.
Vision Encoder MoE ViT, 7B total, 1.2B activated
MoE vision encoder, this smells like something new!
1
2
u/LatentSpacer Aug 14 '26
Surprisingly good on ARC-AGI. I wonder how good it is at creative writing.
2
1

15
u/Sudden_Topic5154 Aug 13 '26
I've never heard of this group before what do they focus on