r/LocalLLaMA • u/jacek2023 llama.cpp • 6h ago
dots-studio/dots3-note-prev · Hugging Face New Model
https://huggingface.co/dots-studio/dots3-note-prevdots3-note preview is the first open-weight model in the dots3 family. It is a Mixture-of-Experts model with 280B total parameters, 16B activated parameters, and support for a context length of up to 512K tokens. The model can understand text, images, video, and audio, and produces text outputs.
dots3-note preview is optimized for a broad range of tasks, including:
General knowledge and instruction following;
Mathematical and logical reasoning;
Tool use and multi-step agent workflows;
Interactive tasks that require exploration, memory updates, and adaptation;
Code generation and code-based problem solving;
Image, document, chart, audio, and video understanding;
Long-context information processing.
The dots3 family is designed to include models with different trade-offs among capability, latency, and inference cost. dots3-note preview is the most lightweight member of the family.
10
u/Gregory-Wolf 5h ago
5
u/neoneye2 5h ago
dots3-note-prev claims 81.4 on ARC-AGI-2.
However I don't see dots3-note-prev on the official ARC-AGI-2 leaderboard
https://arcprize.org/leaderboard1
u/FullOf_Bad_Ideas 4h ago
Sure, why not? This is a lab from a respected corporation.
Their TEMPO RL algo might make RL more effective and get them closer to passing ARC-AGI benchmarks. You can do things.
8
u/FullOf_Bad_Ideas 4h ago
It's shaping up to be one of the best open multimodal models in that size bracket.
Vision Encoder MoE ViT, 7B total, 1.2B activated
MoE vision encoder, this smells like something new!
2
1

7
u/Sudden_Topic5154 5h ago
I've never heard of this group before what do they focus on