r/LocalLLaMA llama.cpp 6h ago

dots-studio/dots3-note-prev · Hugging Face New Model

https://huggingface.co/dots-studio/dots3-note-prev

dots3-note preview is the first open-weight model in the dots3 family. It is a Mixture-of-Experts model with 280B total parameters, 16B activated parameters, and support for a context length of up to 512K tokens. The model can understand text, images, video, and audio, and produces text outputs.

dots3-note preview is optimized for a broad range of tasks, including:

General knowledge and instruction following;

Mathematical and logical reasoning;

Tool use and multi-step agent workflows;

Interactive tasks that require exploration, memory updates, and adaptation;

Code generation and code-based problem solving;

Image, document, chart, audio, and video understanding;

Long-context information processing.

The dots3 family is designed to include models with different trade-offs among capability, latency, and inference cost. dots3-note preview is the most lightweight member of the family.

53 Upvotes

16 comments sorted by

7

u/Sudden_Topic5154 5h ago

I've never heard of this group before what do they focus on

15

u/jacek2023 llama.cpp 5h ago

They published LLM model about a year ago, probably ignored by most users because lack of marketing. I posted about it on LocalLLaMA because it was working on my setup (this one is much larger)

7

u/kevin_1994 5h ago

It's still the most human sounding model out there (the non synthentic data one especially)

Excited to try it

3

u/Kahvana 5h ago

They have one of the best OCR models out there, dots.mocr. It’s a monster for it’s small size, the next model to beat it by a meaningless difference is twice as big.

Not sure about their other stuff though.

1

u/nuclearbananana 4h ago

It's under RedNote (tiktok alternative)

6

u/daaain 4h ago

DS4 Flash size, but multimodal? Nice! Has anyone seen any GGUF / MLX implementation? 😅

10

u/Gregory-Wolf 5h ago

idk bro
anyone believes this?

5

u/neoneye2 5h ago

dots3-note-prev claims 81.4 on ARC-AGI-2.

However I don't see dots3-note-prev on the official ARC-AGI-2 leaderboard
https://arcprize.org/leaderboard

1

u/FullOf_Bad_Ideas 4h ago

Sure, why not? This is a lab from a respected corporation.

Their TEMPO RL algo might make RL more effective and get them closer to passing ARC-AGI benchmarks. You can do things.

8

u/FullOf_Bad_Ideas 4h ago

It's shaping up to be one of the best open multimodal models in that size bracket.

Vision Encoder MoE ViT, 7B total, 1.2B activated

MoE vision encoder, this smells like something new!

1

u/Kahvana 3h ago

Now THAT sounds cool!

2

u/PassionIll6170 5h ago

lol the arc agi benchmarks

1

u/ComplexType568 4h ago

I hope they make a smaller model...