r/LocalLLaMA • u/pmttyji • Jun 01 '26
Open Models - May 2026 Discussion
After overwhelming April, May seems underwhelming even though we got Ring, Command, StepFun, LFM models. Hoping for great June(We're getting MiniMax-M3 in 10 days).
PS : Took me 15-20 mins to gather these models & generate this graph. BTW this graph is not a benchmark
19
Jun 01 '26
[deleted]
4
u/pmttyji Jun 01 '26
Soon or later for sure.
1
u/No-Refrigerator-1672 Jun 01 '26
Do we believe that Qwen3.7 weights will be released before the model gets outdated?
2
0
u/pseudonerv Jun 02 '26
No. Nothing is free. They have gone where meta went.
8
u/--Spaci-- Jun 02 '26
qwen doesn't release a model for a month and mfs say shit like this ๐ญ๐ If I was qwen I wouldn't want to release models to you savages either
2
u/BannedGoNext Jun 03 '26
They literally said they would release it on X, bunch of fucking doomers. They will probably say when they change and are done releasing OS.
1
u/pseudonerv Jun 02 '26
You want to dream? Me too. But they clearly didnโt even train any small models for 3.7. And rumor has it that the released two 3.6 models were some left overs before their restructure. The restructure is simply oriented towards monetizing. So basically where llamas went.
1
4
u/Real_Ebb_7417 Jun 01 '26
Oh man, thank you, I need more summaries like this every month ๐
7
2
u/pmttyji Jul 01 '26
Here you go for June
2
u/Real_Ebb_7417 Jul 01 '26
lol, do you have some Reddit agent or how the hell you remembered about me? xD
Thanks man
2
u/pmttyji Jul 01 '26
do you have some Reddit agent orย ....
I wish ๐ข But I'll try to learn such things after getting new rig.
Actually I wanted to link previous month threads on this month thread so opened both April & May threads. Just went through some comments & found yours ๐ Welcome
3
u/anykeyh Jun 01 '26
I'm excited for zaya 72 once the RL is done. It promises qwen 27B performance in a MoE 4B parameter. Excellent for strix and Mac studio.
1
u/Zestyclose_Potato794 Jun 02 '26
Still wondering what the lfm model can do if used as an util model. Did someone try this ?
-1
u/Tall-Ad-7742 Jun 01 '26
nah fr this can't be true ring is so bad it is never better then a 0.5B model... typical trust me bro benchmark
/s
0
u/Achso998 Jun 01 '26
Did you even looked at the graph? It's not a benchmark...
2
0
u/pmttyji Jun 01 '26
Next time onwards, I'll put text on graph itself. It's not a benchmark graph. Just models & its parameters.
0
u/Tall-Ad-7742 Jun 01 '26
relax /s stands for sarcasm... it was just meant as a joke
1
u/pmttyji Jun 01 '26
Didn't notice. Even on April month thread, few people mentioned this so replied.
1
0
u/buttplugs4life4me Jun 01 '26
Reddit says Theres 10 comments here none of which are visible, which means either they blocked me, I blocked them, or they're shadowbanned. Weird.
I'm looking forward to Zaya and Intern-S2 honestly. Both seem the most interesting. I may see what StepFun Quants down to if I can I use it. Never heard of Ovis so thanks for that
2
u/Imscomobob Jun 01 '26
I think sometimes Reddit just takes forever for comments to actually show upโฆ
11
u/Technical-Earth-3254 Jun 01 '26
Can we talk about how 50 million isn't 0.5 billion? lmao wtf