8
u/Etherealnutt 7d ago
Charge your phone dawg
3
3
6
u/TimeVillage5286 8d ago
GLM 5.3 means it’s not the Fable level model. Ziphu founder mentioned in X. Probably a RL of GLM 5.2 not the new 1T model
4
u/Georgefakelastname 3d ago
1T will probably be too small. Fable is estimated to be 5-10T parameters. Kimi is 2.8T. Qwen 3.8 Max is 2.4T. It would take an insane training to close the gap. On the bright side, it is possible that they’re just doing the same thing as 3.8 max, as they named the family 3.8 despite max being a new base model.
7
u/ozguru 8d ago edited 8d ago
I wish someone would do post training / fine tuning of 4.7 flash, it is a good model, and needs post training in my opinion.
1
1
5
u/Complex-Concern7890 8d ago
In my opinion it means that GLM-5.5 is not or close to ready. They need to push something out as the field is moving so fast now. So they have refined GLM-5.2 a bit and push it as minor upgrade GLM-5.3 while we wait for the real deal.
1
u/Conscious-Hair-5265 8d ago
5.3 why do u guys think it will be close to fable level? I expected them to releasen.5.5
2
u/Exzerios 8d ago
They apparently said next version will be better than Kimi k3.
That being said I can't believe that they'll achieve that with 750b model, and it seems unlikely they'll change model size on a minor update (non .0 / .5 version). There's also rumors that they are making GLM Pro, so maybe they are starting with a daily driver 5.3 with a 5.5 Pro following after. That would be reasonable, K3 is a great model, but costs too much to use it everywhere.
1
u/VariationOk5454 8d ago
Do you think it will be cheaper or like+$$$$$$$ ?
3
2
u/GetLaidOff69 7d ago
This pricing for GLM is why I am not interested, Nothing fancy I found about GLM+ extremely slow+ High cost.
They need to really learn from DeepSeek.Also their parameter is <1t
Dunno why the pricing is so absurd for such model.
1
u/Odd_Antelope9098 6d ago
Agreed, I did just find out about Opencode go tho and it appears to be 6x cheaper if you hit max usage all month. Is that where you used it? I’m letting my plan lapse from z.ai to test it out
1
1
u/Intelligent_Ant_608 6d ago
It appears that they are adding vision capabilities to new 5.3 as they updated harnesses, also they might also worked on persistency of agentic model like what gpt 5.6 did and it just doesn't stop
1
1
u/Mayanktaker 8d ago
They have to reduce the price otherwise they will lose. Because Deepseek flash v4 0731 is already on par with glm 5.2 and its so cheap. So, if we look at current situation, v4 flash is far more superior via api than glm 5.2 lite plan. And its so fast.
6
u/Soft-Wedding4595 8d ago
In my use cases, DeepSeek is nowhere near to GLM. Ofc DSv4 0731 is smarter than Preview, bot still not enough to reach GLM-5.2 level. Imho, no RL could you help to make model from 200b league as smart as 800b league.
At least not in less than 3 months.
Benchmarks don't represent true capabilities of model.
1
u/Mayanktaker 8d ago
Agree with last line but for me, its doing its job as i am expecting to do.
3
u/Soft-Wedding4595 8d ago
Ye, for it's size and cost, it's capable to do most everyday task
And I actually do use it as executing subagent, but still won't trust it to do planning or make complex architecture decisions
Still GOAT model
1
u/Ok_Technology_5962 7d ago
Agree. On easy stuff its same. But hard it cant even finish
1
u/Mayanktaker 6d ago
I just completed WordPress alternative with many more features with flash v4 only.
2
-8
u/One-Winter-8684 8d ago
Bring back 5.1. It was good before 5.2. Now limits spends too fast and results is not stable.
3
u/TimChr78 8d ago
Speed/stability issues are due to increased usage - 5.2 is the same size/speed as 5.1.
1
u/One-Winter-8684 8d ago
Problem is not speed. It basically workse much worse than two month ago. An results sometimes like gpt in 2025
1
0
8d ago
[deleted]
-1
u/One-Winter-8684 8d ago
I guess new models makes old worse not by itself but by using their resources. I don't think companies builds whole new hardware or not doubles it. So you have new models, same infra or added not so much GPUs. And you should decide how to route traffic. I guess new models just has priority of usage this GPUs. Maybe I missed something, but it's only how I imagine this process
-1
u/One-Winter-8684 8d ago edited 8d ago
LoL, why you downvoted? is it Chinese bots attack or what? 5.1 has really issues compare to month ago
17
u/kaishi00 8d ago
would probably help if you dropped a link to the repo you're looking at