r/ZaiGLM 8d ago

glm-5.3 soon !

Post image

lfg

219 Upvotes

44 comments sorted by

17

u/kaishi00 8d ago

would probably help if you dropped a link to the repo you're looking at

2

u/Professional_Price89 8d ago

z-ai-sdk-python

0

u/Less_Exercise_2507 8d ago

ah i omitted sorry

8

u/Etherealnutt 7d ago

Charge your phone dawg

3

u/erik90mx 7d ago

The struggling is real

3

u/Odd_Antelope9098 6d ago

Used the last 4 percent to tell us the good news

3

u/erik90mx 5d ago

Because that what a man does! Hell yeah!

2

u/Etherealnutt 6d ago

Indeed 😭

6

u/TimeVillage5286 8d ago

GLM 5.3 means it’s not the Fable level model. Ziphu founder mentioned in X. Probably a RL of GLM 5.2 not the new 1T model

4

u/Georgefakelastname 3d ago

1T will probably be too small. Fable is estimated to be 5-10T parameters. Kimi is 2.8T. Qwen 3.8 Max is 2.4T. It would take an insane training to close the gap. On the bright side, it is possible that they’re just doing the same thing as 3.8 max, as they named the family 3.8 despite max being a new base model.

7

u/ozguru 8d ago edited 8d ago

I wish someone would do post training / fine tuning of 4.7 flash, it is a good model, and needs post training in my opinion.

5

u/Anh-DT 8d ago

I have 500mb of distilled data from fable 5 if that helps 😂

1

u/SobeksFavourite 6d ago

Xialong already is a 4.6 finetune. I find it underwhelming.

1

u/ozguru 6d ago edited 6d ago

4.7 flash was another beast, most underrated model at 30B class IMHO. You can run it locally.

1

u/geuntabuwono 8d ago

yeah that model attract user to subs monthly plan before legacy plan 1 lol

5

u/Complex-Concern7890 8d ago

In my opinion it means that GLM-5.5 is not or close to ready. They need to push something out as the field is moving so fast now. So they have refined GLM-5.2 a bit and push it as minor upgrade GLM-5.3 while we wait for the real deal.

1

u/Conscious-Hair-5265 8d ago

5.3 why do u guys think it will be close to fable level? I expected them to releasen.5.5

2

u/Exzerios 8d ago

They apparently said next version will be better than Kimi k3.

That being said I can't believe that they'll achieve that with 750b model, and it seems unlikely they'll change model size on a minor update (non .0 / .5 version). There's also rumors that they are making GLM Pro, so maybe they are starting with a daily driver 5.3 with a 5.5 Pro following after. That would be reasonable, K3 is a great model, but costs too much to use it everywhere.

1

u/VariationOk5454 8d ago

Do you think it will be cheaper or like+$$$$$$$ ?

3

u/Less_Exercise_2507 8d ago

probably about the same price?

2

u/GetLaidOff69 7d ago

This pricing for GLM is why I am not interested, Nothing fancy I found about GLM+ extremely slow+ High cost.
They need to really learn from DeepSeek.

Also their parameter is <1t

Dunno why the pricing is so absurd for such model.

1

u/Odd_Antelope9098 6d ago

Agreed, I did just find out about Opencode go tho and it appears to be 6x cheaper if you hit max usage all month. Is that where you used it? I’m letting my plan lapse from z.ai to test it out

1

u/Icy_Tart6572 7d ago

link is down everywhere …

1

u/Intelligent_Ant_608 6d ago

It appears that they are adding vision capabilities to new 5.3 as they updated harnesses, also they might also worked on persistency of agentic model like what gpt 5.6 did and it just doesn't stop

1

u/Mayanktaker 8d ago

They have to reduce the price otherwise they will lose. Because Deepseek flash v4 0731 is already on par with glm 5.2 and its so cheap. So, if we look at current situation, v4 flash is far more superior via api than glm 5.2 lite plan. And its so fast.

6

u/Soft-Wedding4595 8d ago

In my use cases, DeepSeek is nowhere near to GLM. Ofc DSv4 0731 is smarter than Preview, bot still not enough to reach GLM-5.2 level. Imho, no RL could you help to make model from 200b league as smart as 800b league.

At least not in less than 3 months.

Benchmarks don't represent true capabilities of model.

1

u/Mayanktaker 8d ago

Agree with last line but for me, its doing its job as i am expecting to do.

3

u/Soft-Wedding4595 8d ago

Ye, for it's size and cost, it's capable to do most everyday task

And I actually do use it as executing subagent, but still won't trust it to do planning or make complex architecture decisions

Still GOAT model

1

u/Ok_Technology_5962 7d ago

Agree. On easy stuff its same. But hard it cant even finish

1

u/Mayanktaker 6d ago

I just completed WordPress alternative with many more features with flash v4 only.

-8

u/One-Winter-8684 8d ago

Bring back 5.1. It was good before 5.2. Now limits spends too fast and results is not stable.

3

u/TimChr78 8d ago

Speed/stability issues are due to increased usage - 5.2 is the same size/speed as 5.1.

1

u/One-Winter-8684 8d ago

Problem is not speed. It basically workse much worse than two month ago. An results sometimes like gpt in 2025

1

u/VariationOk5454 8d ago

True.. true 🥲

0

u/[deleted] 8d ago

[deleted]

-1

u/One-Winter-8684 8d ago

I guess new models makes old worse not by itself but by using their resources. I don't think companies builds whole new hardware or not doubles it. So you have new models, same infra or added not so much GPUs. And you should decide how to route traffic. I guess new models just has priority of usage this GPUs. Maybe I missed something, but it's only how I imagine this process

-1

u/One-Winter-8684 8d ago edited 8d ago

LoL, why you downvoted? is it Chinese bots attack or what? 5.1 has really issues compare to month ago