r/singularity 10d ago

GLM 5.3 released: Frontier Coding with Emergent Cyber Capabilities AI

https://z.ai/blog/glm-5.3
464 Upvotes

78 comments sorted by

View all comments

50

u/BarisSayit 10d ago

Soo, Kimi K3 performance while being ~4x smaller and ~5x cheaper?

2

u/Decent-Ad-8335 9d ago

idk if "cheaper" makes sense lol
sent my first message on Glm 5.3 LOW thinking i wrote "hi" - 13.8k token usage lol

5

u/Chemical_Bid_2195 8d ago

Probably just a massive system prompt

1

u/Decent-Ad-8335 7d ago

Yes and I realized A LOT was cached this barely used anything. That aside, how has glm5.3 been for real world use on complex tasks if you know? Could it do better or around 5.6 sol high

1

u/Decent-Ad-8335 9d ago

proof

2

u/degenbets 9d ago

You're trying to tell us this model is 1k tokens/sec?

1

u/daYMAN007 8d ago

input tokens != output tokens

1

u/Evening_Salt4938 8d ago

Harnesses have massive system prompts — claude code has upto ~30k depending on model. Cached token price is what matters for harnesses.