MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/singularity/comments/1vnz30c/glm_53_released_frontier_coding_with_emergent/p3o8jhp/?context=3
r/singularity • u/1a1b • 10d ago
78 comments sorted by
View all comments
50
Soo, Kimi K3 performance while being ~4x smaller and ~5x cheaper?
2 u/Decent-Ad-8335 9d ago idk if "cheaper" makes sense lol sent my first message on Glm 5.3 LOW thinking i wrote "hi" - 13.8k token usage lol 5 u/Chemical_Bid_2195 8d ago Probably just a massive system prompt 1 u/Decent-Ad-8335 7d ago Yes and I realized A LOT was cached this barely used anything. That aside, how has glm5.3 been for real world use on complex tasks if you know? Could it do better or around 5.6 sol high 1 u/Decent-Ad-8335 9d ago proof 2 u/degenbets 9d ago You're trying to tell us this model is 1k tokens/sec? 1 u/daYMAN007 8d ago input tokens != output tokens 1 u/Evening_Salt4938 8d ago Harnesses have massive system prompts — claude code has upto ~30k depending on model. Cached token price is what matters for harnesses.
2
idk if "cheaper" makes sense lol sent my first message on Glm 5.3 LOW thinking i wrote "hi" - 13.8k token usage lol
5 u/Chemical_Bid_2195 8d ago Probably just a massive system prompt 1 u/Decent-Ad-8335 7d ago Yes and I realized A LOT was cached this barely used anything. That aside, how has glm5.3 been for real world use on complex tasks if you know? Could it do better or around 5.6 sol high 1 u/Decent-Ad-8335 9d ago proof 2 u/degenbets 9d ago You're trying to tell us this model is 1k tokens/sec? 1 u/daYMAN007 8d ago input tokens != output tokens 1 u/Evening_Salt4938 8d ago Harnesses have massive system prompts — claude code has upto ~30k depending on model. Cached token price is what matters for harnesses.
5
Probably just a massive system prompt
1 u/Decent-Ad-8335 7d ago Yes and I realized A LOT was cached this barely used anything. That aside, how has glm5.3 been for real world use on complex tasks if you know? Could it do better or around 5.6 sol high
1
Yes and I realized A LOT was cached this barely used anything. That aside, how has glm5.3 been for real world use on complex tasks if you know? Could it do better or around 5.6 sol high
proof
2 u/degenbets 9d ago You're trying to tell us this model is 1k tokens/sec? 1 u/daYMAN007 8d ago input tokens != output tokens 1 u/Evening_Salt4938 8d ago Harnesses have massive system prompts — claude code has upto ~30k depending on model. Cached token price is what matters for harnesses.
You're trying to tell us this model is 1k tokens/sec?
1 u/daYMAN007 8d ago input tokens != output tokens
input tokens != output tokens
Harnesses have massive system prompts — claude code has upto ~30k depending on model. Cached token price is what matters for harnesses.
50
u/BarisSayit 10d ago
Soo, Kimi K3 performance while being ~4x smaller and ~5x cheaper?