MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vny9zs/glm_53_released/p3lront/?context=3
r/LocalLLaMA • u/jmorant555 • 18d ago
Official Announcement
https://z.ai/blog/glm-5.3
361 comments sorted by
View all comments
35
interesting pattern of model scoring absolute dogshit when a new benchmark drop and suddenly being frontier in the next update (TerminalBench 3.0)
14 u/ManikSahdev 18d ago The benchmark also act as a direct training guide and reference. If you were able to create a benchmark where all models do shit, you have essentially given these research labs a target to pursue.
14
The benchmark also act as a direct training guide and reference.
If you were able to create a benchmark where all models do shit, you have essentially given these research labs a target to pursue.
35
u/Educational-Fruit854 18d ago
interesting pattern of model scoring absolute dogshit when a new benchmark drop and suddenly being frontier in the next update (TerminalBench 3.0)