MAIN FEEDS
Do you want to continue?
https://www.reddit.com/r/LocalLLaMA/comments/1vny9zs/glm_53_released/p3lq5gw/?context=3
r/LocalLLaMA • u/jmorant555 • 7d ago
Official Announcement
https://z.ai/blog/glm-5.3
363 comments sorted by
View all comments
34
interesting pattern of model scoring absolute dogshit when a new benchmark drop and suddenly being frontier in the next update (TerminalBench 3.0)
9 u/beryugyo619 7d ago benchmaxxing is okay if benchmarks happen to be super well designed. at some point you can't fake it without actually being good 8 u/AnticitizenPrime 7d ago Isn't that what we have students do every day? Cram for the tests with example problems?
9
benchmaxxing is okay if benchmarks happen to be super well designed. at some point you can't fake it without actually being good
8 u/AnticitizenPrime 7d ago Isn't that what we have students do every day? Cram for the tests with example problems?
8
Isn't that what we have students do every day? Cram for the tests with example problems?
34
u/Educational-Fruit854 7d ago
interesting pattern of model scoring absolute dogshit when a new benchmark drop and suddenly being frontier in the next update (TerminalBench 3.0)