Lots of people posting about Z ai "benchmaxxing" to make this model. I think the truth is messy and many faceted: 1. Yes Zai probably cares slightly more about public benchmarks than OpenAI/Ant, helps with marketing 2. Zai is not benchmaxxing to the point where the model is fried (if they did, they'll fix it - they likely use this model internally a lot) 3. Zai likely has a narrower distribution of tasks the model is great at. 5.2 was great at agentic stuff, not the rest 4. GLM has not had vision - single modality is definitely easier 5. Zai is definitely extremely good at what they do, likely well more compute efficient than OpenAI/Ant 6. Time to release for Zai is likely days not months like OpenAI/Ant - this massively flatters them
All together, seems like a perfectly good strategy and congrats on the release - excited for the weights to be out for more broader testing.