Zhipu's GLM-5.2 tops the open-weights leaderboard at 744B parameters — and you can just download it
Z. ai (Zhipu) has released GLM-5.2, a 744-billion-parameter Mixture-of-Experts model with open weights that now leads the Artificial Analysis Intelligence Index among open-weight models, with clear gains over GLM-5.1 on coding and agentic tasks. Within days it was wired into community agent stacks like Nous Research's Hermes Agent. It lands in a crowded open-weights wave that also includes Moonshot's Kimi K2.7 Code and MiniMax's natively multimodal M3.
Why it matters: The open-weights frontier isn't just chasing the proprietary labs anymore — on this index it's leading, and a freely downloadable 744B model changes the build-versus-rent math for anyone serious about controlling their stack. For teams, the payoff is concrete: self-hosting kills per-token fees, keeps sensitive data in-house, and — crucially for agentic workloads that fire thousands of calls per task — removes the metered API as the thing that makes ambitious agents economically painful to run. That so many of these releases (GLM, Kimi, MiniMax) are coming from Chinese labs is the trend beneath the trend: open weights have become the competitive lever for challengers who can't out-distribute the incumbents but can out-give them. The days-to-integration into Hermes shows the other half of why this matters — an open ecosystem compounds faster than any single vendor's roadmap, because the community productizes a release before the lab's own tooling catches up. The strategic caveat for adopters is that self-hosting a 744B model is not free — the fees move from the API bill to your infrastructure and ops team — so the real question is whether you have the muscle to run it, not whether you're allowed to.