Zhipu AI's GLM Gains Traction With 2M Token Context and Open Commercial License
Zhipu AI, a Beijing-based research firm, has released a new iteration of its General Language Model (GLM) that has sparked widespread discussion on Hacker News and developer communities in 2026. The model features a 2-million-token context window using latent attention compression, which reduces memory complexity from quadratic to roughly linear. GLM supports hybrid reasoning, native tool use, and agentic workflows, while remaining runnable on consumer-grade hardware through quantization. The release includes both a 9-billion-parameter dense model and a 47-billion-parameter Mixture-of-Experts variant, the latter activating only 10 billion parameters per token for faster inference. A fully open, commercially permissive license has made the release particularly appealing to developers seeking alternatives to proprietary AI APIs.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in