GLM4 articles

GLM

Articles

  • Zhipu's GLM-5.3-Flash Runs Fully on Domestic Chinese Chips, Challenges NVIDIA Dominance

    Zhipu's GLM-5.3-Flash Runs Fully on Domestic Chinese Chips, Challenges NVIDIA Dominance Zhipu AI's GLM-5.3-Flash model, initially released as the mysterious "Niu Lai" (Ox Alpha) model, has been confirmed to run entirely on domestically produced Chinese accelerator chips, marking a significant milestone in China's AI self-sufficiency efforts. The 320B parameter mixture-of-experts model activates only 18B parameters and achieves performance comparable to Claude Opus 4.8 while operating at 1/40th

    1 min
  • Z.ai releases GLM-5.3-Flash, a 320B parameter hybrid sparse-linear attention model with 18B active parameters

    Z.ai releases GLM-5.3-Flash, a 320B parameter hybrid sparse-linear attention model with 18B active parameters Chinese AI startup Z.ai (formerly Zhipu AI) has released GLM-5.3-Flash, the first natively multimodal model in the GLM-5 series. The model was previously known in stealth as "Ox Alpha" and topped OpenRouter's leaderboard before its official release. GLM-5.3-Flash features a hybrid architecture combining sparse and linear attention with Manifold-Constrained Hyper-Connections (mHC), redu

    1 min
  • Zhipu Confirms Ox Alpha Stealth Model Is GLM Derivative Ahead of Open-Weight Release

    Beijing-based AI lab Z.AI, widely known as Zhipu, has confirmed that the mystery stealth model "Ox Alpha" is an upcoming release in its GLM model series. The company confirmed the model's identity in response to inquiries from Bloomberg on Wednesday and stated it will release the open model weights tonight. Ox Alpha first appeared on August 20, 2026, as an uncredited stealth model on the OpenRouter model routing platform. Operating under a free evaluation tier, the model quickly surged to the t

    1 min
  • Zhipu's GLM-5.2 Narrows the Gap With Anthropic's Fable 5 on Coding Benchmarks

    Zhipu AI's open-weight GLM-5.2 model has placed second globally on the Code Arena coding benchmark, trailing only Anthropic's Claude Fable 5 and sitting within one point of Claude Opus 4.8 on the hardest agentic tasks. The result, recorded after the model's June 13 release under an MIT license, signals a fast-closing capability gap between Chinese open-weight systems and the leading U.S. frontier models. Benchmark Performance On Code Arena's front-end coding leaderboard, GLM-5.2 places second

    1 min