Zhipu AI explores custom silicon as GLM-5.2 demand surges

Chinese AI lab Zhipu AI is in early discussions with domestic chip design houses about building a bespoke processor optimized for its GLM model family, according to a report by The Information. The move comes as daily token usage for GLM-5.2 surged 27-fold during its first week of release, straining compute capacity already squeezed by U.S. export controls on advanced semiconductors. The Beijing-based company, which trades on the Hong Kong Stock Exchange as Z.ai, has made preliminary inquiries

2 min

Chinese AI lab Zhipu AI is in early discussions with domestic chip design houses about building a bespoke processor optimized for its GLM model family, according to a report by The Information. The move comes as daily token usage for GLM-5.2 surged 27-fold during its first week of release, straining compute capacity already squeezed by U.S. export controls on advanced semiconductors.

The Beijing-based company, which trades on the Hong Kong Stock Exchange as Z.ai, has made preliminary inquiries with several Chinese ASIC design firms but has not yet selected a partner. The conversations remain exploratory, and any resulting chip would take more than two years to design, test, and bring to production.

The catalyst is straightforward. GLM-5.2, released in June 2026, became the fastest-growing model on Vercel's model aggregator platform, with daily token usage jumping as much as 27 times during launch week. At the same time, U.S. export restrictions have made it increasingly difficult for Chinese AI labs to acquire Nvidia's most capable GPUs, turning compute availability into a structural constraint rather than a cost issue.

ASICs, or application-specific integrated circuits, are processors engineered for particular model architectures rather than the general-purpose computation that GPUs provide. They typically deliver better energy efficiency and lower per-token inference costs once a model's architecture stabilizes, making them economically attractive for labs running high-volume inference workloads.

Zhipu would be following a well-established path. Google, OpenAI, ByteDance, and Alibaba have all developed proprietary chips to reduce dependence on outside GPU suppliers. Hours before The Information's report, Reuters reported that DeepSeek is also pursuing custom silicon to reduce its reliance on both Huawei and Nvidia.

The broader Chinese ASIC ecosystem has expanded since initial U.S. export restrictions took effect. Cambricon Technologies and Biren Technology are among the domestic firms active in the AI chip space, though neither has been named as a prospective Zhipu partner.

For Nvidia, each Chinese lab that transitions inference to domestic alternatives represents a slice of its China data-center revenue that becomes structurally harder to recover, regardless of how export-control policy evolves. The immediate question for Zhipu is execution: chip design, foundry access, and software adaptation must happen simultaneously, and the lab will need to build or expand a semiconductor team to see the project through.

Sources

Zhipu AI explores custom ASIC chip as GLM-5.2 usage surges 27x - Yahoo Finance / Investing.com

China's AI Lab Zhipu Weighs Custom Chip As Demand for its GLM Model Soars - The Information

Written by

More to read

  • Amazon Data Center Could Be Powered by One of the Nation's Most Polluting Power Plants

    Amazon is investing in a new natural-gas power plant in Pecos County, Texas, to supply a West Texas data center, and the project holds a permit that would allow it to emit more carbon dioxide than any coal plant in the country, according to The Verge and the New York Times. The plant, tracked as GW Ranch by Cleanview, a service that monitors data center power projects, would deploy 35 natural-gas turbines generating about 7.65 gigawatts. At least initially, the plant would not connect to

    1 min
  • Claude Code Defaults to Auto Mode. The Classifier Catches More Than Humans.

    Claude Code Defaults to Auto Mode. The Classifier Catches More Than Humans. Claude Code will ship with Auto Mode enabled by default starting August 14 for Pro, Max, and Team subscribers, shifting the developer role further from active coding toward reviewing AI-generated output. Only Enterprise customers will need to opt in. Auto Mode lets the agent execute steps without waiting for manual approval at each one. A classifier intercepts actions the model judges dangerous or irreversible and paus

    1 min