BTC $77,568.87 -1.42%
ETH $2,414.84 -2.40%
BNB $687.89 -0.81%
XRP $1.35 -2.77%
SOL $100.27 -3.41%
TRX $0.3217 -3.19%
DOGE $0.0814 -2.27%
ADA $0.1980 -1.40%
BCH $249.11 +0.50%
LINK $11.22 -2.10%
HYPE $82.96 -1.61%
AAVE $129.24 +2.65%
SUI $0.7235 -1.49%
XLM $0.1758 -1.22%
ZEC $841.38 -1.76%
BTC $77,568.87 -1.42%
ETH $2,414.84 -2.40%
BNB $687.89 -0.81%
XRP $1.35 -2.77%
SOL $100.27 -3.41%
TRX $0.3217 -3.19%
DOGE $0.0814 -2.27%
ADA $0.1980 -1.40%
BCH $249.11 +0.50%
LINK $11.22 -2.10%
HYPE $82.96 -1.61%
AAVE $129.24 +2.65%
SUI $0.7235 -1.49%
XLM $0.1758 -1.22%
ZEC $841.38 -1.76%

zhipu

All
Article
Flash

first_img The core ARR of China's open-source large models is approximately 6-8 billion USD

Robonomics author FD published the China Open-Source LLM Tracker on September 1, 2026, stating that the total ARR of China's open-source large models is approximately $10-15 billion, with core LLM revenue around $6-8 billion after excluding ByteDance / Seedance. Growth has not relied on price wars; after DeepSeek raised prices by about 3-12 times, usage still increased, and the average price of the Zhipu API rose by 101% while token usage exceeded 40 times within the year. The best estimate for overseas revenue is approximately $2 billion.Zhipu MaaS ARR increased from about $250 million in March to an annualized monthly rate of about $1.6 billion in August, with a weekly annualized rate of $2 billion. August revenue has already surpassed its API revenue for the first half of the year; the API gross margin is 24.6%, and inference costs per token decreased by 80%. MiniMax ARR rose from about $100 million in December 2025 to over $800 million in annualized weekly revenue by August 2026, with B2B accounting for about 80%. Kimi ARR increased from about $100 million in March to about $300 million by mid-June. DeepSeek's revenue from January to July was approximately $70 million, with the latest ARR estimated at about $1.2 billion.ByteDance's annualized AI revenue is approximately $4 billion, of which Seedance accounts for about $2 billion.

Zhipu has launched and open-sourced the "Niu Lai" model GLM-5.3-Flash

Zhipu officially announced the launch and open-sourcing of GLM-5.3-Flash, which is the first native multimodal model in the GLM-5 series.It is reported that the overall performance of GLM-5.3-Flash exceeds that of GLM-5.2, with programming and Agent evaluations approaching Claude Opus 4.8, but at only one-tenth the price of GLM-5.2. It also features a new foundational model, introducing a mixed architecture of sparse attention and linear attention for the first time in the main GLM series, and is pre-trained with 30T Token multimodal data.Zhipu stated that to gather extensive and professional feedback from a wide range of users, large-scale testing was conducted with the anonymous model Ox-Alpha (referred to as "Niu Lai" in the Chinese community) on OpenCode and OpenRouter before the official release. Ox-Alpha quickly became the most popular model of the week, setting a new high for call volume on both platforms, with all request traffic supported by domestic chip computing power.Previously, the community's DeepSWE small sample test for Ox Alpha once achieved 80%, but that was based on only 10 questions. After expanding the sample, the score fell back to about 63%, and testers also actively corrected the initial claim. This score still belongs to the top tier, but is not as exaggerated as the initial 80%. After the weights are open-sourced, developers can directly deploy using frameworks like vLLM, SGLang, KTransformers, without needing to go through the anonymous model's API.

hot_img Zhipu has acquired AI Infra company Zhongke Jiahe for hundreds of millions, fully addressing the shortcomings in underlying heterogeneous computing power engineering

According to "AI Technology Review," China's leading large model company Zhipu has invested hundreds of millions of yuan to acquire the AI heterogeneous computing power software infrastructure company Zhongke Jiahe. This move aims to completely address Zhipu's shortcomings in the underlying engineering and compiler capabilities of large models, in response to the structural shortage of computing power and high-concurrency inference challenges caused by a surge in user numbers.Zhongke Jiahe's technology originates from the Compiler Laboratory of the Institute of Computing Technology, Chinese Academy of Sciences, founded by Dr. Cui Huimin. Its core team has been deeply involved in the development of compilers for several domestic chips, including Loongson, Sunway, Cambricon, and Huawei Ascend. Zhongke Jiahe's core advantage lies in its virtual instruction set technology, which can unify different brands and models of chip ecosystems through middleware software, assembling scattered domestic chips into a unified ultra-large-scale cluster, thereby significantly improving overall computing power utilization; its SigInfer inference engine is claimed by the official source to reduce the inference latency of large models by up to 74 times.Recently, Zhipu's Coding Agent business has experienced explosive growth. The newly released GLM-5.2 large model saw a 27-fold increase in daily Token call volume during its first week on the aggregation platform, leading to the exposure of systemic engineering bottlenecks in its inference infrastructure under high concurrency and long context scenarios. After being placed on the U.S. Entity List, Zhipu has actively promoted domestic alternatives and has now completed inference adaptation for eight major domestic computing power platforms, including Huawei Ascend, PingTouGe, and Moore Threads. The acquisition of Zhongke Jiahe will not only directly improve Zhipu's unit Token inference cost and output quality but also provide core underlying compiler technology support for its previously rumored self-developed custom AI inference chip plan.

hot_img Zhipu ARR breaks 1 billion USD, achieving 15 times explosive growth in 6 months

According to "Intelligent Emergence," multiple independent sources reported that by July 2026, China's leading large model company, Zhipu, had achieved an annual recurring revenue (ARR) of $1 billion. Insiders revealed that between January and July 2026, Zhipu's ARR surged 15 times in just six months. The growth from $100 million ARR to $1 billion took Zhipu only 5 months, surpassing the 15 months it took the American AI lab Anthropic to reach a similar stage. As of the time of publication, Zhipu has not responded to the aforementioned financial data.The rapid rise in Zhipu's revenue is primarily attributed to its concentrated investment in coding and reasoning capabilities. Financial reports show that in the first quarter of this year, even though the API call price for GLM increased by approximately 83%, its overseas subscription prices approached those of Anthropic's Claude Code, yet its total call volume still grew by about 400% against the trend. In June of this year, Zhipu launched its latest open-source large model, GLM-5.2, which has matched or even surpassed mainstream cutting-edge models on several core metrics.Currently, AI coding and video generation have become the fastest commercialized and most revenue-generating tracks for large models globally. As the demand in the coding market becomes more certain, industry competition is intensifying. Domestically, MiniMax released the M3, which focuses on enhancing coding capabilities, in June, while the Dark Side of the Moon launched the K3 open-source model with 2.8 trillion parameters on July 16; internationally, with OpenAI merging ChatGPT and CodeX, it has also shown a momentum to catch up with Anthropic in the coding field, and the global battle for AI productivity tools continues to escalate.
app_icon
ChainCatcher Building the Web3 world with innovations.