BTC $79,475.56 -1.77%
ETH $2,455.63 -1.48%
BNB $716.71 -0.41%
XRP $1.40 -3.66%
SOL $101.40 -3.02%
TRX $0.3295 -0.49%
DOGE $0.0844 -3.83%
ADA $0.2138 -1.96%
BCH $250.86 -1.92%
LINK $11.66 -0.33%
HYPE $84.29 +0.85%
AAVE $131.13 -1.85%
SUI $0.7485 -4.54%
XLM $0.1791 -2.85%
ZEC $977.47 +6.20%
BTC $79,475.56 -1.77%
ETH $2,455.63 -1.48%
BNB $716.71 -0.41%
XRP $1.40 -3.66%
SOL $101.40 -3.02%
TRX $0.3295 -0.49%
DOGE $0.0844 -3.83%
ADA $0.2138 -1.96%
BCH $250.86 -1.92%
LINK $11.66 -0.33%
HYPE $84.29 +0.85%
AAVE $131.13 -1.85%
SUI $0.7485 -4.54%
XLM $0.1791 -2.85%
ZEC $977.47 +6.20%

GLM-5.3-Flash topped the B.AI model call volume rankings, with a cumulative throughput exceeding 2.41 trillion Tokens

2026-09-04 21:41:12

GLM-5.3-Flash has become the most frequently used and popular model on the B.AI platform, with a cumulative token throughput exceeding 2.41 trillion.

As the first native multimodal model in the GLM-5 series, GLM-5.3-Flash features a total of 320 billion parameters and 18 billion active parameters, employing a hybrid architecture that combines sparse and linear attention, supporting 1 million ultra-long contexts, while ensuring rapid response, powerful reasoning, and high cost-effectiveness.

Starting today, developers can still call this model for free through the B.AI platform, covering diverse scenarios such as high-frequency APIs, code writing, complex agents, and ultra-long document processing.

app_icon
ChainCatcher Building the Web3 world with innovations.