BTC $79,696.25 +0.57%
ETH $2,457.84 +0.35%
BNB $769.05 +7.63%
XRP $1.41 +1.19%
SOL $102.82 +1.56%
TRX $0.3332 +1.13%
DOGE $0.0874 +3.49%
ADA $0.2169 +1.94%
BCH $250.17 +0.00%
LINK $11.89 +2.60%
HYPE $85.18 +0.91%
AAVE $130.66 -0.03%
SUI $0.7971 +6.74%
XLM $0.1836 +2.72%
ZEC $1,017.41 +3.44%
BTC $79,696.25 +0.57%
ETH $2,457.84 +0.35%
BNB $769.05 +7.63%
XRP $1.41 +1.19%
SOL $102.82 +1.56%
TRX $0.3332 +1.13%
DOGE $0.0874 +3.49%
ADA $0.2169 +1.94%
BCH $250.17 +0.00%
LINK $11.89 +2.60%
HYPE $85.18 +0.91%
AAVE $130.66 -0.03%
SUI $0.7971 +6.74%
XLM $0.1836 +2.72%
ZEC $1,017.41 +3.44%

GLM-5.3-Flash topped the B.AI model call volume rankings, with a cumulative throughput exceeding 2.41 trillion Tokens

2026-09-04 21:41:12

GLM-5.3-Flash has become the most frequently used and popular model on the B.AI platform, with a cumulative token throughput exceeding 2.41 trillion.

As the first native multimodal model in the GLM-5 series, GLM-5.3-Flash features a total of 320 billion parameters and 18 billion active parameters, employing a hybrid architecture that combines sparse and linear attention, supporting 1 million ultra-long contexts, while ensuring rapid response, powerful reasoning, and high cost-effectiveness.

Starting today, developers can still call this model for free through the B.AI platform, covering diverse scenarios such as high-frequency APIs, code writing, complex agents, and ultra-long document processing.

app_icon
ChainCatcher Building the Web3 world with innovations.