BTC $84,578.70 -2.11%
ETH $2,681.33 -2.42%
BNB $767.60 -1.08%
XRP $1.48 -3.71%
SOL $119.27 -2.15%
TRX $0.3358 +0.45%
DOGE $0.0927 -4.41%
ADA $0.2441 -5.05%
BCH $311.04 -1.60%
LINK $13.93 -3.16%
HYPE $88.00 -2.49%
AAVE $180.64 -2.40%
SUI $1.18 +0.30%
XLM $0.2140 -5.27%
ZEC $1,319.09 -5.04%
AAPL $333.52 +0.81%
AMZN $252.00 +0.90%
GOOGL $343.52 +0.71%
MSFT $517.62 +0.32%
META $728.06 -0.30%
NVDA $234.51 +0.35%
TSLA $370.68 +3.85%
SNDK $1,716.38 -4.75%
INTC $118.10 -3.57%
SPCX $159.03 +6.40%
MU $1,070.43 -3.89%
AMD $632.66 +1.21%
BTC $84,578.70 -2.11%
ETH $2,681.33 -2.42%
BNB $767.60 -1.08%
XRP $1.48 -3.71%
SOL $119.27 -2.15%
TRX $0.3358 +0.45%
DOGE $0.0927 -4.41%
ADA $0.2441 -5.05%
BCH $311.04 -1.60%
LINK $13.93 -3.16%
HYPE $88.00 -2.49%
AAVE $180.64 -2.40%
SUI $1.18 +0.30%
XLM $0.2140 -5.27%
ZEC $1,319.09 -5.04%
AAPL $333.52 +0.81%
AMZN $252.00 +0.90%
GOOGL $343.52 +0.71%
MSFT $517.62 +0.32%
META $728.06 -0.30%
NVDA $234.51 +0.35%
TSLA $370.68 +3.85%
SNDK $1,716.38 -4.75%
INTC $118.10 -3.57%
SPCX $159.03 +6.40%
MU $1,070.43 -3.89%
AMD $632.66 +1.21%

cs-4

All
Article
Flash

hot_img Cerebras releases the fourth generation AI inference system CS-4: performance doubled, power consumption doubled, more flexible deployment

Cerebras released its fourth-generation AI inference system CS-4 this week, based on the same 5nm WSE-3 wafer, achieving double the performance by doubling the clock frequency and power consumption. A single CS-4 cabinet accommodates 3 wafers (CS-3 has 2), featuring a modular "backpack" design that simplifies manufacturing and deployment, with a TDP of approximately 125 to 135kW. The CS-4 can provide an inference speed of nearly 4000 tokens/second/user, about twice that of the CS-3, and supports decomposed inference with heterogeneous systems such as AMD and AWS Trainium.Cerebras claims that the CS-4 offers about 2000 times the on-chip memory bandwidth of NVIDIA's Rubin (43PB/s), but the 44GB SRAM capacity remains unchanged, and long-context inference still requires multi-wafer stacking. For example, with the DeepSeek V4 Pro (1.6T parameters), approximately 20 systems are needed for a 1M context window, and about 40 systems are required for 256 concurrent users, corresponding to a CAPEX exceeding 20 million USD. Cerebras is collaborating with clients such as OpenAI and plans to achieve approximately double performance improvements each year, aiming for a 20-fold throughput increase by 2027. The "backpack" cabinet design of the CS-4 will continue into the next-generation "Nexus" platform.
app_icon
ChainCatcher Building the Web3 world with innovations.