BTC $84,710.62 +1.49%
ETH $2,714.18 +2.54%
BNB $776.79 +1.15%
XRP $1.56 +6.48%
SOL $119.33 +5.27%
TRX $0.3370 -0.75%
DOGE $0.0969 +4.75%
ADA $0.2527 +7.04%
BCH $337.63 +2.02%
LINK $13.90 +14.07%
HYPE $94.22 +3.44%
AAVE $145.87 +6.78%
SUI $1.07 +13.46%
XLM $0.2209 +11.23%
ZEC $1,599.60 +9.23%
AAPL $335.88 -0.33%
AMZN $250.77 +1.21%
GOOGL $343.47 +1.80%
MSFT $497.29 -0.32%
META $778.17 +6.67%
NVDA $225.66 +0.98%
TSLA $380.98 +1.00%
SNDK $1,786.94 +0.51%
INTC $128.93 +8.05%
SPCX $149.33 +0.80%
MU $1,092.50 +3.83%
AMD $641.22 +6.29%
BTC $84,710.62 +1.49%
ETH $2,714.18 +2.54%
BNB $776.79 +1.15%
XRP $1.56 +6.48%
SOL $119.33 +5.27%
TRX $0.3370 -0.75%
DOGE $0.0969 +4.75%
ADA $0.2527 +7.04%
BCH $337.63 +2.02%
LINK $13.90 +14.07%
HYPE $94.22 +3.44%
AAVE $145.87 +6.78%
SUI $1.07 +13.46%
XLM $0.2209 +11.23%
ZEC $1,599.60 +9.23%
AAPL $335.88 -0.33%
AMZN $250.77 +1.21%
GOOGL $343.47 +1.80%
MSFT $497.29 -0.32%
META $778.17 +6.67%
NVDA $225.66 +0.98%
TSLA $380.98 +1.00%
SNDK $1,786.94 +0.51%
INTC $128.93 +8.05%
SPCX $149.33 +0.80%
MU $1,092.50 +3.83%
AMD $641.22 +6.29%

cere

All
Article
Flash

hot_img Cerebras releases the fourth generation AI inference system CS-4: performance doubled, power consumption doubled, more flexible deployment

Cerebras released its fourth-generation AI inference system CS-4 this week, based on the same 5nm WSE-3 wafer, achieving double the performance by doubling the clock frequency and power consumption. A single CS-4 cabinet accommodates 3 wafers (CS-3 has 2), featuring a modular "backpack" design that simplifies manufacturing and deployment, with a TDP of approximately 125 to 135kW. The CS-4 can provide an inference speed of nearly 4000 tokens/second/user, about twice that of the CS-3, and supports decomposed inference with heterogeneous systems such as AMD and AWS Trainium.Cerebras claims that the CS-4 offers about 2000 times the on-chip memory bandwidth of NVIDIA's Rubin (43PB/s), but the 44GB SRAM capacity remains unchanged, and long-context inference still requires multi-wafer stacking. For example, with the DeepSeek V4 Pro (1.6T parameters), approximately 20 systems are needed for a 1M context window, and about 40 systems are required for 256 concurrent users, corresponding to a CAPEX exceeding 20 million USD. Cerebras is collaborating with clients such as OpenAI and plans to achieve approximately double performance improvements each year, aiming for a 20-fold throughput increase by 2027. The "backpack" cabinet design of the CS-4 will continue into the next-generation "Nexus" platform.
app_icon
ChainCatcher Building the Web3 world with innovations.