BTC $85,403.22 -1.03%
ETH $2,694.36 -1.05%
BNB $780.99 -1.90%
XRP $1.49 -1.36%
SOL $119.77 -0.92%
TRX $0.3355 -0.02%
DOGE $0.0943 -2.15%
ADA $0.2658 -2.20%
BCH $314.61 -1.66%
LINK $13.73 -3.20%
HYPE $93.07 +2.47%
AAVE $181.86 +0.75%
SUI $1.18 -3.82%
XLM $0.2138 -4.81%
ZEC $1,337.19 +0.32%
AAPL $333.02 -0.07%
AMZN $252.10 +0.02%
GOOGL $347.16 +0.92%
MSFT $525.98 +1.96%
META $743.97 +2.23%
NVDA $239.71 +1.73%
TSLA $380.30 +2.13%
SNDK $1,692.99 -2.23%
INTC $116.11 -1.29%
SPCX $171.70 +7.24%
MU $1,058.50 -1.90%
AMD $631.77 -1.03%
BTC $85,403.22 -1.03%
ETH $2,694.36 -1.05%
BNB $780.99 -1.90%
XRP $1.49 -1.36%
SOL $119.77 -0.92%
TRX $0.3355 -0.02%
DOGE $0.0943 -2.15%
ADA $0.2658 -2.20%
BCH $314.61 -1.66%
LINK $13.73 -3.20%
HYPE $93.07 +2.47%
AAVE $181.86 +0.75%
SUI $1.18 -3.82%
XLM $0.2138 -4.81%
ZEC $1,337.19 +0.32%
AAPL $333.02 -0.07%
AMZN $252.10 +0.02%
GOOGL $347.16 +0.92%
MSFT $525.98 +1.96%
META $743.97 +2.23%
NVDA $239.71 +1.73%
TSLA $380.30 +2.13%
SNDK $1,692.99 -2.23%
INTC $116.11 -1.29%
SPCX $171.70 +7.24%
MU $1,058.50 -1.90%
AMD $631.77 -1.03%

cs-4

All
Article
Flash

hot_img Cerebras releases the fourth generation AI inference system CS-4: performance doubled, power consumption doubled, more flexible deployment

Cerebras released its fourth-generation AI inference system CS-4 this week, based on the same 5nm WSE-3 wafer, achieving double the performance by doubling the clock frequency and power consumption. A single CS-4 cabinet accommodates 3 wafers (CS-3 has 2), featuring a modular "backpack" design that simplifies manufacturing and deployment, with a TDP of approximately 125 to 135kW. The CS-4 can provide an inference speed of nearly 4000 tokens/second/user, about twice that of the CS-3, and supports decomposed inference with heterogeneous systems such as AMD and AWS Trainium.Cerebras claims that the CS-4 offers about 2000 times the on-chip memory bandwidth of NVIDIA's Rubin (43PB/s), but the 44GB SRAM capacity remains unchanged, and long-context inference still requires multi-wafer stacking. For example, with the DeepSeek V4 Pro (1.6T parameters), approximately 20 systems are needed for a 1M context window, and about 40 systems are required for 256 concurrent users, corresponding to a CAPEX exceeding 20 million USD. Cerebras is collaborating with clients such as OpenAI and plans to achieve approximately double performance improvements each year, aiming for a 20-fold throughput increase by 2027. The "backpack" cabinet design of the CS-4 will continue into the next-generation "Nexus" platform.
app_icon
ChainCatcher Building the Web3 world with innovations.