#wafer-scale

1 articles about wafer-scale

Cerebras Introduces CS-4 — Up to 30x Faster Inference, 1,000+ Tokens/sec for 10T-Parameter Models
AI Infrastructure Aug 18, 2026

Cerebras Introduces CS-4 — Up to 30x Faster Inference, 1,000+ Tokens/sec for 10T-Parameter Models

Cerebras unveils CS-4, its fourth-generation system with three WSE-3 Turbo processors, delivering up to 30x faster inference than GPUs and 10x more throughput per watt than CS-3.