4 articles about AI hardware
OpenAI's custom inference chip Jalapeño shows industry-leading speed and efficiency in early benchmarks, signaling a new era of purpose-built AI hardware.
Groq announces it will be among the first adopters of NVIDIA Groq 3 LPX with Vera Rubin NVL72, deploying through Dell Technologies to power its inference cloud with 3,400 tokens/sec on agentic workloads.
Cerebras unveils CS-4, its fourth-generation system with three WSE-3 Turbo processors, delivering up to 30x faster inference than GPUs and 10x more throughput per watt than CS-3.
Anthropic confirmed it is building an in-house chip design team, hiring engineers with silicon experience to co-design hardware and models so Claude runs faster and more efficiently. The move follows similar efforts by OpenAI, Google, and Meta to reduce dependency on Nvidia.