When it comes to inference, compute is key, but memory bandwidth is king. With 128 accelerators OpenAI's (& Broadcom) Jalapeño offers ~2 PB/s of HMB4 B/W — more than either Nvidia's Vera Rubin or AMD's Helios
My 1,000+ word analysis only @TheRegister
www.theregister.com/systems/2026...
theregister.com
OpenAI's upcoming Jalapeño chip looks like it'll be an inference beast
128 chips, 1.7 exaFLOPS, and 27 TB of HBM give Altman and crew a leg up over Blackwell, and maybe even Rubin