hardware AMD EPYC & Cerebras WSE Deliver Ultra-Fast AI Inference in Helios Rack-Scale Systems The Helios system is already running GPT-3 scale models with sub-millisecond latency, and that's a direct shot at Nvidia's inference throne.
Diego Alvarez · 7h ago