What Is Disaggregated Inference — and How Does It Speed Up AI?
Disaggregated inference splits an AI model's two work stages onto separate chips built for each job — a technique now expanding beyond Nvidia to AMD and Cerebras hardware.
Read more →This project is suspended: no new articles or editions will be published. The archive stays available.
Every AI News story tagged with both AI Chips and AMD — the two topics side by side, updated as new articles publish.
4 articles
Disaggregated inference splits an AI model's two work stages onto separate chips built for each job — a technique now expanding beyond Nvidia to AMD and Cerebras hardware.
Read more →AMD and Cerebras unveiled a joint AI inference system that splits chip work between AMD's Helios servers and Cerebras' wafer-scale engine, claiming up to 5x higher efficiency per watt.
Read more →
AMD Instinct is AMD's line of data-center GPUs built to train and run AI models — the main hardware challenger to Nvidia's dominance in AI computing.
Read more →
AMD unveiled its Instinct MI400 AI chips and Helios rack systems on July 23, claiming 34x faster inference, with OpenAI, Meta and Anthropic on hand.
Read more →