AMD Acquires Taalas, an AI Inference Chip Startup
AMD has agreed to acquire Taalas, a Toronto startup that etches AI models directly into silicon, aiming to boost inference speed and efficiency across its chip lineup.
Read more →Every AI News story tagged with both AI Inference and News — the two topics side by side, updated as new articles publish.
9 articles
AMD has agreed to acquire Taalas, a Toronto startup that etches AI models directly into silicon, aiming to boost inference speed and efficiency across its chip lineup.
Read more →OpenAI says its GPT-5.6 Sol model autonomously rewrote the GPU kernels behind its inference stack, cutting serving costs 20% and helping fund an 80% price cut for the lightweight Luna tier.
Read more →AMD and Cerebras unveiled a joint AI inference system that splits chip work between AMD's Helios servers and Cerebras' wafer-scale engine, claiming up to 5x higher efficiency per watt.
Read more →
AMD unveiled its Instinct MI400 AI chips and Helios rack systems on July 23, claiming 34x faster inference, with OpenAI, Meta and Anthropic on hand.
Read more →Chipmaker Etched is negotiating a funding round that would value it at $20 billion, quadrupling its December price, as investors chase alternatives to Nvidia for AI inference, the Wall Street Journal reported.
Read more →SambaNova closed the first tranche of a $1 billion Series F round led by General Atlantic, and said JPMorganChase will deploy its chips for on-premises AI inference.
Read more →Qualcomm on June 24 unveiled the Dragonfly data center chip family — a 250-core Arm-based CPU and a new inference accelerator — with Meta signed on to deploy the hardware in its server fleet from 2028.
Read more →Qualcomm will acquire Modular — creator of the Mojo language and MAX inference engine — in a $3.9 billion all-stock deal aimed at letting developers deploy AI on any hardware without rewrites.
Read more →OpenAI and Broadcom revealed Jalapeño on June 24 — a purpose-built inference chip designed in nine months with AI assistance, targeting deployment by end of 2026.
Read more →