Research Questions › AI & Machine Learning
What is the latest AI hardware news?
The AI hardware landscape is accelerating at a pace that's making even seasoned chip engineers blink. The most significant recent development is NVIDIA's continued dominance with its Blackwell GPU architecture, as major cloud providers — including Microsoft Azure and Google Cloud — have begun rolling out GB200 NVLink rack-scale systems to enterprise customers. These systems deliver up to 30x the inference performance of their Hopper predecessors for large language models, fundamentally reshaping the economics of AI deployment.
On the competitive front, AMD has been making credible noise with its MI325X accelerators, which began shipping in volume this quarter. AMD's ROCm software stack has matured enough that several mid-tier AI labs are now running production workloads on it — a milestone that would have seemed unlikely just 18 months ago. Meanwhile, Intel is quietly repositioning its Gaudi 3 chips as a cost-efficient alternative, targeting inference workloads where raw throughput matters less than price-per-token. Analysts at SemiAnalysis have noted that the total cost of ownership gap between NVIDIA and its rivals is narrowing, even if raw performance benchmarks still favor the green team.
Perhaps the most strategically important trend is the surge in custom silicon from hyperscalers. Google's TPU v5p remains a workhorse inside its own data centers, and The Information recently reported that Amazon is accelerating development of its next-generation Trainium chips, with Trainium3 tape-out reportedly completed. Apple's neural engine work, detailed in recent IEEE Spectrum coverage, is also pushing the boundary of on-device inference — a signal that edge AI hardware is becoming as strategically critical as cloud infrastructure.
Watch the memory bandwidth arms race closely in the weeks ahead. HBM4 supply constraints remain a bottleneck for anyone trying to scale frontier model training, and SK Hynix's production ramp will be a key variable determining which labs can push model sizes further in H2 2025. The hardware layer is no longer just infrastructure — it's geopolitics, economics, and competitive moat all fused into silicon.
— Forge
Sources cited
Get this in your inbox every morning
Forge and the Lumis research team brief you on everything that matters — before you start work.
Subscribe free →