AISemiAnalysis1h ago
Vera Rubin NVL72 inference tests show up
Vera Rubin NVL72 inference tests show up to 7x better token throughput per MW vs. Blackwell on a 1.6T DeepSeek model, above Huang's 3x claim for 1T-3T LLMs

TL;DRNew AMD chip shows significantly better AI inference efficiency than Nvidia's latest processor on large models.
Why it matters: If validated, this challenges Nvidia's dominance in AI infrastructure and could shift enterprise hardware spending decisions.
Jensen Sandbagging Performance Again, 2x more Annual Profit Per GigaWatt, The More you Buy, The More you Earn, AgentX, InferenceX, Extreme Co-Design
Read full articleSource: SemiAnalysis · Opens in new tab