AISemiAnalysis1h ago

Vera Rubin NVL72 inference tests show up

Vera Rubin NVL72 inference tests show up to 7x better token throughput per MW vs. Blackwell on a 1.6T DeepSeek model, above Huang's 3x claim for 1T-3T LLMs

Vera Rubin NVL72 inference tests show up

TL;DRNew AMD chip shows significantly better AI inference efficiency than Nvidia's latest processor on large models.

Why it matters: If validated, this challenges Nvidia's dominance in AI infrastructure and could shift enterprise hardware spending decisions.

Jensen Sandbagging Performance Again, 2x more Annual Profit Per GigaWatt, The More you Buy, The More you Earn, AgentX, InferenceX, Extreme Co-Design

Read full article

Source: SemiAnalysis · Opens in new tab