Live Feeds
โ— LIVE Updated 1h ago ยท 9 sources tracked

Vera Rubin NVL72 Agentic Inference: 67x better Performance per Dollar

NVIDIA has introduced the Vera Rubin NVL72 system, targeting agentic AI workloads with significant gains in energy and cost efficiency. On September 16, 2026, NVIDIA published MLPerf Inference v6.1 preview results showing the system delivers up to 3.7x performance improvements. CoreWeave has already deployed multi-rack Vera Rubin NVL72 clusters into its cloud environment to support scale-out agentic AI. The system focuses on optimizing tokens per watt for AI factories, with SemiAnalysis reporting a 67x increase in performance per dollar for agentic inference.

๐ŸŽ™๏ธ

Listen to Live Briefing

Real-time synthesized voice briefing ยท Live Feeds Desk

โฑ ~3 min
Speed:
RSS Source map (9)
โšก Key Developments & Real-Time Context
Text size:
  • โœ“ NVIDIA published MLPerf Inference v6.1 submission results for the Vera Rubin NVL72 on September 16, 2026.
  • โœ“ The Vera Rubin NVL72 system delivered up to 3.7x performance according to MLPerf preview results.
  • โœ“ CoreWeave deployed multi-rack NVIDIA Vera Rubin NVL72 clusters on its cloud platform on September 16, 2026.
  • โœ“ The Vera Rubin and DSX platform focus on optimizing tokens per watt for AI factories.
๐Ÿ›ก๏ธ Source Corroboration: 9 independent reporting domains (90% confidence) โฑ Read time: ~2 min

What changed

NVIDIA released MLPerf Inference v6.1 preview results and CoreWeave deployed multi-rack Vera Rubin NVL72 clusters.

Live updates

  1. NVIDIA Vera Rubin NVL72 Launches with High Agentic Inference Efficiency

    NVIDIA has introduced the Vera Rubin NVL72 system, targeting agentic AI workloads with significant gains in energy and cost efficiency. On September 16, 2026, NVIDIA published MLPerf Inference v6.1 preview results showing the system delivers up to 3.7x performance improvements. CoreWeave has already deployed multi-rack Vera Rubin NVL72 clusters into its cloud environment to support scale-out agentic AI. The system focuses on optimizing tokens per watt for AI factories, with SemiAnalysis reporting a 67x increase in performance per dollar for agentic inference.

    Why it matters

    The shift toward agentic AI requires chip architectures that optimize for specific inference patterns rather than just raw training power. NVIDIA is integrating these accelerators into its DSX AI Factory platform to maximize energy efficiency. This transition represents a broader transformation in data center design to avoid obsolescence.

    What is confirmed

    • NVIDIA published MLPerf Inference v6.1 submission results for the Vera Rubin NVL72 on September 16, 2026.
    • The Vera Rubin NVL72 system delivered up to 3.7x performance according to MLPerf preview results.
    • CoreWeave deployed multi-rack NVIDIA Vera Rubin NVL72 clusters on its cloud platform on September 16, 2026.
    • The Vera Rubin and DSX platform focus on optimizing tokens per watt for AI factories.

    Still unconfirmed

    • The Vera Rubin AI accelerator delivers 2X profit per GW compared to Blackwell.

    What to watch next

    • Full MLPerf Inference v6.1 final results
    • Additional cloud provider deployments of Rubin NVL72 clusters
    • Official NVIDIA performance per dollar benchmarks for agentic AI
    Sources used for this update (9)
    1. SemiAnalysis โ€” Vera Rubin NVL72 Agentic Inference: 67x better Performance per Dollar
    2. NVIDIA Blog โ€” AI Infra Summit: NVIDIA Vera Rubin and DSX Platform Advances Showcase Energy Efficiencies of Optimizing Tokens Per Watt for AI Factories
    3. Forbes โ€” The CPU Is Back Thanks To Agents, And What That Means For You
    4. Broadband Breakfast โ€” How Shifts in Computing from AI Are Changing Chip Architectures
    5. Unite.AI โ€” NVIDIA Reports Early Production Results for DSX AI Factory Platform
    6. 36Kr โ€” NVIDIA Leading AI Data Center Major Transformation: Chinese Manufacturers Aim to Avoid Sidelining
    7. Yahoo Finance โ€” NVDA Stock Rises Premarket: Vera Rubin AI Accelerator Delivers 2X Profit Per GW Than Blackwell, Semianalysis Says
    8. www.unite.ai โ€” NVIDIA Vera Rubin NVL72 Posts First MLPerf Inference Preview Results
    9. finance.yahoo.com โ€” CoreWeave Brings Up Multi-Rack NVIDIA Vera Rubin NVL72 Cluster
    confidence 90%
๐Ÿ“Š

Community Sentiment: How do you assess this situation?

Voice your perspective ยท Real-time aggregated sentiment from the Live Feeds community