Live Feeds
● LIVE Updated 2h ago · 25 sources tracked

AgentX

OpenAI's Jalapeño inference chip outperforms Nvidia's Blackwell system in power efficiency and latency according to public benchmarks. The chip utilizes TSMC's N3P process and features 15.4TB/s of memory bandwidth. While Nvidia reported strong FY2027 Q2 results with 962 billion USD in revenue and 117% growth in data center earnings, the Jalapeño chip threatens Nvidia's CUDA ecosystem and the upcoming Rubin architecture. OpenAI claims the hardware provides 1.9x more throughput per watt and 3.6x lower latency than the Blackwell flagship.

RSS Source map (25)

What changed

SemiAnalysis reports that Jalapeño's performance and cost-efficiency may threaten Nvidia's future Rubin architecture.

Live updates

  1. OpenAI Jalapeño Chip Challenges Nvidia Blackwell and Rubin Performance

    OpenAI's Jalapeño inference chip outperforms Nvidia's Blackwell system in power efficiency and latency according to public benchmarks. The chip utilizes TSMC's N3P process and features 15.4TB/s of memory bandwidth. While Nvidia reported strong FY2027 Q2 results with 962 billion USD in revenue and 117% growth in data center earnings, the Jalapeño chip threatens Nvidia's CUDA ecosystem and the upcoming Rubin architecture. OpenAI claims the hardware provides 1.9x more throughput per watt and 3.6x lower latency than the Blackwell flagship.

    Why it matters

    Nvidia currently leads the agentic AI infrastructure market with hardware like the Vera Rubin NVL72. OpenAI is attempting to reduce its dependence on Nvidia by developing custom silicon. This shift represents a broader industry trend where AI software leaders build proprietary hardware to optimize cost and performance.

    What is confirmed

    • Nvidia reported FY2027 second quarter revenue of 962 billion USD, a 106% increase year-over-year.
    • Nvidia data center revenue grew 117% year-over-year to 890 billion USD.
    • OpenAI's Jalapeño chip exceeds Nvidia's Blackwell system in power efficiency and latency in public benchmarks.

    Still unconfirmed

    • The Jalapeño chip utilizes TSMC's N3P process and offers 15.4TB/s of memory bandwidth.
    • Jalapeño is expected to enter mass production in 2027.
    • Jalapeño's performance and cost-efficiency threaten the upcoming Rubin architecture.

    What to watch next

    • Confirmation of Jalapeño mass production dates
    • Independent third-party verification of internal OpenAI test data
    • Nvidia's official response to Jalapeño benchmark results
    Sources used for this update (5)
    1. wingsoverscotland.com — Wings Over Scotland| The world's most-read Scottish politics website
    2. www.manilatimes.net — SOLOWIN HOLDINGS (NASDAQ: AXG) Announces AI Infrastructure Expansion with 100 MW High-Performance Computing Target
    3. stock.10jqka.com.cn — 营收暴涨106%,却先跌3%:英伟达的超预期,为什么不及格?利空
    4. www.mirrormedia.mg — 黃仁勳惡夢成真?OpenAI晶片實測嗆爆輝達
    5. zaikei.co.jp — OpenAI独自推論チップ「Jalapeño」、Nvidia「Blackwell」比で電力効率最大1.9倍 測定条件には留意
    confidence 80%
  2. Nvidia Vera Rubin NVL72 boosts agentic AI throughput 30-fold

    Nvidia is scaling its agentic AI infrastructure with the Vera Rubin NVL72, which delivers 30 times higher throughput per megawatt than the GB300 NVL72. This hardware supports agents that research, code, and reason through multiple steps. While Nvidia maintains a cost-efficiency lead over AMD in coding-agent benchmarks, OpenAI is challenging this dominance with its custom Jalapeño inference chip. OpenAI claims the Jalapeño chip provides 1.9x more throughput per watt and 3.6x lower latency than Nvidia's Blackwell flagship.

    Why it matters

    Agentic AI requires iterative reasoning and massive context lengths, creating significant power and cost constraints for data centers. Nvidia is combining Vera CPUs, Rubin GPUs, and Groq 3 LPX accelerators to establish a foundational framework for this era. OpenAI is pursuing vertical integration of its AI stack ahead of a planned IPO.

    What is confirmed

    • The Vera Rubin NVL72 delivers up to 30 times higher throughput per megawatt for agentic AI workloads compared to GB300 NVL72 systems.

    Still unconfirmed

    • Nvidia's Vera Rubin NVL72 provides a 35x reduction in cost per token when running DeepSeek V4 Pro models compared to the GB300 NVL72.

    What to watch next

    • Independent verification of Jalapeño chip benchmarks against Blackwell
    • Deployment timelines for Vera Rubin NVL72 on Starmind AI satellites
    • Official IPO filing from OpenAI
    Sources used for this update (7)
    1. uk.news.yahoo.com — One Agent Benchmark Puts Nvidia 5x Ahead Of AMD On Cost
    2. eu.36kr.com — NVIDIA's Stunning First Benchmark of Vera Rubin: DeepSeek Throughput Surges 30-Fold
    3. officechai.com — OpenAI’s New Jalapeno Chip Beats NVIDIA’s Blackwell On Some Parameters, Company Says
    4. technosports.co.in — NVIDIA’s New Chip Delivers 30x More AI Work Per Watt
    5. glitchwire.com — OpenAI's Jalapeño Chip Posts Spicy Benchmark Results That Challenge Nvidia. Here's What It Means for Users.
    6. datacenters.economictimes.indiatimes.com — Nvidia Vera Rubin Boosts AI Throughput Per Megawatt
    7. finance.yahoo.com — Solowin Holdings (AXG) Stock Price, News, Quote & History - Yahoo Finance
    confidence 80%
  3. Nvidia Launches Groq 3 LPX and Vera Rubin for Agentic AI

    Nvidia has entered full production of Groq 3 LPX inference accelerators to support agentic AI, which requires iterative reasoning and massive context lengths. The company is deploying these alongside Vera CPUs and Rubin GPUs. SpaceXAI is adopting the Vera CPU for agent orchestration and plans to deploy Vera Rubin NVL72 hardware on Starmind AI satellites. Performance data for the Vera Rubin NVL72 shows a 30x increase in throughput per megawatt and a 35x reduction in cost per token when running DeepSeek V4 Pro models compared to the GB300 NVL72.

    Why it matters

    Agentic AI differs from standard chatbots by executing hundreds of reasoning steps, calling tools, and managing contexts of hundreds of thousands of tokens. Nvidia is shifting its infrastructure focus from raw training power to token generation speed, energy efficiency, and cost. This transition aims to redefine the economic viability of autonomous AI agents.

    What is confirmed

    • Nvidia has started full production of Groq 3 LPX inference accelerators.
    • SpaceXAI is using NVIDIA Vera CPUs to accelerate agentic AI.
    • The Vera Rubin NVL72 achieves up to 30 times the throughput per megawatt of the GB300 NVL72 when running DeepSeek V4 Pro models.
    • Token costs for DeepSeek V4 Pro can be reduced by up to 35 times on the Vera Rubin platform.
    • Groq 3 LPX reached 3,400 output tokens per second in tests using the Gemma 4 31B model with a 100,000 token context.
    • Nvidia's Groq racks will be online this year following a 20 billion dollar purchase.

    Still unconfirmed

    • Nvidia is 5x ahead of AMD on cost according to one agent benchmark.
    • SpaceXAI will deploy Vera Rubin NVL72 on first-generation Starmind AI satellites.
    • The Vera CPU is 1.8 times faster than x86 for agent AI orchestration.

    What to watch next

    • Deployment of Vera Rubin hardware in Nebius data centers
    • Confirmation of Starmind AI satellite orbital computing performance
    • Further cost-comparison benchmarks between Nvidia and AMD for agentic workloads
    Sources used for this update (14)
    1. SemiAnalysis — AgentX - InferenceXv3: Does CUDA Moat Hold up in Agentic Inferencing?
    2. CNBC — Nvidia says Groq racks will be online this year following $20 billion purchase
    3. NVIDIA Newsroom — SpaceXAI Adopts NVIDIA Vera CPU to Accelerate Agentic AI at Massive Scale
    4. Yahoo Finance — NVIDIA Groq 3 LPX Now in Full Production With World-Class Speed for Agentic AI
    5. The Information — Nvidia Announces New Customers For Vera CPU, Groq LPX Racks
    6. consent.yahoo.com — One Agent Benchmark Puts Nvidia 5x Ahead Of AMD On Cost
    7. news.qq.com — 英伟达新一代算力平台补齐拼图 结合DeepSeek模型Token成本可降35倍
    8. tech.ifeng.com — 英伟达宣布Groq 3 LPX全面投产:Vera Rubin推理效率实现数倍提升
    9. www.stnn.cc — 万亿美元“买债救市”要来了?金价飙升
    10. news.cnyes.com — 輝達宣布Groq 3 LPX全面投產 Vera Rubin推理效率大提升 SpaceX加碼AI代理與太空AI
    11. news.pedaily.cn — 英伟达震撼首测Vera Rubin,DeepSeek吞吐暴涨30倍
    12. www.wikitree.co.kr — 엔비디아 ‘베라’ CPU, 에이전트 AI 오케스트레이션 전담해 x86보다 1.8배 빠르다
    confidence 95%