AI news story

NVIDIA CEO Jensen Huang at Dell Technologies World: ‘Demand Is Going Parabolic, Utterly Parabolic’

Agentic AI inference at one-tenth the cost per token with NVIDIA Vera Rubin NVL72. Agent sandboxes run 50% faster on NVIDIA Vera than traditional CPUs — while enterprise data queries are up to 3x faster with the Vera CPU. And 5,000 enterprises like L

  • Hardware
  • Source: NVIDIA AI Blog
  • Published: 2026-05-18
  • Signal score: 4
  • 31 sources

Editor's take

NVIDIA's CEO, Jensen Huang, declared an unprecedented surge in demand for their AI inference hardware, specifically highlighting the new Vera Rubin NVL72. This acceleration is driven by the promise of significantly reduced inference costs for agentic AI, reportedly one-tenth the cost per token, coupled with substantial performance gains for enterprise data queries and agent sandboxes compared to CPU-based solutions.

This development is critical as it addresses a key bottleneck in deploying AI agents at scale. The economic viability of widespread AI agent adoption hinges on lowering operational expenses, and NVIDIA's Vera platform appears positioned to deliver this. The reported improvements directly impact businesses seeking to leverage AI for complex tasks like data analysis and automated workflows, potentially accelerating the integration of AI into mainstream enterprise operations.

Future observations should focus on whether NVIDIA can meet this "parabolic" demand while maintaining supply chain integrity. Specific metrics to track include the actual cost savings realized by early adopters and the rate at which other hardware vendors can offer comparable performance and efficiency for agentic inference. The success of Vera could also dictate the pace of development and deployment for AI companies reliant on efficient, cost-effective inference.

Signal score: 4

This event was corroborated by 31 independent sources. The signal score weighs cross-source corroboration, recency, source weight and topic salience. How we rank stories.

More Hardware stories

  1. Firebird Takes Its AI Factory Platform Global With a 2-Gigawatt Pipeline

    Unite.AI · 2026-08-08

    Firebird opened its first AI factory in Hrazdan, Armenia, on August 8, 2026, and used the ceremony to lay out the rest of the map: a second market in Kazakhstan with 125 megawatts

  2. NVIDIA AI Releases NOOA: An Object-Oriented Python Framework That Turns an AI Agent Into a Single Python Class

    MarkTechPost · 2026-08-07

    NVIDIA Labs has open-sourced NOOA (NVIDIA Object-Oriented Agents), a model-agnostic Python framework for building AI agents.

  3. Firebird Launches CIS Region’s Largest AI Factory in Armenia

    NVIDIA AI Blog · 2026-08-08

    The global buildout of AI infrastructure reached a new milestone today — Firebird, an emerging AI cloud, launched the CIS region’s largest AI factory in Armenia

  4. d-Matrix Buys Wallaroo to Orchestrate Inference Across Chips

    Unite.AI · 2026-08-03

    d-Matrix has acquired Wallaroo.ai, a maker of software for deploying and orchestrating AI inference, in a deal the Santa Clara chip company announced on August 3, 2026.

  5. ASML Supplier Zeiss Says It Can Handle Demand for Key AI Parts

    Bloomberg · 2026-08-03

    One of the critical suppliers in the semiconductor industry, Germany’s Zeiss Group, pushed back on investor concerns about bottlenecks in the AI supply chain and said it’s

  6. Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B Active Open Weights Multimodal MoE Model

    MarkTechPost · 2026-08-02

    Inkling-Small matches Inkling at a quarter the size, and its NVFP4 checkpoint runs on one NVIDIA B300 GPU