ZenAI
Back to AI News

Jensen Huang: The AI Bottleneck Has Shifted from Chips to Power — Agentic AI Demands 10x More Compute

Industry Insights · May 14, 2026

·May 15, 2026·3 min read

At the 29th Milken Institute Global Conference, Nvidia CEO Jensen Huang made a case that cuts to the heart of where the AI industry is actually headed. His argument: generative AI was the opening act. The real competition — and the real value creation — happens in the age of Agentic AI.

Huang put a striking number on it: compute demand for Agentic AI has grown roughly 1,000% compared to generative AI two years ago. The reason, he explained, is a fundamental shift in what AI is being asked to do. Generative AI handles a single prompt and returns a response. Agents, by contrast, work through a problem iteratively — planning, reasoning, calling external tools, checking their own work, and correcting course — all within a single task. The compute load compounds at every step.

Chips Are No Longer the Constraint

Perhaps the more consequential signal in Huang's remarks was this: the primary bottleneck in AI infrastructure has moved from chips to power. That's a meaningful shift. It means the scarcity is migrating up the stack — away from silicon and toward the grid, the cooling systems, and the real estate that data centers sit on.

This aligns closely with a recent Barclays research note, which found that GPU rack power density has surged from roughly 25 kW per rack in 2020 to 150 kW under the current Blackwell architecture — with projections exceeding 600 kW after Nvidia's Rubin Ultra launch in 2027. At that scale, power supply and thermal management become the long poles in the tent for any data center buildout.

Huang Isn't Talking About the Future. He's Describing the Present.

One detail in Huang's remarks is easy to miss but worth sitting with: he didn't say Agentic AI will arrive. He said it is happening. That's not marketing language — it's a description of current demand.

The most concrete validation is Nvidia's own order book. Enterprises don't commit to large-scale compute purchases on speculation. Earlier this month, Anthropic signed a deal with SpaceX to source more than 300 megawatts of capacity from its Memphis data center — the equivalent of roughly 300,000 H100 GPUs. That infrastructure isn't being provisioned for chatbots. It's being built to run Claude at enterprise scale, across complex, multi-step Agent workflows.

"The demand for agentic compute is expanding AI's reach from the software industry into the $50 trillion physical economy — and this isn't a future event. It's happening now."
— Jensen Huang, Milken Institute Global Conference, May 2026

The 1,000% compute figure isn't just a headline — it has real implications for any business evaluating an AI Agent strategy. Running an Agent is orders of magnitude more expensive than a standard API call. What used to cost a fraction of a cent per query can now run ten times that, depending on task complexity and the number of reasoning steps involved.

That raises a question that the industry hasn't spent nearly enough time on: in which use cases does deploying an AI Agent actually pencil out? Can the productivity gains — fewer headcount hours, faster cycle times, reduced errors — justify the step-change in compute costs?

Goldman Sachs, in its recently published Decoding the Agent Economy report, projects that global AI token demand will grow 24x by 2030. Taken together, these forecasts point to something more significant than a growth story. They signal a fundamental restructuring of cost economics across industries. The businesses that figure out the ROI math early — and build their Agent strategies around it — will be the ones with a durable edge when this market matures.


Sources: Milken Institute Global Conference / Barclays Research / Goldman Sachs Research


Was this article helpful?

Related Articles

Weathered bankrupt airline jet parked on the tarmac, headline reads "An Airline Went Bankrupt, But Google Bought Its Internal Data For $10 Million to Train AI," with a Google logo, a data asset purchase agreement document and $10M price tag on the right, an AI chip icon, and several data documents (Flight Data, Customer Info, Financial Reports, Operations Logs) streaming into the agreement via glowing digital light trails, ZEN logo in the top left corner.

An Airline Went Bankrupt. Google Just Paid $10 Million for Its Internal Data to Train AI.

According to Yahoo Finance, Alphabet, Google's parent company, has won a bankruptcy auction for defunct carrier Spirit Airlines' internal business data with a $10 million bid, saying it will use the data for product development and AI model training. The trove includes 100 million employee emails, 500 million Microsoft Teams chat records, and more than 175,000 employee records dating back to 1986. The deal still needs court approval, expected in September.

Read More
News cover image with a dark blue background featuring the EU flag and star circle beside a Parliament building silhouette, a gavel, and an "AI ACT REGULATION" document in the foreground. A side panel shows a glowing brain icon labeled "GPAI" with callouts for transparency and risk management, plus a "3% of global revenue" penalty badge. Headline: "EU AI Act GPAI Rules Take Effect, Fines Up to 3% of Global Revenue." Bottom callouts: Rules Now Active, Steep Penalties, Broader Industry Impact.

EU AI Act's General-Purpose Model Rules Just Got Teeth — Fines Can Now Hit 3% of Global Revenue

On August 2, 2026, the European Union's enforcement powers over providers of general-purpose AI (GPAI) models under the AI Act officially took effect. According to MediaLaws, this means the European Commission can now actually exercise investigative powers, demand remediation, and issue fines against major model providers like OpenAI, Google, and Meta. These obligations were written into law back in August 2025, but providers were given a full year to adjust — and now that grace period has run out.

Read More
News cover image showing a dark blue courtroom scene with lightning striking down between two silhouetted men facing off before a justice scale statue. To the right, a floating "LLM" display beside a glowing neural-network brain, with prototype hardware — a smart speaker, a square gadget, and AR glasses — plus a "COMPLAINT" document and gavel on the desk. Headline: "Apple Sues OpenAI, Seeks to Halt Its Hardware Plans." Three callouts below: Legal Clash Escalates, IP Dispute, Hardware Plans Blocked.

Apple Just Sued OpenAI — and Now Wants a Judge to Freeze Its Hardware Plans

On July 10, 2026, Apple filed suit against OpenAI in the US District Court for the Northern District of California, accusing OpenAI's hardware chief Tang Yew Tan and former engineer Chang Liu of systematically stealing Apple trade secrets to accelerate development of OpenAI's own AI hardware. In early August, the case escalated further: Apple asked the court for a preliminary injunction that would halt OpenAI's AI hardware development altogether while the case proceeds. Two tech giants that struck a high-profile partnership in 2024 now find themselves in open legal warfare.

Read More