Skip to content
MarketScale
‹ Back to IndustriesSoftware & Technology

Palo Alto Networks CEO puts a number on the AI cost problem: 90% token price drop needed

Nikesh Arora, CEO of Palo Alto Networks, stated that for enterprise AI to scale, token costs must decrease by 90% within two years. He highlighted that high costs have already impacted companies like Uber, which spent its full-year AI budget by April.

This story was produced through MarketScale. See how Software & Technology teams put it to work with Executive Thought Leadership.

By MarketScale Newsroom · Palo Alto NetworksEnterprise AiToken CostsAi Adoption
Share
Learn this in 60 seconds

Key facts, context, and what it means, in one minute.

:60
0:001:00
Palo Alto Networks CEO puts a number on the AI cost problem: 90% token price drop needed

Key takeaways

01

Token costs for AI need to decline by 90% in two years for scalability.

02

Uber exhausted its annual AI budget by April due to high costs.

Palo Alto Networks CEO Nikesh Arora put precise numbers on the enterprise AI cost problem July 9, telling CNBC's Squawk on the Street that token prices need to fall 20% within a year and 90% the year after before companies can realistically scale AI workloads. The remarks land as mounting evidence shows enterprises are already pulling back spending they committed to earlier in 2026.

The cost ceiling is real and already being hit

Uber is the clearest data point. According to PYMNTS, the company burned through its entire 2026 AI budget by April. Chief Operating Officer Andrew Macdonald said Uber would weigh token costs directly against the cost of hiring engineers, a comparison that would have seemed far-fetched two years ago. CTO Praveen Neppalli Naga described the situation as being "back to the drawing board."

Uber's situation is not isolated. PYMNTS reported in June that companies that once encouraged broad internal AI tool adoption, when costs were lower, are now rationing access through usage caps, nudging employees toward task-appropriate models, and routing lower-stakes work to older, cheaper options. The economics shifted faster than most IT and procurement teams planned for.

When Arora was asked about OpenAI CEO Sam Altman's claim that OpenAI's latest model is 54% more efficient for coding, Arora said the improvement is a good start but not sufficient. "I think we probably need another turn at it," he said, per CNBC. The comment signals that even headline efficiency gains from frontier model vendors are not closing the gap fast enough for enterprise buyers.

Agentic tools amplify the exposure

Standard chatbot interactions generate a single inference call per exchange. Agentic coding tools, which complete multi-step tasks autonomously, generate many inference calls per session. That structural difference means enterprises that deployed agentic tools based on chatbot-era cost assumptions are seeing usage bills that scale non-linearly with adoption, according to PYMNTS.

For operations and IT leaders, this is a procurement design problem. Budgets built on per-seat or per-user assumptions break down when the actual unit of cost is inference volume, which varies sharply by use case, user behavior, and model selection.

Cheaper alternatives are gaining ground

The cost pressure is creating an opening for lower-priced alternatives. PYMNTS reported in June that Chinese AI labs are attracting attention from enterprise buyers because their more efficient models and China's lower energy costs let them undercut U.S. providers on price. Procurement teams evaluating AI vendors in 2026 are now treating price per token as a primary selection criterion alongside capability benchmarks.

Open-source models are also seeing renewed interest. Companies are deploying them for internal or lower-risk tasks where a frontier model's performance advantage does not justify the cost premium. That tiered-model approach is becoming standard practice for cost-conscious AI programs.

Budget discipline is replacing blank-check experimentation

The PYMNTS Intelligence Enterprise AI Benchmark Report found that enterprises across financial services, insurance, healthcare, and media and advertising are continuing to increase AI budgets in 2026. But the report also noted a meaningful shift in posture: companies are becoming more selective, deciding which projects warrant real capital and which still need to prove their value before receiving it.

That selectivity is the direct operational consequence of token shock. Arora's 90% cost-reduction benchmark gives procurement and IT leaders a concrete yardstick: at current prices, broad deployment is financially constrained. At prices 90% lower, the economics of many use cases flip.

What this means for your team

  • Audit your AI cost structure by use case now. Separate agentic workloads from single-turn interactions in your tracking; they have fundamentally different cost profiles and need separate budget lines.
  • Build model-tiering into your AI procurement policy. Define which tasks require frontier models and which can run on older, open-source, or lower-cost alternatives. Cost governance should be a design requirement, not an afterthought.
  • Add price-per-token to your vendor evaluation scorecard. Capability benchmarks alone no longer tell the full story. Efficiency metrics and pricing trajectories are equally material for multi-year contracts.
  • Establish a cost-reduction trigger in your AI roadmap. Arora's 20%/90% timeline gives you a concrete signal to watch. If token prices hit those thresholds on schedule, use cases that are marginal today may become viable, and your deployment plan should account for that shift.

Featured companies

About the author

MarketScale Newsroom
MarketScale NewsroomEditorial Team, MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

Software & Technology: are you visible to AI?

Before they reach out, Software & Technology buyers ask AI engines which vendors to trust. See how AI describes your company today, and where competitors show up instead.

Free workspace

You just read one Software & Technology expert. Imagine publishing your whole team.

This article was produced through MarketScale. Create a free workspace and turn your own team's Software & Technology expertise into the articles, video, and social content B2B marketing buyers in your industry are searching for. No credit card, no demo required.

NPS +73 · 1,000+ creators · 38+ countries

What you get, free

Your own MarketScale Studio workspace
One video edit a month, on us
AI writing, editing, and publishing tools
In-platform coaching to learn the system

More Software & Technology Insights

AI startups are proving they can build real businesses, and Forbes' 2026 lists show exactly where the money is going

AI startups are proving they can build real businesses, and Forbes' 2026 lists show exactly where the money is going

Forbes' 2026 AI rankings highlight how AI startups are effectively building substantial businesses. These rankings reveal the concentration of enterprise AI investments, ranging from startups with $30 billion in revenue run rates to those valued under $1 billion.

  • 01AI startups are achieving up to $30 billion in revenue run rates.
  • 02The Forbes 2026 rankings identify where enterprise AI investments are focusing.
  • 03AI unicorns valued under $1 billion are also emerging.

Aug 4, 2026

QCi splits its CRO role in two, hiring Susan Hunt to lead revenue as Pouya Dianat moves to a new chief product officer seat

QCi splits its CRO role in two, hiring Susan Hunt to lead revenue as Pouya Dianat moves to a new chief product officer seat

Quantum Computing Inc. has restructured its executive team by appointing Susan Hunt as the Chief Revenue Officer and creating a Chief Product Officer role for Pouya Dianat. This move separates the responsibilities of commercial execution and product strategy.

  • 01Quantum Computing Inc. has appointed Susan Hunt as Chief Revenue Officer to lead revenue generation.
  • 02The company created a new Chief Product Officer position for Pouya Dianat to focus on product strategy.
  • 03This organizational change reflects a strategic separation of commercial and product development responsibilities.

Aug 4, 2026

QCi splits its CRO role in two, hiring Susan Hunt to run revenue as Pouya Dianat moves to a new chief product officer seat

QCi splits its CRO role in two, hiring Susan Hunt to run revenue as Pouya Dianat moves to a new chief product officer seat

Quantum Computing Inc. has appointed Susan Hunt as Chief Revenue Officer and Pouya Dianat as Chief Product Officer. This organizational change reflects a strategic focus on distinguishing sales from product development to enhance operational effectiveness.

  • 01Quantum Computing Inc. has separated its CRO role into two positions: Chief Revenue Officer and Chief Product Officer.
  • 02Susan Hunt will focus on driving revenue as the newly appointed Chief Revenue Officer.
  • 03Pouya Dianat will lead product development and innovation as the Chief Product Officer.

Aug 3, 2026

Explore More Software & Technology Insights

Read more expert perspectives from across Software & Technology.

Browse Software & Technology Hub

About the Expert

MarketScale Newsroom
MarketScale Newsroom

Editorial Team

MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

For B2B teams

Your experts could be publishing here

Stories like this one run on content MarketScale captures from real practitioners. See how your team's expertise becomes coverage in Software & Technology and beyond.

Book a 15-minute demo

Or call us. No forms required. We pick up. 214-945-2512