Skip to content
MarketScale
‹ Back to IndustriesSoftware & Technology

Palo Alto Networks CEO puts a number on the AI cost problem: 90% token price drop needed

Nikesh Arora, CEO of Palo Alto Networks, stated that for enterprise AI to scale, token costs must decrease by 90% within two years. He highlighted that high costs have already impacted companies like Uber, which spent its full-year AI budget by April.

This story was produced through MarketScale. See how Software & Technology teams put it to work with Executive Thought Leadership.

By MarketScale Newsroom · Palo Alto NetworksEnterprise AiToken CostsAi Adoption
Share
Learn this in 60 seconds

Key facts, context, and what it means, in one minute.

:60
0:001:00
Palo Alto Networks CEO puts a number on the AI cost problem: 90% token price drop needed

Key takeaways

01

Token costs for AI need to decline by 90% in two years for scalability.

02

Uber exhausted its annual AI budget by April due to high costs.

Get featured

Want to get featured in MarketScale Software & Technology?

Create a free MarketScale workspace and get your company's expertise featured across our Software & Technology coverage. No credit card, no demo required.

Request an invite

Palo Alto Networks CEO Nikesh Arora put precise numbers on the enterprise AI cost problem July 9, telling CNBC's Squawk on the Street that token prices need to fall 20% within a year and 90% the year after before companies can realistically scale AI workloads. The remarks land as mounting evidence shows enterprises are already pulling back spending they committed to earlier in 2026.

The cost ceiling is real and already being hit

Uber is the clearest data point. According to PYMNTS, the company burned through its entire 2026 AI budget by April. Chief Operating Officer Andrew Macdonald said Uber would weigh token costs directly against the cost of hiring engineers, a comparison that would have seemed far-fetched two years ago. CTO Praveen Neppalli Naga described the situation as being "back to the drawing board."

Uber's situation is not isolated. PYMNTS reported in June that companies that once encouraged broad internal AI tool adoption, when costs were lower, are now rationing access through usage caps, nudging employees toward task-appropriate models, and routing lower-stakes work to older, cheaper options. The economics shifted faster than most IT and procurement teams planned for.

When Arora was asked about OpenAI CEO Sam Altman's claim that OpenAI's latest model is 54% more efficient for coding, Arora said the improvement is a good start but not sufficient. "I think we probably need another turn at it," he said, per CNBC. The comment signals that even headline efficiency gains from frontier model vendors are not closing the gap fast enough for enterprise buyers.

Agentic tools amplify the exposure

Standard chatbot interactions generate a single inference call per exchange. Agentic coding tools, which complete multi-step tasks autonomously, generate many inference calls per session. That structural difference means enterprises that deployed agentic tools based on chatbot-era cost assumptions are seeing usage bills that scale non-linearly with adoption, according to PYMNTS.

For operations and IT leaders, this is a procurement design problem. Budgets built on per-seat or per-user assumptions break down when the actual unit of cost is inference volume, which varies sharply by use case, user behavior, and model selection.

Cheaper alternatives are gaining ground

The cost pressure is creating an opening for lower-priced alternatives. PYMNTS reported in June that Chinese AI labs are attracting attention from enterprise buyers because their more efficient models and China's lower energy costs let them undercut U.S. providers on price. Procurement teams evaluating AI vendors in 2026 are now treating price per token as a primary selection criterion alongside capability benchmarks.

Open-source models are also seeing renewed interest. Companies are deploying them for internal or lower-risk tasks where a frontier model's performance advantage does not justify the cost premium. That tiered-model approach is becoming standard practice for cost-conscious AI programs.

Budget discipline is replacing blank-check experimentation

The PYMNTS Intelligence Enterprise AI Benchmark Report found that enterprises across financial services, insurance, healthcare, and media and advertising are continuing to increase AI budgets in 2026. But the report also noted a meaningful shift in posture: companies are becoming more selective, deciding which projects warrant real capital and which still need to prove their value before receiving it.

That selectivity is the direct operational consequence of token shock. Arora's 90% cost-reduction benchmark gives procurement and IT leaders a concrete yardstick: at current prices, broad deployment is financially constrained. At prices 90% lower, the economics of many use cases flip.

What this means for your team

  • Audit your AI cost structure by use case now. Separate agentic workloads from single-turn interactions in your tracking; they have fundamentally different cost profiles and need separate budget lines.
  • Build model-tiering into your AI procurement policy. Define which tasks require frontier models and which can run on older, open-source, or lower-cost alternatives. Cost governance should be a design requirement, not an afterthought.
  • Add price-per-token to your vendor evaluation scorecard. Capability benchmarks alone no longer tell the full story. Efficiency metrics and pricing trajectories are equally material for multi-year contracts.
  • Establish a cost-reduction trigger in your AI roadmap. Arora's 20%/90% timeline gives you a concrete signal to watch. If token prices hit those thresholds on schedule, use cases that are marginal today may become viable, and your deployment plan should account for that shift.

Featured companies

Your experts belong here

Every story in MarketScale Software & Technology starts with a company putting its solutions engineers, product teams, and customer engineers on the record. Buyers are already reading this topic. The only question is whose experts they find.

Buyers ask AI engines who to consider, and published expert answers are what those engines cite.

Get your team featuredSee how it works15 minutes, straight to a calendar.

About the author

MarketScale Newsroom
MarketScale NewsroomEditorial Team, MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

Follow Software & Technology Insights

Get new expert content in your inbox.

Software & Technology: are you visible to AI?

Before they reach out, Software & Technology buyers ask AI engines which vendors to trust. See how AI describes your company today, and where competitors show up instead.

Free workspace

You just read one Software & Technology expert. Your company is full of them.

This article was produced through MarketScale. The same platform turns your solutions engineers, product teams, and customer engineers into the articles, video, and social content Software & Technology buyers are searching for. Create a free workspace and see it with your own people. No credit card, no demo required.

NPS +73 · 1,000+ creators · 38+ countries

What you get, free

Your own MarketScale Studio workspace
One video edit a month, on us
AI writing, editing, and publishing tools
In-platform coaching to learn the system

More Software & Technology Insights

OpenAI's GPT-6 Astra pitch is to skip integrations and run the software UI itself

OpenAI's GPT-6 Astra pitch is to skip integrations and run the software UI itself

OpenAI shipped GPT-6 Astra on Sept. 3 with a "computer use" capability that lets the model operate existing software UIs directly instead of requiring custom API integrations. The staged rollout through Daybreak, ChatGPT tiers, and AWS shifts the automation bottleneck from building connectors to governing UI-driven sessions, with speed measured in minutes per task as the cost input.

  • 01Astra operates software via pixels, keyboard, and mouse interactions to bypass API integration work on the long tail of internal tools without clean API access
  • 02OpenAI reported Astra at 40 minutes per task (47% faster than GPT-5.6 Sol at 75 minutes), making task time the practical proxy for compute cost modeling and throughput evaluation
  • 03Enterprises must define governance before broad rollout: eligible workflows for UI automation, audit logging systems, identity and secrets handling, and fallback procedures when UIs change or sessions break

Sep 3, 2026

Cloud is set to take 26% of IT budgets, and hiring is shifting toward platforms

Cloud is set to take 26% of IT budgets, and hiring is shifting toward platforms

Foundry’s 2026 Cloud Computing Study, cited by CIO, reports IT leaders expect 26% of IT budgets to go to cloud computing in the next year and that 74% accelerated cloud migrations in the last 12 months. At the same time, CIO’s coverage of Robert Half Technology’s 2026 IT salary report shows AI/ML engineers at a $170,750 median salary and a striking capability gap, with only 7% of leaders saying they have the capabilities to complete prioritized projects and 65% expecting to upskill existing staff. Separate reporting from Nextgov on senior appointments in the Pentagon CIO office, and Government Technology’s account of Illinois’ multi-agency data-sharing MOU, indicate large organizations are staffing for organizational change, governance, and cross-domain data sharing, not just for lift-and-shift migrations. For enterprise operators, the signal is that cloud programs increasingly depend on internal platform engineering, adoption enablement, and finance-aligned governance, because those functions connect migration speed to day-two operations and business results.

  • 01A useful benchmark for 2026 planning: Foundry’s survey puts cloud at 26% of IT budget next year, so showback/FinOps and workload-level chargeback governance can’t stay a side project (per CIO citing Foundry).
  • 02Talent pricing has become an input to architecture decisions. A CIO, citing Robert Half, compared the median pay for AI/ML engineers ($170,750) with systems administrators ($98,000), arguing that platform standardization and self-service guardrails are as much about labor strategy as technology.
  • 03If 65% of leaders expect to upskill to close gaps, the differentiator becomes your internal adoption system, in-app guidance, runbooks, and training telemetry, not the third new tool in the stack (per CIO citing Robert Half; illustrated by Ferring’s Whatfix rollout reported by BankInfoSecurity).

Sep 2, 2026

AI agents are pushing access controls and testing into the data layer

AI agents are pushing access controls and testing into the data layer

Snowflake is warning that dashboard-era access controls do not hold up once AI agents can query and act across datasets, pushing governance closer to the data layer, according to TechTarget’s Computer Weekly. In parallel, TechTarget reported that enterprise AI-agent testing needs to expand beyond pre-deployment checks into continuous monitoring so agents don’t drift beyond prescribed instructions. InformationWeek’s reporting on CISOs at Intuit, Smartsheet and ETS adds the operational risk: unmanaged “AI orphans” and identity sprawl as agent count grows, which shifts near-term workload onto IAM, data governance and platform engineering teams building the guardrails.

  • 01If an AI agent can reach multiple tools, the real control plane becomes the data layer and identity, not the BI dashboard permissions model.
  • 02Agent rollouts that stop at pre-production testing are likely to miss the failure mode operators actually see, post-deploy tool changes that alter what the agent can do.
  • 03“AI orphans” is a practical inventory problem: if teams can’t enumerate agents, they can’t set ownership, secrets rotation, or access reviews on a schedule.

Sep 2, 2026

Explore More Software & Technology Insights

Read more expert perspectives from across Software & Technology.

Browse Software & Technology Hub

About the Expert

MarketScale Newsroom
MarketScale Newsroom

Editorial Team

MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

For B2B teams

Your experts could be publishing here

Stories like this one run on content MarketScale captures from real practitioners. See how your team's expertise becomes coverage in Software & Technology and beyond.

Book a 15-minute demo

Or call us. No forms required. We pick up. 214-945-2512