Skip to content
‹ Back to IndustriesSoftware & Technology

AI budgets are burning out before year-end, and CFOs are rethinking every token

Enterprise AI costs are exceeding their allocated budgets quickly, prompting CFOs to reassess spending strategies. Despite the high expenditure, the return on investment remains questionable as companies often use up their budgets in just a few months. CFOs are now balancing resource allocation between workforce and AI-related token expenses.

This story was produced through MarketScale. See how Software & Technology teams put it to work with Executive Thought Leadership.

By MarketScale Newsroom · · Enterprise AiAi Cost ManagementModel RoutingCfo Strategy
Share
Listen to the audio brief

Key facts, context, and what it means.

AUDIO
0:00—
AI budgets are burning out before year-end, and CFOs are rethinking every token

Key takeaways

01

Enterprise AI budgets are being exhausted rapidly.

02

CFOs are reevaluating the balance between headcount and AI spending.

03

ROI on AI investments remains uncertain.

Free workspace

Turn your Software & Technology expertise into content.

Record interviews, organize footage, and write with AI on a free trial of the MarketScale platform for qualifying companies. No demo required, no credit card.

Try it Free

Annual AI budgets are running dry in one to two months at major U.S. companies. That is not a projection; it is what enterprise technology buyers are reporting directly to their vendors right now, according to CNBC. The dynamic is reshaping how operations and finance leaders think about AI spending and, increasingly, about headcount.

Arvind Jain, CEO of enterprise AI company Glean, told CNBC that AI cost is now the top concern across every enterprise conversation his team has. The culprit: frontier model pricing has not fallen the way buyers expected. Instead, each successive model release has come in at roughly twice the per-token cost of the previous one, putting what Jain described to CNBC as an unsustainable trajectory in front of CFOs who set budgets based on older, cheaper baselines.

Tech priced like people

The cost dynamic has opened a comparison that simply did not exist in previous technology cycles. Jain told CNBC this is the first time he can recall that the price of technology is comparable to the cost of a person, forcing a direct trade-off that historically never came up because technology was always a fraction of operating costs. Growing AI budgets are now coming, at least in part, in lieu of future headcount additions.

Matan Grinberg, CEO of Factory AI, framed it as a concrete resource allocation question playing out inside leadership teams: whether to optimize the number of employees or the AI spend per employee. Speaking to CNBC, he described three distinct phases companies have moved through in roughly a year. First, boards demanded action on AI. Then came what he called tokenmaxxing, deploying AI by any means necessary regardless of cost. Now, leadership teams are in a third phase, actively reassessing whether premium frontier models are necessary for every task.

Do we need to be using Opus-level intelligence for every single task? You just don't need to., Matan Grinberg, CEO, Factory AI, via CNBC

95% on the expensive tier, and an obvious fix

One concrete number stands out in the CNBC reporting: roughly 95% of enterprise AI usage is still running on the most expensive frontier models, even for tasks that cheaper alternatives could handle without meaningful quality loss. Jain put the available savings from smarter model routing at roughly 10x for organizations willing to direct simpler workloads to lower-cost tiers.

Factory AI's product is built around exactly that premise, automatically matching each task to the most cost-appropriate model. Grinberg illustrated to CNBC how marginal the real-world performance gap often is between successive frontier releases, comparing the difference to that between a professor with 13 years of experience and one with 15. For most enterprise tasks, the distinction is imperceptible.

Adoption is not the same as value

The cost squeeze lands on top of a separate but related problem: widespread AI adoption has not reliably translated into measurable business impact. Commencis, in its 2026 State of AI: Hype, Reality, and What Comes Next report, argues that individual productivity gains, the typical first-wave benefit, do not automatically compound into enterprise-level value. The report's thesis is that realizing that next level of return requires organizations to redesign the structure of work itself, not just deploy AI tools within existing workflows.

Taken together, the two pressures point in the same direction. Buying more AI access at current prices, without changing how work is organized or how models are selected for each task, produces neither cost control nor scalable value. Operations leaders who treat AI as a software subscription rather than as a variable-cost infrastructure requiring active management are the ones most likely to hit the budget wall.

What this means for your team

  • Audit model usage now: identify what share of your AI workloads are running on frontier-tier models and whether simpler tasks could be routed to cheaper alternatives without quality impact.
  • Rebuild the budget model: per-token pricing has been rising, not falling; finance and IT teams should stress-test AI line items against continued cost escalation rather than assuming the relief that was previously expected.
  • Move beyond tool deployment: evaluate whether current AI initiatives are changing how work flows at the process level, not just adding AI features to individual roles, as system-level redesign is where measurable value is more likely to emerge.
  • Tie headcount planning to AI cost projections: with CFOs now explicitly weighing token spend against future hiring, operations and HR leaders need visibility into AI consumption data to participate meaningfully in that trade-off conversation.

Featured companies

Your experts belong here

Every story in MarketScale Software & Technology starts with a company putting its solutions engineers, product teams, and customer engineers on the record. Buyers are already reading this topic. The only question is whose experts they find.

Buyers ask AI engines who to consider, and published expert answers are what those engines cite.

Book DemoSee how it works15 minutes, straight to a calendar.

About the author

MarketScale Newsroom
MarketScale NewsroomEditorial Team, MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

B2B Weekly

The week in Software & Technology, and sixteen other industries, every Monday.

Ten stories, one-line takes, five minutes. Free.

Software & Technology: are you visible to AI?

Before they reach out, Software & Technology buyers ask AI engines which vendors to trust. Explore how your experts, customers, and partners can become useful content for buyers and AI search.

Free Trial

You just read one Software & Technology expert. Your company is full of them.

This article was produced through MarketScale. The same platform turns your solutions engineers, product teams, and customer engineers into the articles, video, and social content Software & Technology buyers are searching for. Start a free trial and see it with your own people. For qualifying companies, no credit card, no demo required.

NPS +73 · 1,000+ creators · 38+ countries

What your free trial includes

Hands-on access to the MarketScale platform
Media requests to your crowd, remote recording, AI writing tools
No demo required. No credit card.
For qualifying companies. Company confirmation required.

More Software & Technology Insights

Hammond tells MSPs to pick a vertical and one problem

Hammond tells MSPs to pick a vertical and one problem

N-able Head Nerd Stefanie Hammond lays out a 10-step growth playbook for MSPs that have built spare capacity, starting with one industry and one problem and a 100-account target list. She argues MSPs have a sales problem more than a lead problem, and points to her go-to-market accelerator program, with classes starting in November.

  • 01Hammond starts with revenue targets, then one industry and one problem, then a 100-account list split 20, 30 and 50.
  • 02Hammond says MSPs have a sales problem more than a lead problem: find where proposals stall and whether the right decision makers are in the room before paying for more leads.
  • 03The Dream 100 name traces to a case Chet Holmes International describes: 167 of 2,200 newspaper advertisers bought 95% of the ads, and the first close took five months of mail and calls.

Oct 5, 2026

SDLC Corp launches Pulastya AI, a voice agent platform that answers business calls from a company's own documents

SDLC Corp launched Pulastya AI, a voice agent platform that answers and places business calls 24/7 using company documents without requiring a long integration process. The platform connects to existing phone numbers and OpenAI accounts, handles calls that it cannot answer by transferring to staff, and saves full transcripts for context.

  • 01Most teams can complete setup in under 30 minutes once Twilio/Exotel, OpenAI, and documents are ready
  • 02Internal tests showed ~500 ms response; it won’t guess and hands off to staff
  • 03Designed for clinics, banks, real estate offices, hotels, schools and support teams handling administrative and informational calls like appointments, bookings, order status and inquiries

Oct 3, 2026

Google Just Put AI Chips in Orbit. The Real Story Is the Power Bill on the Ground.

Google Just Put AI Chips in Orbit. The Real Story Is the Power Bill on the Ground.

Google launched a prototype satellite carrying four Trillium TPUs to test whether the chips can survive launch and operate under orbital radiation and heat constraints, driven by power limits on the ground. It is not a data center but a survival test for durability, radiation resistance, and heat dissipation, signaling that energy availability—not chips or models—is becoming the limiting factor for AI infrastructure growth.

  • 01Project Suncatcher's first satellite is a survival test for hardware durability, not operational compute capacity for actual workloads
  • 02Orbital solar can deliver up to 8x more power, and Google research suggests launch costs could drop below $200/kg by the mid-2030s, bringing space build costs closer to some Earth equivalents.

Oct 1, 2026

Explore More Software & Technology Insights

Read more expert perspectives from across Software & Technology.

Browse Software & Technology Hub

About the Expert

MarketScale Newsroom
MarketScale Newsroom

Editorial Team

MarketScale

The MarketScale Newsroom reports on the companies, technologies, and trends shaping 16 B2B industries. It turns primary sources and expert commentary into clear, useful coverage for the people doing the work.

For B2B teams

Your experts could be publishing here

Stories like this one run on content MarketScale captures from real practitioners. See how your team's expertise becomes coverage in Software & Technology and beyond.

Book a Demo

Or call us. No forms required. We pick up. 214-945-2512