Insights  / Blogs  /
Your AI Bill Is a Workforce Planning Problem
Your AI Bill Is a Workforce Planning Problem
AI token prices are falling. AI budgets are still climbing. The answer is a planning discipline finance already knows: driver-based planning.

Your AI Bill Is a Workforce Planning Problem

AI token prices are falling. AI budgets are still climbing. The answer is a planning discipline finance already knows: driver-based planning.

The AI Paradox Finance Didn’t Plan For

Here is the strangest line item in enterprise finance right now. Over the past year, the blended market price of AI dropped roughly 67%, from about $18.40 to $6.07 per million tokens. Measured against late 2022, the cost of GPT-4-level output has fallen from around $20 to $0.40 per million tokens. By any normal procurement logic, AI should be getting dramatically cheaper to run.

And yet: 73% of enterprises report that their AI costs exceeded original projections, and more than half admit they don’t understand the full scope of what they’re spending. The unit price fell by two-thirds, and the budgets still broke. Newer data shows the gap widening, not closing: in a February 2026 survey of 500 finance leaders, 79% said their organizations experienced AI cost overruns in the past twelve months.

For finance leaders, that contradiction points to a clear conclusion: the issue is not simply unit price. It is consumption volume, and consumption is a planning problem, not a procurement problem.

Finance has managed this cost pattern before. Payroll is shaped by decisions made across the business, and a single top-line budget number rarely explains the result. Workforce planning works because it models the drivers beneath the expense. That same discipline is the right starting point for getting AI cost under control.

When AI costs rise while token prices fall, the missing variable is consumption. Driver-based planning is built to manage exactly that.

Why AI Spending Behaves More Like Payroll Than Software

Most organizations initially book AI as software expense, which is understandable: the invoice arrives from a technology vendor. But traditional software licensing is largely contractual: annual terms, known renewal dates, and relatively predictable expense. AI inference is consumption-based. Every prompt, user, product feature, workflow change, and model selection can change the run rate. You are not simply buying software; you are funding computation by the unit.

Finance does not budget payroll by entering one salary number. It models the underlying drivers: role mix, compensation bands, hiring timing, geography, benefits load, and attrition. Payroll is the output; the drivers are the model.

AI token budgeting has the same architecture with different inputs. Monthly inference expense is the output. The drivers are active users, requests per user, average tokens per request, model mix, and price per token, across every product and workflow that uses AI. Budget the output without modeling the drivers, and variance becomes both predictable and difficult to explain.

The planning parallel is direct:

Planning dimension Workforce planning AI token budgeting
Cost driver Headcount × compensation Consumption × unit price
Variability Low month to month; spikes at hiring events High; can change with launches or model changes
Budget owner Finance + HR Finance + Engineering + Product
What it measures People capacity Computational capacity
Forecast method Headcount plans, role-based rates Usage patterns, pricing, and growth projections
Waste signal Unfilled headcount, underutilization Runaway calls, redundant models, inefficient prompts
Scenario inputs Hiring pace, attrition, salary bands Model mix, usage growth, prompt efficiency
Source of surprise Unplanned backfill, comp adjustments Usage spikes, model changes, new AI features
In NSPB today? Yes (native workforce planning module) Yes (custom dimensions + driver model)

Both are variable, consumption-based, cross-functional costs. Both need driver-based planning, regular reforecasting, and clear ownership to be managed well.

Where the Analogy Breaks Down

The workforce parallel is the right starting point. But three differences should shape the AI cost model from the outset.

1. AI moves faster than people

Hiring takes months and attrition is roughly predictable. AI consumption can double in a week: a product launch, a viral internal tool, one engineer restructuring prompts. Agentic workflows compound this: a single autonomous agent workflow can consume 10–50 times the tokens of a simple query, which is why Goldman Sachs projects total token consumption will grow roughly 24-fold by 2030 even as prices fall. An annual AI budget is stale before the second quarter. Rolling forecasts aren’t a best practice here; they’re the only format that makes sense.

2. Ownership crosses every line on the org chart

Payroll rolls up relatively cleanly: departments own headcount, and headcount rolls to the P&L. AI rarely follows the org chart so neatly. A single model can support customer service, marketing, product, and finance. Without tags and dimensions, the provider invoice cannot be allocated, and costs that cannot be allocated cannot be governed. Finance needs visibility into how engineering instruments API calls, not just the bill that arrives at month-end.

3. The unit economics won’t hold still

Salary bands typically change on an annual cadence. Model pricing, capabilities, and the right model for a task can change overnight. A budget built on today’s model mix may be materially wrong by the next quarter even if usage is unchanged. AI planning must scenario-model the unit cost as well as the volume, for example by testing the impact of a lower-cost model becoming viable midyear.

A Practical AI Cost Model in NSPB

NetSuite Planning and Budgeting (NSPB) does not need a dedicated AI cost module to support AI token budgeting. Its driver-based planning capabilities provide the core structure: define the dimensions, model the consumption drivers, forecast scenarios, and compare the plan with actual usage. In practice, the work follows four steps.

Step 1: Define the cost dimensions

Start with the axes that make AI spend explainable: provider (for example, OpenAI, Anthropic, Google, or Azure), model or model tier, consuming team or business unit, product or workflow, and environment (production versus development). These dimensions become both the planning structure and the allocation logic for actuals.

Step 2: Build the driver tree

The driver tree mirrors a headcount model with different math. Instead of headcount × compensation rate, the core calculation is:

Active users × requests per user × average tokens per request × model price per token = monthly AI expense

Each driver should be maintained independently. That makes the forecast useful when pricing shifts, usage changes, or a new workflow goes live.

Step 3: Model the scenarios

Workforce scenarios test hiring pace and attrition. AI scenarios should test model selection, usage growth, and prompt efficiency: What if usage doubles? What if a high-volume workflow moves to a model at one-third the price? What if prompt optimization reduces tokens per request by 20%? These decisions need engineering inputs, but their financial impact belongs in the planning model before the decision is made, not after.

Step 4: Connect the actuals

The model earns trust when actuals flow back into it. Major providers expose usage data through APIs or exports. The practical challenge is tagging API calls with the chosen dimensions and moving provider billing data into the GL structure so NSPB can report budget versus actual at the level of real consumption. The difficult part is rarely the data source; it is agreeing, across finance and engineering, who owns each tag.

Six Questions to Ask Before the Next Budget Cycle

If your organization already spends materially on AI inference (or expects to within the next twelve months), these questions reveal whether the planning foundation is ready:

  1. Do we know which teams and products generate AI costs, and in what proportion, or do we see only one aggregate invoice?
  2. Are AI costs tracked in the GL at a dimensional level that supports allocation and variance analysis?
  3. Who in engineering owns the instrumentation that tags usage by model, team, and workflow?
  4. Have we modeled a 2× usage increase and a midyear model-price change?
  5. Is our AI forecast a fixed annual number or a rolling model that updates as usage data arrives?
  6. Do we have a scenario for migrating high-volume workflows to lower-cost models as they become viable?

Organizations that cannot answer these questions are managing AI reactively. That may be tolerable while AI remains an innovation-budget line item. It is not tolerable once the bill becomes visible to leadership and the board. The stakes are rising on the value side as well: in the same February 2026 survey, only 15% of finance leaders said they could calculate AI ROI without significant bottlenecks, while 83% expect clear, quantifiable returns within twelve months. A cost model built on explainable drivers is the prerequisite for closing that gap.

What Myers-Holum Is Seeing

Across our NSPB client base, the question has flipped in about eighteen months: from “how can we use AI?” to “how do we govern AI financially?” That mirrors the broader market: 98% of FinOps practitioners now manage AI spend, up from 63% a year earlier. The pattern we see is consistent: a company has been booking AI as a lump sum in a technology expense account, the number has grown large enough to appear in board discussions, and now finance needs a model.

Once the dimensional structure is agreed and instrumentation is available, an NSPB AI cost model can often be built in four to six weeks. The more important work is the finance-and-engineering conversation: what data exists today, what tags are attached to API calls, and what is required to create a reliable actuals feed. That is fundamentally a planning conversation: the same kind of alignment finance and HR establish for workforce planning.

The organizations making the fastest progress are not necessarily the ones spending the most on AI. They are the ones that establish the planning model before the cost becomes material.

The Takeaway

Every major technology shift eventually becomes a finance problem. Cloud computing required new infrastructure planning. Subscription models changed revenue forecasting. AI is now reshaping operating-expense planning, but it is not asking finance to invent a new discipline.

AI is not introducing a new planning problem. It is introducing a new consumption driver. Finance already knows how to plan variable, driver-based costs; the opportunity is to apply that discipline before surprises become recurring variances.

Teams that wait for engineering alone to solve AI cost visibility will keep explaining budget variances after the fact. Teams that build the model now, even an imperfect first version, can manage the spend as it changes.

Ready to Build an AI Cost Planning Model in NSPB?

Myers-Holum helps finance teams define the dimensions, build the driver model, and connect provider-billing actuals to NSPB’s budget-versus-actual framework. If AI spending is growing faster than your ability to forecast it, the next step is a diagnostic conversation about the data, drivers, and governance a practical model requires.

This post is part of Myers-Holum’s EPM series for NetSuite Planning and Budgeting leaders: practical planning frameworks you can take back to your team, not product overviews.

References

  1. Blended AI cost per million tokens fell from $18.40 to $6.07 between Q1 2025 and Q1 2026. https://optimumpartners.com/insight/ai-token-costs-and-how-they-might-wreck-your-budget/        
  2. Epoch AI, LLM inference price trends. https://epoch.ai/data-insights/llm-inference-price-trends
  3. inOps Foundation, State of FinOps 2026 (survey of 1,192 organizations representing $83B in annual technology spend). https://data.finops.org/
  4. DoiT International / Sapio Research survey of 500 US and UK finance leaders at organizations with 1,000+ employees, February 2026. https://www.doit.com/blog/ai-spending-survey
  5. Cockroach Labs, The Bill Arrives: How to Manage Agentic AI Costs at Scale. https://www.cockroachlabs.com/blog/agentic-ai-costs-at-scale/ Token growth projection: Goldman Sachs, AI Agents Forecast to Boost Tech Cash Flow as Usage Soars. https://www.goldmansachs.com/insights/articles/ai-agents-forecast-to-boost-tech-cash-flow-as-usage-soars
  6. DoiT International / Sapio Research, February 2026 (see note 6). https://www.doit.com/blog/ai-spending-survey
  7. Linux Foundation / FinOps Foundation press release: 98% of FinOps practitioners now manage AI spend, up from 63% in 2025. https://www.linuxfoundation.org/press/state-of-finops-survey-ai-value-and-skills-top-priorities-as-finops-matures-across-technology-value-98-manage-ai-90-saas-64-licensing-48-data-center-1
//
Related POsts
IBS Electronics + Myers-Holum, Inc.: Winter 2024 Alliance Partner Spotlight Award Winner
Myers-Holum announces top Solution Architecture team who will help guide business process optimization implementing NetSuite cloud solutions.
Read article
Bose Professional + Myers-Holum, Inc.: Summer 2024 Alliance Partner Spotlight Award Winner
Myers-Holum announces top Solution Architecture team who will help guide business process optimization implementing NetSuite cloud solutions.
Read article
Oracle Integration Cloud Release 3
Oracle Integration Cloud Release 3 makes OIC the last NetSuite integrator your enterprise will ever need. Learn why.
Read article

Let’s discuss your next project