Wakalix/Guides/A question to take to your next review

Where does your AI spend actually go?

Most engineering organisations can say what they paid for AI last quarter. Far fewer can say what it bought. This guide walks through the questions that close that gap, and ends with a short exercise you can run on your own numbers.

An eight-minute readFor engineering and finance leadersJump to the exercise ↓

The situation

The question nobody in the room can answer.

It's the last week of the quarter. Finance has the AI line in front of them: 140 seats, a usage bill that grew faster than the team did, and eleven expense claims for tools nobody bought centrally. Nearly $50,000 in all. The question is short. What did we get for it?

Engineering answers with the numbers it has: how many engineers use the tools each week, and a survey where most of them say they feel faster. All of it is true, and none of it answers the question. Adoption says the tools are being used. It doesn't say what they were used on, or whether that work ever shipped.

That isn't carelessness. It's how AI spend arrives: in three shapes, and none of them carries the answer on its own.

Why it's hard

Three shapes of spend, three different blind spots.

Per person, per month

Seats

Tells you
Who has access, and roughly how often they use it.
Doesn't tell you
What any of it was used on. A seat costs the same whether it built a feature or sat idle.
Per token, per key

Usage

Tells you
How much was consumed, and through which key or account.
Doesn't tell you
Which piece of work it served. A token carries no ticket number.
On someone's own card

Outside procurement

Tells you
Very little, and usually late, when the expense claims arrive.
Doesn't tell you
That it exists at all, until somebody adds it up.

The ladder

Four answers, and each one needs the one before it.

Every question anyone asks about AI spend is one of four. They come in order: you can't say whether spend was worth it until you know what it bought, and you can't know that until you know which work it paid for.

  1. 0

    What did we pay?

    NeedsThe invoice. Everyone has this one.

  2. 1

    Who spent it?

    NeedsEvery seat and every key tied to a team.

  3. 2

    On what?

    NeedsEvery run tied to the piece of work it served, at the time it runs.

  4. 3

    Did it ship?

    NeedsThat work's outcome on the same record as its spend.

  5. 4

    Was it worth it?

    NeedsAn estimate made before the work was approved, to measure against.

The questions

Five questions to ask before your next review.

Put these to whoever owns your AI tools. If the answers take a week of spreadsheets, that tells you where you are on the ladder.

  1. Which team spent it?

    Answer 1

    Seats are billed to people and usage to keys. If two teams share a key, or a platform team holds keys on everyone's behalf, the answer is a guess.

  2. Which piece of work did it pay for?

    Answer 2

    Unless each run is tied to the work item it served while it runs, no report afterwards can put the two back together.

  3. How much went on attempts you threw away?

    Answer 3

    An agent that needs three attempts costs three times as much, and the invoice still shows one line. Retries, abandoned branches and reverted changes are real spend with nothing to show for it.

  4. Did the work it bought actually ship?

    Answer 3

    Spend on a change that was never merged, or was reverted a week later, bought nothing. You only see that when spend and outcome sit on the same record.

  5. Could you have known the cost before you said yes?

    Answer 4

    Without an estimate there's no such thing as an overrun, only a bill. With one, a figure can be high or low, and someone can say why.

Try this on your organisation

How far up the ladder does your spend go?

Enter a typical month and answer four questions about how your spend is recorded today. It takes two minutes, and nothing you type leaves this page.

A typical month
$
$
$
How it's recorded today

These are example answers for one organisation. Replace them with yours.

Add your monthly figures to see where your answers stop.

0 · What did we pay?
1 · Who spent it?
2 · On what?
3 · Did it ship?
4 · Was it worth it?
How this is worked out
  • Seat spend reaches answer 1 if you know which team holds each seat. It never goes further: a seat is paid for by the month, not by the piece of work.
  • Usage spend reaches answer 1 if each team has its own key, or if runs are tagged with their work item.
  • Usage spend reaches answer 2 only when runs are tagged with their work item, answer 3 when you can also see which items shipped, and answer 4 when those items were estimated first.
  • Spend outside procurement stays at 0. Nobody can explain spend they can't see.

This counts what you can explain, not what you wasted. Spend you can't explain may well be buying good work. You just can't show it.

With Wakalix

Spend that's attached to the work while the work runs.

The four answers can't be rebuilt from an invoice afterwards. They have to be recorded as the work happens. That is what Wakalix does, on the AI tools your teams already use.

  1. 1
    Who spent itEvery work item belongs to a team's cycle, so its spend does too, whichever AI tool that team works in.
  2. 2
    On whatEach run is measured from the AI tool's own usage records and attributed to the work item and the agent that did it. It's never an agent's account of itself, and it works the same on a flat-rate subscription.
  3. 3
    Did it shipEvery attempt is counted, and each item carries its outcome (accepted, in review or sent back) right next to what it used.
  4. 4
    Was it worth itEvery item is estimated when it's planned, from the typical cost of each agent's skills, so you approve the cost with the plan and watch the measured figure against it.
Reports › Spend
The spend report for one cycle: AI usage measured, spend against the forecast, work accepted on the first attempt, and a table of each work item with its agent, attempts, forecast, measured usage and outcome.
Product screen · Reports › Spend. Left to right: the work item, the agent, attempts, forecast, measured usage and outcome.

Your AI subscriptions and their bills stay your own. Wakalix never runs the model, and never adds anything of its own to the figure.

Next

Bring your answer to a conversation.

Run the exercise with your finance partner and note where your answers stop. Then tell us. We'll walk you through the same four answers in Wakalix, and what it would take to reach them across your teams.