Knowledge Base › Measuring AI

Measuring AI

Measure AI ROI without fooling yourself

How-toLast reviewed Jul 8, 20267 min read

In short

AI ROI is easy to say and easy to fake. Start with one number your business already watches. Keep it narrow. Measure it before the build, then again after on fresh work a human checks. If you can't show the sample size and caveats, you can't claim the return.

Most "AI ROI" numbers fall apart when you poke them. They're measured on demos or cherry-picked examples. Bad start. Some compare against a "before" nobody recorded, like last quarter's support backlog. Use a method your team can rerun without trusting anyone's word.

  1. Pick one number your business already cares about

    Not five goals. One. Use hours spent on a task, cost per case, error rate, or time-to-answer. If your ops lead wouldn't care, it's the wrong number.

  2. Set a clean baseline before you build

    "Before vs after" means nothing without a real before. Measure today's work and write it down. This is the skipped step that makes ROI believable, like timing 50 intake calls before automation.

  3. Measure after on fresh, human-checked data

    Test on work the system hasn't seen, not the examples used to tune it. Have a person check a slice by hand. A vendor's private score isn't evidence.

  4. Write down the sample size and the caveats

    A win on 400 cases means something. A win on 4 doesn't. Note the sample size, who checked it, and what the number leaves out.

  5. Turn the number into money or time

    Translate it into plain math: minutes saved per case × cases per month × the cost of that time. That's the return a budget owner can feel.

The vanity metrics that fake ROI

Some numbers look serious and still don't tell you much. Watch these. They're the usual suspects.

  • Model accuracy on a clean test set. Your real Tuesday inputs are messy, and a tidy benchmark score often won't survive them.
  • Benchmark wins. "Beats model X on task Y" is a press release. Nobody's afternoon got shorter because of it.
  • Usage and engagement. People clicking the AI button doesn't mean the work got faster. Measure the invoice, ticket, or call.
  • A single amazing example. One great output proves it can be right, not that it usually is. Measure a sample, not a screenshot.

The honesty test

Before you start, write down what result would make you call the project a failure. If you can't name that number, you're not measuring ROI. You're writing a story that can only end in success.

Why the method is the point

Anyone can show you a number. The method is what matters. You need a metric the business cares about, a real baseline, and a fresh human-checked sample. Put the caveats in plain sight. Do that, and the ROI has weight. Skip it, and you've got theater with a decimal point.

Apply this

DPR agrees the one honest number with you before any build. We set the baseline and publish how it's measured. You don't have to take the return on faith.

See how a build works Talk to us