How to tell if your AI really finished.

By the Olomni team · Updated

Short answer

Check whether a test (tests, build or lint) ran after the last change and passed. If it wasn’t tested, it isn’t finished even if the AI says “done”; if the test fails, it isn’t either. Olo automates this on your machine: green if a test passed after the last change, amber if it wasn’t tested, red if it fails or touched a path it shouldn’t.

Step by step

  1. Check what really changed with git status and git diff, not with the AI’s summary.
  2. Run the project’s tests and build after the last change (or ask the AI to and read the result).
  3. Check it didn’t touch sensitive paths: migrations, configuration, deployment, key files.
  4. Be wary if it repeats the same error several times: ask for another approach or hand the work to another AI.
  5. With Olo, write once in the project contract which commands show it’s ready; Olo runs them on your machine and marks the turn green, amber or red.

Why an AI’s “done” isn’t enough

An agent describes what it thinks it did. It may have edited files without running the tests, or run a test before its last change. The useful evidence is what was observed on your machine after the last change.

What each color means in Olo

  • Green: an observed test passed after the last change.
  • Amber: nothing was tested after the last change.
  • Red: the test fails, it touched an untouchable path, or the AI is looping on the same failure.

What it doesn’t do

It doesn’t replace code review or guarantee there are no bugs: it only checks what your tests cover. Olo only runs the commands you wrote in the project contract.

Frequently asked questions

Does it work with any AI?

The verdict is computed from what Olo observes on your machine, through each AI’s integration. Claude Code is the best-tested; the others work with limits.

Does Olo run code from the AI?

No. It only runs the commands you wrote in the contract, from a closed list of command types (tests, build, lint).

What if I have no tests?

The turn stays amber: no tests, no evidence. A good reason to add at least one.

One control for all your AIs.

Free during the beta, for Mac (Windows in beta). Ready in three minutes.

Start free