Compare
Excellent vs Tusk: one writes the tests, one proves the attempt
Tusk writes the tests. Excellent proves the attempt — and can take Tusk's tests as the evidence.
Tusk
Tusk in one line
Tusk calls itself a verification layer for coding agents. It writes tests from production traffic and business context, runs them from the CLI, repeats until they pass, and repairs tests that have gone stale.
- Best for
- Teams whose real gap is coverage: the suite is thin, and they want tests generated from real traffic and run against every agent-authored change.
- Limits
- A test suite is not a record. Tusk does not hold the agent's attempts, the claim the agent made, or the non-test evidence — build, browser, CI — that sits behind a finished piece of work.
- Stronger than us at
- Creating tests, which Excellent does not do at all. Tusk generates them from production traffic and business context; Excellent runs the checks you already have. Tusk also publishes that its tests catch real-world regressions in 43% of PRs — we have published no comparable number, and we will not invent one.
- Price 2026-09-23
- $50 / active dev / mo (Team)
Excellent
Excellent in one line
A verification system for agentic tasks — it records what the agent attempted, runs your checks against the claim, and binds the result to evidence you can open.
- Best for
- Teams that want agent work checked before they trust or promote it.
- Limits
- There is no hosted option. You run it on your machines. Verification evidence stays on the machine that produced it — it does not sync to teammates yet. No SSO, and no Mac app has shipped. It does not write tests, and it does not review your diff.
- Stronger than them at
- Keeping the attempt record, running your own checks as the oracle, and trialling prompt, model and tool changes on held-out work from your repository.
- Price 2026-09-23
- Free to install and run (Consulting from $10,000 / month)
Head to head
The capabilities that decide it
Capability by capability — Tusk on the left, Excellent on the right. Prices are the vendor's own, read on 2026-09-23.
| Capability | Tusk | Excellent |
|---|---|---|
Writes new tests for you | Yes — generated from production traffic and business context | |
Runs tests as part of the change | ||
Records each agent attempt and the claim it made | ||
Evidence beyond tests | Tests | Tests, builds, browser checks and your existing GitHub Actions or GitLab CI |
Result bound to a receipt you can re-open | ||
Tests prompt, model and tool changes on held-out work | ||
Where it runs | Tests are generated by the service and run from the CLI | Local CLI and MCP; checks run on your machines |
Published price (checked 2026-09-23) | Free $0 (no seat minimum); Team $50 per active developer per month (no seat minimum); Enterprise custom, 200-seat minimum | Free to install and run. Consulting starts at $10,000 / month. |
Together
They are not mutually exclusive
These two stack rather than collide. Tusk's generated tests are exactly the kind of oracle Excellent wants: run them as a check on an attempt and their pass or fail becomes part of the evidence behind the result. If you already run Tusk, Excellent does not replace it — it records what happened when those tests ran, and what the agent claimed before they did.
Verdict
When to pick which
If the problem is that you have no tests, start with Tusk — it makes them, and we do not. If the problem is that you cannot tell whether the agent did what it said it did, start with Excellent. Teams running agents at volume tend to end up wanting both: Tusk supplies the oracle, Excellent keeps the record of what the oracle said.
Pick Tusk if
Teams whose real gap is coverage: the suite is thin, and they want tests generated from real traffic and run against every agent-authored change.
Add Excellent if
You want agent work captured as attempts, checked against evidence you can open, and promoted only when the result is defensible to someone outside the team.
Still shopping? See the full list of Tusk alternatives →
More matchups
See Excellent against the rest of the category
Trust the evidence, not the agent's confidence
Install Excellent, run it beside the agent you already use, and inspect what the checks actually said.