Cut your agent bill by fixing what's burning it.

Mutagent clusters thousands of your traces and reads the ones that matter. It finds the loop burning your tokens and drafts the fix as a PR you approve.

Free, runs locally in your own coding agent. Nothing leaves your machine.

Integrates with every observability platform and agent framework

Vercel AI SDK
OpenAI
LangChain
LangGraph
Mastra
Langfuse
LangSmith
Braintrust
Phoenix
Datadog
Claude Code
Cursor
Codex
OpenCode
Vercel AI SDK
OpenAI
LangChain
LangGraph
Mastra
Langfuse
LangSmith
Braintrust
Phoenix
Datadog
Claude Code
Cursor
Codex
OpenCode
Sound familiar?

You see the symptom. Mutagent names the cause.

At thousands of traces, the cause is buried. Mutagent clusters them and pins it.

Your bill is one monthly total, no idea which path did it
Every run clustered by cost, the worst trajectories ranked
It loops on the same call and burns tokens for hours
The retry loop, with the cache fix as a PR you approve
A five-step run costs far more than five calls
Context re-sent every step, and exactly where to trim it
It spends 50k tokens on work that needed 5k
The dead-end exploration, named and priced
See it run

A real 66-second diagnosis

Point it at your traces, then watch it cluster, find the root cause, and draft the fix.

From a wall of traces to a named root cause

Run it in your own coding agent, and it turns your whole trace history into a short list of named, located causes. Scroll through the steps.

  1. 01

    Run it

  2. 02

    Scan every trace

  3. 03

    Sort signal from noise

  4. 04

    Group and fan out

  5. 05

    Walk to the root cause

  6. 06

    Name it and propose the fix

No platform switch

Works with your stack. Fixes your stack.

Mutagent reads the traces wherever they live and lands the fix on whatever your agent runs on.

Sources in
Langfuse
OpenTelemetry
Datadog
{ }Raw JSONL
Claude Code logs
Codex logs
MUTAGENT/diagnostics
Claude CodeCodexCursorHermes
Targets out
Claude skills & subagents
OpenCode
Codex agents
Vercel AI SDK
Mastra & LangGraph
Deep Agents
</>Prompts
Proof

What it found

89%
of one agent's cost was a single loop
$12,490/mo
spend on a customer agent we analyzed
60%
of that agent's spend was waste
~33% off
on that one agent, month one, no model change

From one recruiting-AI team teardown: 7,524 agent executions across a 24-hour production window on Langfuse, March 2026. Anonymized at their request, the numbers are theirs. The 89% is from pointing Mutagent at our own agent.

A diagnosis you can't reproduce is an opinion with a progress bar.

Questions

Is it really free?

Yes. You install Mutagent into your own coding agent and run the diagnosis on your own traces. No card, no gate.

Does my data leave my machine?

No. It runs in your own environment, through your own coding agent. Your traces stay where they are.

What do I need to run it?

Traces in Langfuse, OpenTelemetry, raw JSONL, or your Claude Code or Codex session logs, plus a coding agent like Claude Code, Codex, Cursor, or OpenCode.

What do I get back?

A ranked list of root causes, each named what, why, and where with cited evidence, and the prompt or guard fix as a PR you approve.

Will it change my code on its own?

No. It proposes; you own the judgment. It lands a change only after you approve it, as a PR on an isolated branch.

What happens on the call?

We walk through your own report together, show you the fixes it would ship, and scope what the full engineer would catch across your stack. No deck.

Find the leak before next month's bill.

Run the diagnosis on your own agent, free. See the first finding, then book a call to see the full engineer.