Latten
Latten for Developers · Cost

Cut your LLM bill.
Keep the speed.

Your token spend keeps climbing — more agents, bigger context, more calls — and it's hard to see where the money goes. Latten finds the spend you can recover, and shows you the private data headed into those calls — down to the exact call site. Keep the throttle where it is; we point at the money on the table, and at what's leaving that shouldn't.

npm install @latten/cost See pricing ↓

Free, local, no account. Nothing leaves your machine.

See what you can recover

Point it at a coding-agent session and it answers in dollars — recoverable, never "wasted":

Latten spotted about $0.42 you could recover across 6 turns (oversized context). Want the fix?

That number is the point. Most sessions have one — the question is how big, and whether you act on it.

How it reduces LLM costs

It detects the common, expensive token-waste patterns on-device — each with a copy-paste fix:

Redundant round-trips
The agent re-polls a tool whose result didn't change.
Stuck loops
Repeated output that isn't making progress.
Oversized context
Large prompts paying full price every turn instead of cached.
Uncached repeated context
The same big context re-sent, uncached, call after call.
Missing env bootstrap
Early turns burned re-discovering the environment.

Local-first and values-free: it works on token counts, costs, and equality-only hashes. Your prompts and code never leave the machine.

Two halves of one bill

Cost is what your AI spends. Your data is what it can cost you. Latten watches both: @latten/cost finds the spend you can recover, and @latten/redact detects PII and secrets leaking into your LLM calls — on-device, types and counts, never values. Same local-first engine, pointed at the other risk.

See both packages → · How your data stays safe →

Reduce costs in your production app

Free

Keep your own coding-agent spend in view. On your machine, forever.

Paid

See and cut what every crossing in your production app costs, across your whole team. Same engine, pointed at your app.

Pricing

Flat and predictable. We never charge by the call — metering your spend would mean charging you more to be more efficient, which is backwards. Your first week is everything, free. After that, knowing where you stand stays free; understanding and acting is paid.

Free

$0

Your coding agent, forever.

  • On-device recoverable-spend detection — npm, no account
  • Console mirror: live spend, PII counts, one spotlight finding
Install free
Recommended

Solo

$49 / mo

Protect one production app.

  • Everything in Free
  • The full graph — every finding ranked by $, fix prompts, verified savings
Get Solo →

Team

$199 flat

One app, your whole team.

  • Everything in Solo · up to 5 seats
  • Audit exports + weekly digest
Get Team →

Enterprise for larger orgs — talk to us.

Honest about what it is

Detection isn't magic — we catch the patterns we know, and we label fidelity plainly (measured vs estimated). The savings ledger only ever counts fixes you actually applied. We'd rather under-claim and keep your trust than inflate a number.

Spend big on AI. Just keep the bill in view.

npm install @latten/cost Get Solo →