When your agents get complex
Simulation-based AI agent testing and evaluation that turns unpredictable agents into reliable production systems.
This is the 4th launch from LangWatch. View more
Claude Code usage tracking by LangWatch
Launched this week
Track Claude Code usage: cost, cache, session replay. Run `npx langwatch claude` once. Every session gets cost with cache reads/writes as separate token classes, every bash and MCP call as a span, theoretical vs billed for your Max plan, and a full terminal replay in the UI. Works for Codex too.









Free
Launch Team




LangWatch
@imtiaj_ahmad That's exactly where we're headed, this is the groundwork for it, so what kinds of suggestions would actually be useful to you?
Lancepilot
Can the platform can identify inefficient prompting patterns across multiple sessions?
Interesting. What about the agent-to-agent usage tracking?
This feels like the equivalent of DevTools, but for AI-assisted development. Congratulations on the launch @manouk_dr
LangWatch
@manouk_dr @sidraarifali 💯
LiveDemo
Congrats on the launch!
Token tracking is becoming crutial with everyday usage of Claude... maybe eventually it will turn out that junior developers are less expensive after all :)
P.S:
I have built an interactive livedemo for you, feel free to check it out
https://app.livedemo.ai/livedemos/6a6afa27cc86a442a8fd9d17
LangWatch
@gapostolov Thanks quick, thanks for sharing that, we'll have a look at it at the same time, please feel free to start using and let us know how it helped you!
Pazi
Knowing exactly what each Claude Code session costs is a game changer. Congrats on the launch, this is awesome! 🚀
LangWatch
@zvonimir_sabljic1 thanks, this is just the beginning!
LangWatch
@zvonimir_sabljic1 thanks 🙏 let us know once you’ve tried
Triforce Todos
LangWatch
@abod_rehman whoop, let us know once you've set it up, there is so much more to explore! thanks for ssharing :)