Inferock Bench - An independent receipt for every LLM API call
by•
Inferock-bench is a local proxy that sits between your app and OpenAI, Anthropic, Gemini, or OpenRouter shaped calls. It captures per-call token usage, failures, and retries, then generates an independent receipt showing what you were billed and how much you're actually overpaying for.


Replies
The bill dispute question is interesting. have you personally managed to get a provider to credit a charge after showing them one of these per call receipts or is that still something you are testing?
Inferock
@jeremy_loomis honest answer: no credit to brag about yet, that's exactly why we threw the question to the community. What we can stand behind today is the receipt itself, knowing which call failed and what it cost, instead of arguing from a monthly total.
Does anyone know of any tools that can work with web-based logins (claud.ai etc)
Inferock
@jay_janarthanan1 we don't touch that layer. inferock-bench works where there's an API key and a baseURL to point somewhere, and web app subscriptions don't expose the per call detail we'd need. If someone has cracked that curious to see it too.
Macaly
auditing your own inference bill is such an obvious gap 🔍 nice one. seeing much overbilling in the wild yet?
Inferock
DROP
Retries are the part I’d put in giant font. ‘$0.04/call’ means nothing if a successful job secretly takes 6 calls. I want cost per accepted outcome.
Inferock
Lancepilot
Congrats on the launch of Inferock Bench.
Love the idea of having an independent receipt for every LLM API call. Transparent usage, billing, and overpayment tracking can be a huge win for teams working with multiple LLM providers.
Inferock