agentburn
Socialpranker/agentburn
Local profiler that shows which 5-hour window burned your usage, and where tokens went
agentburn is a local, zero-dependency usage profiler for Claude Code, OpenClaw, Hermes Agent and other coding agents that identifies which rolling usage window or which hours actually drove your cost, instead of only a running total.
What agentburn does
Subscription users hit a rolling usage window, not a flat daily cap, so agentburn surfaces the peak 5-hour window against the median of your own active windows, what filled it by model, by source (you, subagents or scheduled work) and by kind (cache reads, cache writes or output), and measures that peak against your own recorded cut-off moments rather than inventing a threshold Anthropic has not published. Per-token payers, including those running Hermes, instead get a straight cost view: where the money went overnight, which loops or retries drove it, and ready config fixes.
Everything runs locally and read-only against your assistant's own logs, with zero external dependencies and nothing leaving the machine; uvx agentburn runs it directly, and a browser demo requires no install at all. Codex CLI users additionally get the provider's own rate_limits.used_percent reading paired against agentburn's local weighted-usage estimate for the same window.
Key features
- Peak-vs-typical 5-hour usage window comparison, not just a running total
- Cost-by-source breakdown (direct use, subagents, scheduled jobs) for per-token billing
- Measured usage ceiling built from your own recorded cut-off moments, never invented
- Zero dependencies, fully local and read-only against existing agent logs
- Separate limits and context views for subscription ceilings and long-context cost
When to use it
- Finding which 5-hour window actually caused a Claude Code or Hermes rate-limit cutoff
- Seeing how much of an overnight bill came from cron jobs versus subagents versus direct chat
- Deciding whether a /clear at a given context size would have saved meaningful cost
Who it is for: Anyone running Claude Code, Hermes Agent, OpenClaw or similar agents who wants to know exactly which window or source drove their usage, not just a total.
How it fits with Hermes Agent
agentburn lists Hermes Agent among its supported agents and treats per-token Hermes usage the same way it treats OpenClaw: as a cost problem to localize by source and time, rather than a window-limit problem.
How to install agentburn
These commands are copied from the project's README. Check the repository for the latest steps before you run them.
uvx agentburn FAQ
What is agentburn?
agentburn is a local, zero-dependency profiler that identifies which specific usage window or source, cron jobs, subagents or direct chat, actually drove an AI agent's cost or rate-limit cutoff.
Does agentburn work with Hermes Agent?
Yes, it lists Hermes Agent among its supported agents and reads its local logs the same normalized way it reads Claude Code, Codex CLI, Gemini CLI, opencode and OpenClaw logs.
How do I run agentburn?
Run uvx agentburn directly for the cost view, uvx agentburn limits for the subscription rolling-window view, or try it with no install at all through the browser demo linked in the README.
Similar models for Hermes Agent
All modelsContext reuse layer that cuts redundant tokens before they reach the model
dyedd LensSelf-hosted multi-protocol LLM gateway with one base URL across many providers
SouthpawIN TurbofitAdaptive local-inference provider that fits the best sustainable model to your hardware
tuxevil tuxevil-rotatorOpenAI-compatible proxy that rotates free-tier LLM accounts with per-model quota routing
piyush-tyagi-13 llm-keypoolFree-tier LLM API key pool with rotation, 429 cooldowns and a local OpenAI-compatible proxy
InfiniteWhispers HermesAgent-MultiModelFully local multi-model setup for Hermes Agent with Ollama, Mixture-of-Agents routing and a tuning guide
Related guides: How to install Hermes Agent