Local MCP server and CLI for tracking Claude, GPT, and Gemini token costs and identifying optimization savings.
Copy the AI prompt to install this server into Claude Code, Cursor, or another agent — or use 1-click editor setup below.
We haven't yet run this listing's install command through our automated sandbox check. This isn't a red flag — we're steadily working through the catalog.
💡 Paste the JSON block into your client's configuration file under mcpServers, then restart the application.
Inspect callable tools, capabilities, and parameters exposed to AI agents by TokenBurnRate.
log_usageLog an API call — auto-calculates cost
get_summarySummary for today / week / month / all
get_hintsRanked optimization hints with $ savings
get_hint_detailDeep-dive on a specific hint
set_budgetSet a daily / weekly / monthly spend limit
list_sessionsSessions ranked by cost
See where your AI tokens go — and how to spend less of them.
TokenBurnRate is an MCP server + CLI that logs every Claude / GPT / Gemini API call locally, shows you a cost dashboard in your terminal, and tells you exactly how to reduce that cost.
Or run without installing:
Or add manually to ~/Library/Application Support/Claude/claude_desktop_config.json:
Restart Claude Desktop. Done.
| Command | Description |
|---|---|
token-tracker report | Full 7-day dashboard |
token-tracker report --period month | Monthly report |
token-tracker today | Today only |
token-tracker hints | Optimization hints ranked by $ saving |
token-tracker hint <id> | Deep-dive on one hint |
token-tracker status | One-line: cost · cache % · top hint |
token-tracker budget | Budget gauges |
token-tracker models | Pricing table for all models |
token-tracker export > out.csv | Raw CSV export |
| Tool | Description |
|---|---|
log_usage | Log an API call — auto-calculates cost |
get_summary | Summary for today / week / month / all |
get_hints | Ranked optimization hints with $ savings |
get_hint_detail | Deep-dive on a specific hint |
set_budget | Set a daily / weekly / monthly spend limit |
list_sessions | Sessions ranked by cost |
list_models | Pricing table |
export_csv | CSV dump |
8 deterministic rules — no LLM calls, runs instantly on your local data:
| Hint | Triggers when |
|---|---|
cache-utilization | Cache hit rate < 30% |
model-swap-testing | Test gen running on Sonnet / Opus |
model-swap-debug | Debugging on Opus |
verbose-outputs | Output / input ratio > 0.35 |
session-spike | Any session costs 3× your average |
context-bloat | Avg tokens / call > 8K |
retry-loops | Sessions with 30+ high-token calls |
single-model-dependency | 100% traffic on one expensive model |
Each hint includes severity · evidence · recommended action · estimated monthly saving.
All data stored at ~/.token-tracker/usage.db (SQLite).
Nothing leaves your machine. No telemetry, no account required.
MIT © 2026 Nikhil T
Factual signals from GitHub, npm, and our automated checks — not a rating.
No reviews yet — be the first to share how this listing worked for you.
Showcase your server listing on GitHub or your project documentation. Embed this dynamic SVG badge to highlight official listing status and live engagement.
[](https://allmcps.com/mcp/nikhilnt1234-tokenburnrate)<a href="https://allmcps.com/mcp/nikhilnt1234-tokenburnrate"><img src="https://allmcps.com/api/badge/nikhilnt1234-tokenburnrate?style=directory" alt="TokenBurnRate on AllMCPs" /></a>