Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
88 entries in this view.
Tool ranking
Ranked by measured growth, normalized across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
voly
AgentControl plane for AI coding agents: route tasks, reduce token spend, run multi-agent workflows, fallback executors, and track cost per task.
githubmeasured growthOpen source ↗
16GitHub starsstable - 2Active
Statically audits MCP tool surfaces for token cost, schema quality, and design issues.
mcpmeasured growthOpen source ↗
Remote server, to be declared in the MCP configuration
https://mcplint-web.vercel.app/api/mcp0GitHub starsstable - 3Active
Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.
githubmeasured growthOpen source ↗
1GitHub starsstable - 4Dormant
Spiderbrain-V3
SkillSpiderBrain v3 is a multi-platform skill/framework to reduce token usage and AI hallucinations across Claude, Cursor, and other AI tools.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/abhishek-performdigital/Spiderbrain-V3 ~/.claude/skills/Spiderbrain-V365GitHub starsstable - 5Active
permafrost
SkillFreeze Claude Code's prompt prefix so DeepSeek's automatic cache always hits — alignment proxy + coalescing + keepalive, installable as a CC plugin. Measured 64% cheaper on real Claude Code traffic.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add jianzhichun/permafrost21GitHub starsstable - 6Active
Usage-aware status line for Claude Code: session & weekly rate-limit gauges with reset countdowns and a precise context-window meter, plus tokens, git & weather. Template-driven and extensible — set it up by chatting with Claude or via a…
githubmeasured growthOpen source ↗
1GitHub starsstable - 7Active
somtum
SkillLocal-first memory and prompt-cache layer for Claude Code
githubmeasured growthOpen source ↗
Install
/plugin marketplace add riz007/somtum7GitHub starsstable - 8Dormant
bonsai
SkillDrop-in Cursor rules to cut AI response tokens (no filler/hedging, preserve code + technical strings). Includes local, reproducible benchmarks.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/marcomondelli/bonsai ~/.claude/skills/bonsai0GitHub starsstable - 9Active
Token cost math for LLM API calls: verified per-1M-token rates for 69 models, 17 providers.
mcpmeasured growthOpen source ↗
Install
claude mcp add comparedge-llm-cost -- npx @comparedge/llm-cost-mcp2GitHub starsstable - 10Active
A MCP that turns PDFs into structured text, tables, figures, metadata, annotations, and page-level evidence for selective retrieval by MCP-compatible AI agents. Improves the information quality AI agents receive while also reducing token…
githubmeasured growthOpen source ↗
17GitHub starsstable - 11Active
Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen
mcpmeasured growthOpen source ↗
Install
claude mcp add vision-squeezer -- npx vision-squeezer2GitHub starsstable - 12Active
webfetch
SkillOwn your LLM's web search: a local search->fetch->rank pipeline that replaces hosted web-search tools. Measured: matches hosted accuracy at 66% lower cost and up to 88% fewer tokens, plus a precision-tuned semantic caching with…
githubmeasured growthOpen source ↗
Install
git clone https://github.com/firish/webfetch ~/.claude/skills/webfetch55GitHub starsstable - 13Active
mcp-recall
Skillmcp-recall compresses MCP tool outputs (94 KB → 3.5 KB · 96%) and stores full results in SQLite for retrieval — up to 30x more tool calls per session for heavy MCP workloads.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add sakebomb/mcp-recall8GitHub starsstable - 14Active
Context GC for LLM agents: offload large tool outputs and recall them to save tokens.
mcpmeasured growthOpen source ↗
Install
claude mcp add lethe -- uvx lethe-llm-context0GitHub starsstable - 15Active
cycgraph
AgentAgent context compressor and memory persistent orchestration
githubmeasured growthOpen source ↗
2GitHub starsstable - 16Active
lazy-cat
SkillClaude Code skills for developers who code like cats — never more effort than the problem requires.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/albertobarnabo/lazy-cat ~/.claude/skills/lazy-cat50GitHub starsstable - 17Dormant
DAG-based Lossless Context Management for Claude Code. Every message preserved, summaries cascade as a DAG, patterns extracted via Lossless Dream — full recall and reflection across sessions.
githubmeasured growthOpen source ↗
15GitHub starsstable - 18Active
catalyst
SkillToken-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add coalesce-labs/catalyst20GitHub starsstable - 19Active
compactor-skill
OtherClaude Code skill: reduce token usage 60-80% — targeted reads, model switching, context management. One-line install.
githubmeasured growthOpen source ↗
2GitHub starsstable - 20Active
Verifies token/cost savings claimed by AI context-reduction proxies via MCP tools.
mcpmeasured growthOpen source ↗
Install
claude mcp add tokentrust -- uvx tokentrust-cli0GitHub starsstable - 21Active
Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable
mcpmeasured growthOpen source ↗
Install
claude mcp add token_optimizer -- uvx slimctx1GitHub starsstable - 22Active
claude-lean-skill
OtherTurn your strongest model into a conductor: route each step to the cheapest capable model, and escalate only load-bearing pieces to a pricier reviewer that catches pitfalls early. Subscription, API, or hybrid.
githubmeasured growthOpen source ↗
2GitHub starsstable - 23Dormant
token-reducer
Skill⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add Madhan230205/token-reducer44GitHub starsstable - 24Active
basecamp
SkillA production-ready Claude Code setup. Global CLAUDE.md, coding rules, per-project memory, and token optimization.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add janmaaarc/basecamp7GitHub starsstable - 25Active
SpanishCaveMan
SkillClaude Code skill: caveman-style token efficiency, spoken with a real Spanish regional accent
githubmeasured growthOpen source ↗
Install
/plugin marketplace add Mun1to/SpanishCaveMan2GitHub starsstable