Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
96 entries in this view.
Tool ranking
Ranked by normalized popularity across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.
githubmeasured growthOpen source ↗
1GitHub starsstable - 2Active
Usage-aware status line for Claude Code: session & weekly rate-limit gauges with reset countdowns and a precise context-window meter, plus tokens, git & weather. Template-driven and extensible — set it up by chatting with Claude or via a…
githubmeasured growthOpen source ↗
1GitHub starsstable - 3Active
Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable
mcpmeasured growthOpen source ↗
Install
claude mcp add token_optimizer -- uvx slimctx1GitHub starsstable - 4Active
Cut AI agent token costs 5-15x — routes only relevant code symbols instead of full files.
mcpmeasured growthOpen source ↗
Install
claude mcp add agent-booster -- uvx agent-booster1GitHub starsstable - 5Dormant
Analyze your MCP setup: token costs, grades, duplicates, and optimization tips
mcpmeasured growthOpen source ↗
Install
claude mcp add mcp-checkup -- npx mcp-checkup1GitHub starsstable - 6Dormant
contextforge
SkillAgent context gate for Codex, Claude Code, Copilot, MCP, Cursor, Cline, Gemini and Windsurf repos
githubmeasured growthOpen source ↗
Install
git clone https://github.com/grnbtqdbyx-create/contextforge ~/.claude/skills/contextforge1GitHub starsstable - 7Active
A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).
githubmeasured growthOpen source ↗
Install
/plugin marketplace add Da-ya7/PromptCompressor-skill1GitHub starsstable - 8Active
distil
MCPContext optimization for LLM agents: tool registry, result masking, budgeting, compaction.
mcpmeasured growthOpen source ↗
Install
claude mcp add distil -- npx @munhq/distil1GitHub starsstable - 9Active
sidequest
SkillCut Claude Code token costs by delegating bulk LLM work to cheap models. 600+ models via NanoGPT, OpenRouter, Groq or local Ollama — batched, resumable, with real per-call cost tracking.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add bicced/sidequest1GitHub starsstable - 10Active
Statically audits MCP tool surfaces for token cost, schema quality, and design issues.
mcpmeasured growthOpen source ↗
Remote server, to be declared in the MCP configuration
https://mcplint-web.vercel.app/api/mcp0GitHub starsstable - 11Dormant
bonsai
SkillDrop-in Cursor rules to cut AI response tokens (no filler/hedging, preserve code + technical strings). Includes local, reproducible benchmarks.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/marcomondelli/bonsai ~/.claude/skills/bonsai0GitHub starsstable - 12Active
Context GC for LLM agents: offload large tool outputs and recall them to save tokens.
mcpmeasured growthOpen source ↗
Install
claude mcp add lethe -- uvx lethe-llm-context0GitHub starsstable - 13Active
Verifies token/cost savings claimed by AI context-reduction proxies via MCP tools.
mcpmeasured growthOpen source ↗
Install
claude mcp add tokentrust -- uvx tokentrust-cli0GitHub starsstable - 14Active
MCP server that cuts AI coding agent token usage via framework-aware context optimization
mcpmeasured growthOpen source ↗
Install
claude mcp add ai-optimizer -- npx @ai-optimizer/core0GitHub starsstable - 15Active
Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.
mcpmeasured growthOpen source ↗
Install
claude mcp add compactprompt -- uvx compactprompt0GitHub starsstable
Learning resources
Ranked by normalized popularity across sources. These resources remain available separately and do not take part in the main tool ranking.
- 1ActiveOpen source ↗
SUMMARY-md
SkillA cheat sheet that stops AI agents from scanning your whole repo and saves you tokens.
githubInstall
Resourcemeasured growthgit clone https://github.com/toshon-jennings/SUMMARY-md ~/.claude/skills/SUMMARY-md0GitHub starsstable
Listed without a public metric
Ranked by normalized popularity across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1UnknownLast commit: unavailable
Semantic context compression and cognitive memory layer for LLM swarms.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://recallmax-mcp.vercel.app/api/mcp—No public metric provided - 2UnknownLast commit: unavailable
MCP server for Skim code transformation. Compresses code 60-95% for LLM context optimization.
mcpmetric not publishedOpen source ↗
Install
claude mcp add skim-mcp-server -- npx skim-mcp-server—No public metric provided - 3UnknownLast commit: unavailable
Cut LLM token costs: count tokens, estimate cost, slim prompts, and pick the cheapest capable model.
mcpmetric not publishedOpen source ↗
Install
claude mcp add mcp-token-optimizer -- npx mcp-token-optimizer—No public metric provided - 4UnknownLast commit: unavailable
LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://cache.api.ainode.tech/mcp—No public metric provided - 5UnknownLast commit: unavailable
sqz
MCPPre-injection context compression for coding agents. Zero LLM calls, zero telemetry, offline-safe.
mcpmetric not publishedOpen source ↗
—No public metric provided