Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
101 entries in this view.
Tool ranking
Ranked by measured growth, normalized across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
llmtrim
MCPMCP server and proxy that compresses LLM prompts, tool output, and replies to cut token cost.
mcpmeasured growthOpen source ↗
Install
claude mcp add llmtrim -- npx @llmtrim/cli222GitHub stars+6 (+2.8 %) - 2UnknownLast commit: unavailable
AI agent token-cost telemetry + 429 prediction. Per-agent attribution, anomaly + routing + quota.
mcpmeasured growthOpen source ↗
Install
claude mcp add openclaw-cost-tracker-mcp -- uvx openclaw-cost-tracker-mcp129downloads over 7 daysstable - 3Active
Reduces AI agent token usage by 90% via context compression and task checkpoint persistence.
mcpmeasured growthOpen source ↗
Install
claude mcp add smart-context-mcp -- npx smart-context-mcp4GitHub starsstable - 4Dormant
Save tokens while coding — your AI agent gets structured code context, not file dumps.
mcpmeasured growthOpen source ↗
Install
claude mcp add codeweave -- npx @codeweave/mcp4GitHub starsstable - 5Dormant
Analyze your MCP setup: token costs, grades, duplicates, and optimization tips
mcpmeasured growthOpen source ↗
Install
claude mcp add mcp-checkup -- npx mcp-checkup1GitHub starsstable - 6Active
Cut AI agent token costs 5-15x — routes only relevant code symbols instead of full files.
mcpmeasured growthOpen source ↗
Install
claude mcp add agent-booster -- uvx agent-booster1GitHub starsstable - 7Active
Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen
mcpmeasured growthOpen source ↗
Install
claude mcp add vision-squeezer -- npx vision-squeezer2GitHub starsstable - 8Active
Token cost math for LLM API calls: verified per-1M-token rates for 69 models, 17 providers.
mcpmeasured growthOpen source ↗
Install
claude mcp add comparedge-llm-cost -- npx @comparedge/llm-cost-mcp2GitHub starsstable - 9Active
Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.
mcpmeasured growthOpen source ↗
Install
claude mcp add compactprompt -- uvx compactprompt0GitHub starsstable - 10Active
MCP server that cuts AI coding agent token usage via framework-aware context optimization
mcpmeasured growthOpen source ↗
Install
claude mcp add ai-optimizer -- npx @ai-optimizer/core0GitHub starsstable - 11Active
Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable
mcpmeasured growthOpen source ↗
Install
claude mcp add token_optimizer -- uvx slimctx1GitHub starsstable - 12Active
rtk
CLICLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
githubmeasured growthOpen source ↗
78 241GitHub stars+697 (+0.90 %) - 13Active
headroom
MCPCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
githubmeasured growthOpen source ↗
68 350GitHub stars+630 (+0.93 %) - 14Active
Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add NodeNestor/claude-rolling-context31GitHub stars+1 (+3.3 %) - 15Active
ratel
MCPContext engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
githubmeasured growthOpen source ↗
433GitHub stars+6 (+1.4 %) - 16Dormant
pith
SkillPith is the hook that makes Claude Code sessions last 3x longer.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add abhisekjha/pith98GitHub starsstable - 17Active
claude-power-skills
Skill20 power-user skills for Claude Code: session memory and consolidation, context compression, multi-agent orchestration, adversarial bug hunting, security review, calibrated estimation and decision archaeology. Drop-in SKILL.md files. Built…
githubmeasured growthOpen source ↗
Install
git clone https://github.com/irfad7/claude-power-skills ~/.claude/skills/claude-power-skills4GitHub starsstable - 18Active
super-token-saver
SkillRestore prior Claude Code AND Codex sessions with zero LLM calls. Cross-tool session handoff, cache-expiry prevention, real cost dashboard. One plugin, both hosts.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add ww-w-ai/super-token-saver31GitHub starsstable - 19Active
claude-code-eco
SkillEco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured hardest on Claude Fable 5 (fable5), deep-studied on Sonnet 5, works on Opus 4.8 too. We publish our…
githubmeasured growthOpen source ↗
Install
/plugin marketplace add sup3x/claude-code-eco33GitHub starsstable - 20Dormant
ppt-report-skills
SkillA Claude Code skill for web report decks — richer ECharts charts, one file per slide (edit one page, save tokens), and complex structure diagrams. 16:9 & vertical, PDF/image export.网页版 PPT 汇报的 Claude Code skill:ECharts 把数据画得更好看、每页一个文件改一页省…
githubmeasured growthOpen source ↗
Install
git clone https://github.com/myunwang/ppt-report-skills ~/.claude/skills/ppt-report-skills81GitHub starsstable - 21Active
voly
AgentControl plane for AI coding agents: route tasks, reduce token spend, run multi-agent workflows, fallback executors, and track cost per task.
githubmeasured growthOpen source ↗
16GitHub starsstable - 22Active
A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).
githubmeasured growthOpen source ↗
Install
/plugin marketplace add Da-ya7/PromptCompressor-skill1GitHub starsstable - 23Dormant
unclog
MCPAudit MCP servers, skills, and CLAUDE.md bloat eating your Claude Code context window.
githubmeasured growthOpen source ↗
35GitHub starsstable - 24Active
basecamp
SkillA production-ready Claude Code setup. Global CLAUDE.md, coding rules, per-project memory, and token optimization.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add janmaaarc/basecamp7GitHub starsstable
Listed without a public metric
Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1UnknownLast commit: unavailable
Semantic context compression and cognitive memory layer for LLM swarms.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://recallmax-mcp.vercel.app/api/mcp—No public metric provided