Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
101 entries in this view.
Tool ranking
Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).
githubmeasured growthOpen source ↗
Install
/plugin marketplace add Da-ya7/PromptCompressor-skill1GitHub starsstable - 2Active
Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.
githubmeasured growthOpen source ↗
1GitHub starsstable - 3Dormant
Analyze your MCP setup: token costs, grades, duplicates, and optimization tips
mcpmeasured growthOpen source ↗
Install
claude mcp add mcp-checkup -- npx mcp-checkup1GitHub starsstable - 4Active
Cut AI agent token costs 5-15x — routes only relevant code symbols instead of full files.
mcpmeasured growthOpen source ↗
Install
claude mcp add agent-booster -- uvx agent-booster1GitHub starsstable - 5Active
Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable
mcpmeasured growthOpen source ↗
Install
claude mcp add token_optimizer -- uvx slimctx1GitHub starsstable - 6Dormant
nekte
OtherAgent-to-Agent coordination that doesn't waste your context window. Token-efficient protocol with progressive discovery, zero-schema invocation, gRPC transport, and task lifecycle management.
githubmeasured growthOpen source ↗
1GitHub starsstable - 7Active
Usage-aware status line for Claude Code: session & weekly rate-limit gauges with reset countdowns and a precise context-window meter, plus tokens, git & weather. Template-driven and extensible — set it up by chatting with Claude or via a…
githubmeasured growthOpen source ↗
1GitHub starsstable - 8Dormant
contextforge
SkillAgent context gate for Codex, Claude Code, Copilot, MCP, Cursor, Cline, Gemini and Windsurf repos
githubmeasured growthOpen source ↗
Install
git clone https://github.com/grnbtqdbyx-create/contextforge ~/.claude/skills/contextforge1GitHub starsstable - 9Active
distil
MCPContext optimization for LLM agents: tool registry, result masking, budgeting, compaction.
mcpmeasured growthOpen source ↗
Install
claude mcp add distil -- npx @munhq/distil1GitHub starsstable - 10Active
sidequest
SkillCut Claude Code token costs by delegating bulk LLM work to cheap models. 600+ models via NanoGPT, OpenRouter, Groq or local Ollama — batched, resumable, with real per-call cost tracking.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add bicced/sidequest1GitHub starsstable - 11Dormant
bonsai
SkillDrop-in Cursor rules to cut AI response tokens (no filler/hedging, preserve code + technical strings). Includes local, reproducible benchmarks.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/marcomondelli/bonsai ~/.claude/skills/bonsai0GitHub starsstable - 12Active
Statically audits MCP tool surfaces for token cost, schema quality, and design issues.
mcpmeasured growthOpen source ↗
Remote server, to be declared in the MCP configuration
https://mcplint-web.vercel.app/api/mcp0GitHub starsstable - 13Active
Verifies token/cost savings claimed by AI context-reduction proxies via MCP tools.
mcpmeasured growthOpen source ↗
Install
claude mcp add tokentrust -- uvx tokentrust-cli0GitHub starsstable - 14Active
Context GC for LLM agents: offload large tool outputs and recall them to save tokens.
mcpmeasured growthOpen source ↗
Install
claude mcp add lethe -- uvx lethe-llm-context0GitHub starsstable - 15Active
MCP server that cuts AI coding agent token usage via framework-aware context optimization
mcpmeasured growthOpen source ↗
Install
claude mcp add ai-optimizer -- npx @ai-optimizer/core0GitHub starsstable - 16Active
Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.
mcpmeasured growthOpen source ↗
Install
claude mcp add compactprompt -- uvx compactprompt0GitHub starsstable - 17Active
dao-code
OtherOpen-source TypeScript terminal coding agent for DeepSeek-V4 — builds on DeepSeek's strong price-performance and ultra-cheap cache pricing, engineering byte-stable prefixes and cache-reusing forks so cross-session memory and a continuous…
githubestimated momentumOpen source ↗
1 254GitHub stars— - 18Active
quiv
SkillThe Agent Knowledge Kit & Architecture System — 10–32x token-efficient architecture discovery for AI coding agents
githubestimated momentumOpen source ↗
Install
git clone https://github.com/chama-x/quiv ~/.claude/skills/quiv3GitHub stars— - 19Active
Major-AI-Skills
SkillInstallable agentic skills (SKILL.md) for Claude Code, Cursor, Codex, Gemini & Antigravity — 292+ product, efficiency, and common-sense skills. npx major-ai-skills
githubestimated momentumOpen source ↗
Install
git clone https://github.com/alivirgo/Major-AI-Skills ~/.claude/skills/Major-AI-Skills1GitHub stars— - 20Active
urusilla
AgentExperimental no-install semantic language for AI agents: typed action/state, safe NL/JSON fallback, and falsifiable public evals. Try the one-file probe.
githubestimated momentumOpen source ↗
1GitHub stars—
Learning resources
Ranked by measured growth, normalized across sources. These resources remain available separately and do not take part in the main tool ranking.
- 1ActiveOpen source ↗
SUMMARY-md
SkillA cheat sheet that stops AI agents from scanning your whole repo and saves you tokens.
githubInstall
Resourcemeasured growthgit clone https://github.com/toshon-jennings/SUMMARY-md ~/.claude/skills/SUMMARY-md0GitHub starsstable
Listed without a public metric
Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1UnknownLast commit: unavailable
Semantic context compression and cognitive memory layer for LLM swarms.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://recallmax-mcp.vercel.app/api/mcp—No public metric provided - 2UnknownLast commit: unavailable
sqz
MCPPre-injection context compression for coding agents. Zero LLM calls, zero telemetry, offline-safe.
mcpmetric not publishedOpen source ↗
—No public metric provided - 3UnknownLast commit: unavailable
MCP server for Skim code transformation. Compresses code 60-95% for LLM context optimization.
mcpmetric not publishedOpen source ↗
Install
claude mcp add skim-mcp-server -- npx skim-mcp-server—No public metric provided - 4UnknownLast commit: unavailable
LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://cache.api.ainode.tech/mcp—No public metric provided