Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

101 entries in this view.

Tool ranking

Ranked by measured growth, normalized across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1

    MCP
    Active

    MCP server and proxy that compresses LLM prompts, tool output, and replies to cut token cost.

    mcpmeasured growthOpen source ↗

    Install claude mcp add llmtrim -- npx @llmtrim/cli

    222GitHub stars+6 (+2.8 %)
  2. 2
    UnknownLast commit: unavailable

    AI agent token-cost telemetry + 429 prediction. Per-agent attribution, anomaly + routing + quota.

    mcpmeasured growthOpen source ↗

    Install claude mcp add openclaw-cost-tracker-mcp -- uvx openclaw-cost-tracker-mcp

    129downloads over 7 daysstable
  3. 3
    Active

    Reduces AI agent token usage by 90% via context compression and task checkpoint persistence.

    mcpmeasured growthOpen source ↗

    Install claude mcp add smart-context-mcp -- npx smart-context-mcp

    4GitHub starsstable
  4. 4
    Dormant

    Save tokens while coding — your AI agent gets structured code context, not file dumps.

    mcpmeasured growthOpen source ↗

    Install claude mcp add codeweave -- npx @codeweave/mcp

    4GitHub starsstable
  5. 5
    Dormant

    Analyze your MCP setup: token costs, grades, duplicates, and optimization tips

    mcpmeasured growthOpen source ↗

    Install claude mcp add mcp-checkup -- npx mcp-checkup

    1GitHub starsstable
  6. 6
    Active

    Cut AI agent token costs 5-15x — routes only relevant code symbols instead of full files.

    mcpmeasured growthOpen source ↗

    Install claude mcp add agent-booster -- uvx agent-booster

    1GitHub starsstable
  7. 7
    Active

    Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen

    mcpmeasured growthOpen source ↗

    Install claude mcp add vision-squeezer -- npx vision-squeezer

    2GitHub starsstable
  8. 8
    Active

    Token cost math for LLM API calls: verified per-1M-token rates for 69 models, 17 providers.

    mcpmeasured growthOpen source ↗

    Install claude mcp add comparedge-llm-cost -- npx @comparedge/llm-cost-mcp

    2GitHub starsstable
  9. 9
    Active

    Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.

    mcpmeasured growthOpen source ↗

    Install claude mcp add compactprompt -- uvx compactprompt

    0GitHub starsstable
  10. 10
    Active

    MCP server that cuts AI coding agent token usage via framework-aware context optimization

    mcpmeasured growthOpen source ↗

    Install claude mcp add ai-optimizer -- npx @ai-optimizer/core

    0GitHub starsstable
  11. 11
    Active

    Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable

    mcpmeasured growthOpen source ↗

    Install claude mcp add token_optimizer -- uvx slimctx

    1GitHub starsstable
  12. 12

    CLI
    Active

    CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies

    githubmeasured growthOpen source ↗

    78 241GitHub stars+697 (+0.90 %)
  13. 13

    MCP
    Active

    Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

    githubmeasured growthOpen source ↗

    68 350GitHub stars+630 (+0.93 %)
  14. 14
    Active

    Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add NodeNestor/claude-rolling-context

    31GitHub stars+1 (+3.3 %)
  15. 15

    MCP
    Active

    Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.

    githubmeasured growthOpen source ↗

    433GitHub stars+6 (+1.4 %)
  16. 16

    Skill
    Dormant

    Pith is the hook that makes Claude Code sessions last 3x longer.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add abhisekjha/pith

    98GitHub starsstable
  17. 17
    Active

    20 power-user skills for Claude Code: session memory and consolidation, context compression, multi-agent orchestration, adversarial bug hunting, security review, calibrated estimation and decision archaeology. Drop-in SKILL.md files. Built…

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/irfad7/claude-power-skills ~/.claude/skills/claude-power-skills

    4GitHub starsstable
  18. 18
    Active

    Restore prior Claude Code AND Codex sessions with zero LLM calls. Cross-tool session handoff, cache-expiry prevention, real cost dashboard. One plugin, both hosts.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add ww-w-ai/super-token-saver

    31GitHub starsstable
  19. 19
    Active

    Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured hardest on Claude Fable 5 (fable5), deep-studied on Sonnet 5, works on Opus 4.8 too. We publish our…

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sup3x/claude-code-eco

    33GitHub starsstable
  20. 20
    Dormant

    A Claude Code skill for web report decks — richer ECharts charts, one file per slide (edit one page, save tokens), and complex structure diagrams. 16:9 & vertical, PDF/image export.网页版 PPT 汇报的 Claude Code skill:ECharts 把数据画得更好看、每页一个文件改一页省…

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/myunwang/ppt-report-skills ~/.claude/skills/ppt-report-skills

    81GitHub starsstable
  21. 21

    Agent
    Active

    Control plane for AI coding agents: route tasks, reduce token spend, run multi-agent workflows, fallback executors, and track cost per task.

    githubmeasured growthOpen source ↗

    16GitHub starsstable
  22. 22
    Active

    A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Da-ya7/PromptCompressor-skill

    1GitHub starsstable
  23. 23

    MCP
    Dormant

    Audit MCP servers, skills, and CLAUDE.md bloat eating your Claude Code context window.

    githubmeasured growthOpen source ↗

    35GitHub starsstable
  24. 24

    Skill
    Active

    A production-ready Claude Code setup. Global CLAUDE.md, coding rules, per-project memory, and token optimization.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add janmaaarc/basecamp

    7GitHub starsstable

Listed without a public metric

Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    UnknownLast commit: unavailable

    Semantic context compression and cognitive memory layer for LLM swarms.

    mcpmetric not publishedOpen source ↗

    Remote server, to be declared in the MCP configuration https://recallmax-mcp.vercel.app/api/mcp

    No public metric provided