Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

101 entries in this view.

Tool ranking

Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    Active

    A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Da-ya7/PromptCompressor-skill

    1GitHub starsstable
  2. 2
    Active

    Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  3. 3
    Dormant

    Analyze your MCP setup: token costs, grades, duplicates, and optimization tips

    mcpmeasured growthOpen source ↗

    Install claude mcp add mcp-checkup -- npx mcp-checkup

    1GitHub starsstable
  4. 4
    Active

    Cut AI agent token costs 5-15x — routes only relevant code symbols instead of full files.

    mcpmeasured growthOpen source ↗

    Install claude mcp add agent-booster -- uvx agent-booster

    1GitHub starsstable
  5. 5
    Active

    Reversible context compression for AI agents: cut token usage up to 94%, originals retrievable

    mcpmeasured growthOpen source ↗

    Install claude mcp add token_optimizer -- uvx slimctx

    1GitHub starsstable
  6. 6

    Other
    Dormant

    Agent-to-Agent coordination that doesn't waste your context window. Token-efficient protocol with progressive discovery, zero-schema invocation, gRPC transport, and task lifecycle management.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  7. 7
    Active

    Usage-aware status line for Claude Code: session & weekly rate-limit gauges with reset countdowns and a precise context-window meter, plus tokens, git & weather. Template-driven and extensible — set it up by chatting with Claude or via a…

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  8. 8

    Skill
    Dormant

    Agent context gate for Codex, Claude Code, Copilot, MCP, Cursor, Cline, Gemini and Windsurf repos

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/grnbtqdbyx-create/contextforge ~/.claude/skills/contextforge

    1GitHub starsstable
  9. 9

    MCP
    Active

    Context optimization for LLM agents: tool registry, result masking, budgeting, compaction.

    mcpmeasured growthOpen source ↗

    Install claude mcp add distil -- npx @munhq/distil

    1GitHub starsstable
  10. 10

    Skill
    Active

    Cut Claude Code token costs by delegating bulk LLM work to cheap models. 600+ models via NanoGPT, OpenRouter, Groq or local Ollama — batched, resumable, with real per-call cost tracking.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add bicced/sidequest

    1GitHub starsstable
  11. 11

    Skill
    Dormant

    Drop-in Cursor rules to cut AI response tokens (no filler/hedging, preserve code + technical strings). Includes local, reproducible benchmarks.

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/marcomondelli/bonsai ~/.claude/skills/bonsai

    0GitHub starsstable
  12. 12
    Active

    Statically audits MCP tool surfaces for token cost, schema quality, and design issues.

    mcpmeasured growthOpen source ↗

    Remote server, to be declared in the MCP configuration https://mcplint-web.vercel.app/api/mcp

    0GitHub starsstable
  13. 13
    Active

    Verifies token/cost savings claimed by AI context-reduction proxies via MCP tools.

    mcpmeasured growthOpen source ↗

    Install claude mcp add tokentrust -- uvx tokentrust-cli

    0GitHub starsstable
  14. 14
    Active

    Context GC for LLM agents: offload large tool outputs and recall them to save tokens.

    mcpmeasured growthOpen source ↗

    Install claude mcp add lethe -- uvx lethe-llm-context

    0GitHub starsstable
  15. 15
    Active

    MCP server that cuts AI coding agent token usage via framework-aware context optimization

    mcpmeasured growthOpen source ↗

    Install claude mcp add ai-optimizer -- npx @ai-optimizer/core

    0GitHub starsstable
  16. 16
    Active

    Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.

    mcpmeasured growthOpen source ↗

    Install claude mcp add compactprompt -- uvx compactprompt

    0GitHub starsstable
  17. 17

    Other
    Active

    Open-source TypeScript terminal coding agent for DeepSeek-V4 — builds on DeepSeek's strong price-performance and ultra-cheap cache pricing, engineering byte-stable prefixes and cache-reusing forks so cross-session memory and a continuous…

    githubestimated momentumOpen source ↗

    1 254GitHub stars
  18. 18

    Skill
    Active

    The Agent Knowledge Kit & Architecture System — 10–32x token-efficient architecture discovery for AI coding agents

    githubestimated momentumOpen source ↗

    Install git clone https://github.com/chama-x/quiv ~/.claude/skills/quiv

    3GitHub stars
  19. 19
    Active

    Installable agentic skills (SKILL.md) for Claude Code, Cursor, Codex, Gemini & Antigravity — 292+ product, efficiency, and common-sense skills. npx major-ai-skills

    githubestimated momentumOpen source ↗

    Install git clone https://github.com/alivirgo/Major-AI-Skills ~/.claude/skills/Major-AI-Skills

    1GitHub stars
  20. 20

    Agent
    Active

    Experimental no-install semantic language for AI agents: typed action/state, safe NL/JSON fallback, and falsifiable public evals. Try the one-file probe.

    githubestimated momentumOpen source ↗

    1GitHub stars

Learning resources

Ranked by measured growth, normalized across sources. These resources remain available separately and do not take part in the main tool ranking.

  1. 1
    ActiveOpen source ↗

    A cheat sheet that stops AI agents from scanning your whole repo and saves you tokens.

    github

    Install git clone https://github.com/toshon-jennings/SUMMARY-md ~/.claude/skills/SUMMARY-md

    Resourcemeasured growth
    0GitHub starsstable

Listed without a public metric

Mixed ranking: measured growth takes priority; entries without two snapshots are estimated. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    UnknownLast commit: unavailable

    Semantic context compression and cognitive memory layer for LLM swarms.

    mcpmetric not publishedOpen source ↗

    Remote server, to be declared in the MCP configuration https://recallmax-mcp.vercel.app/api/mcp

    No public metric provided
  2. 2

    MCP
    UnknownLast commit: unavailable

    Pre-injection context compression for coding agents. Zero LLM calls, zero telemetry, offline-safe.

    mcpmetric not publishedOpen source ↗

    No public metric provided
  3. 3
    UnknownLast commit: unavailable

    MCP server for Skim code transformation. Compresses code 60-95% for LLM context optimization.

    mcpmetric not publishedOpen source ↗

    Install claude mcp add skim-mcp-server -- npx skim-mcp-server

    No public metric provided
  4. 4
    UnknownLast commit: unavailable

    LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.

    mcpmetric not publishedOpen source ↗

    Remote server, to be declared in the MCP configuration https://cache.api.ainode.tech/mcp

    No public metric provided