Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

90 entries in this view.

Tool ranking

Ranked by measured growth, normalized across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1

    Other
    Dormant

    Agent-to-Agent coordination that doesn't waste your context window. Token-efficient protocol with progressive discovery, zero-schema invocation, gRPC transport, and task lifecycle management.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  2. 2
    Active

    Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  3. 3
    Active

    A skill that shrink noisy log/output 60-90% before read — filter, group, dedupe, truncate. Cut token, cut cost, keep signal. By Dhayalan (aka Joy).

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Da-ya7/PromptCompressor-skill

    1GitHub starsstable
  4. 4
    Active

    Token cost math for LLM API calls: verified per-1M-token rates for 69 models, 17 providers.

    mcpmeasured growthOpen source ↗

    Install claude mcp add comparedge-llm-cost -- npx @comparedge/llm-cost-mcp

    2GitHub starsstable
  5. 5
    Active

    Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen

    mcpmeasured growthOpen source ↗

    Install claude mcp add vision-squeezer -- npx vision-squeezer

    2GitHub starsstable
  6. 6

    Agent
    Active

    Agent context compressor and memory persistent orchestration

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  7. 7
    Active

    Claude Code skill: reduce token usage 60-80% — targeted reads, model switching, context management. One-line install.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  8. 8
    Active

    Turn your strongest model into a conductor: route each step to the cheapest capable model, and escalate only load-bearing pieces to a pricier reviewer that catches pitfalls early. Subscription, API, or hybrid.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  9. 9

    Skill
    Active

    🤫 Claude shuts up until it's done. A Claude Code plugin: one short answer at the end, the file to open, and no log dumps in your chat.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add V-Songbird/hush

    49GitHub stars+5 (+11.4 %)
  10. 10
    Active

    MCP server that cuts AI coding agent token usage via framework-aware context optimization

    mcpmeasured growthOpen source ↗

    Install claude mcp add ai-optimizer -- npx @ai-optimizer/core

    0GitHub starsstable
  11. 11
    Active

    Claude Code skill: caveman-style token efficiency, spoken with a real Spanish regional accent

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Mun1to/SpanishCaveMan

    2GitHub starsstable
  12. 12
    Dormant

    ⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Madhan230205/token-reducer

    44GitHub starsstable
  13. 13
    Active

    Pay-per-call agent APIs over x402: web scraping, token compression, and semantic cache.

    mcpmeasured growthOpen source ↗

    Install claude mcp add gate402-mcp -- npx gate402-mcp

    2GitHub starsstable
  14. 14
    Active

    Persistent memory for any AI — zero token cost until recall

    mcpmeasured growthOpen source ↗

    44GitHub starsstable
  15. 15
    Active

    Lossless context compression: 2-8x fewer tokens, byte-exact recovery, search inside payloads

    mcpmeasured growthOpen source ↗

    Install claude mcp add densely -- uvx densely

    6GitHub starsstable
  16. 16

    Other
    Dormant

    Translate any language to English before Claude processes it. ~49% token savings combined with caveman.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  17. 17

    Skill
    Dormant

    Claude Code plugin that offloads large outputs to filesystem and retrieves when required.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sheeki03/Few-Word

    38GitHub starsstable
  18. 18
    Active

    A Claude Code plugin that shows exactly where your AI coding session wasted tokens — and how to fix it. Flags waste, never quality. Runs locally.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Jimmynycu/token-efficiency

    2GitHub starsstable
  19. 19

    MCP
    Dormant

    Audit MCP servers, skills, and CLAUDE.md bloat eating your Claude Code context window.

    githubmeasured growthOpen source ↗

    35GitHub starsstable
  20. 20
    Dormant

    Save tokens while coding — your AI agent gets structured code context, not file dumps.

    mcpmeasured growthOpen source ↗

    Install claude mcp add codeweave -- npx @codeweave/mcp

    4GitHub starsstable
  21. 21
    Dormant

    Convert long AI conversations into portable conversation state graphs for LLM handoffs.

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/Adityapal67/context-graph-compressor ~/.claude/skills/context-graph-compressor

    35GitHub starsstable
  22. 22
    Active

    Review and compact prompts, docs, and AI-agent skills to save tokens, preserving structure.

    mcpmeasured growthOpen source ↗

    Install claude mcp add compactprompt -- uvx compactprompt

    0GitHub starsstable
  23. 23
    Active

    Reduces AI agent token usage by 90% via context compression and task checkpoint persistence.

    mcpmeasured growthOpen source ↗

    Install claude mcp add smart-context-mcp -- npx smart-context-mcp

    4GitHub starsstable
  24. 24
    Active

    Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured hardest on Claude Fable 5 (fable5), deep-studied on Sonnet 5, works on Opus 4.8 too. We publish our…

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sup3x/claude-code-eco

    33GitHub starsstable
  25. 25
    Active

    Compact, source-linked context packs that reduce token waste for coding agents.

    mcpmeasured growthOpen source ↗

    Install claude mcp add capsule -- npx capsulectx

    4GitHub starsstable