Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

90 entries in this view.

Tool ranking

Ranked by measured growth, normalized across sources. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    Active

    Measure and shrink the token cost of Claude Code skills, tier by tier: the always-on system-prompt tax, the per-invocation body, and on-demand references, each priced at the rate it is actually billed.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  2. 2

    Other
    Dormant

    Agent-to-Agent coordination that doesn't waste your context window. Token-efficient protocol with progressive discovery, zero-schema invocation, gRPC transport, and task lifecycle management.

    githubmeasured growthOpen source ↗

    1GitHub starsstable
  3. 3
    Active

    Token cost math for LLM API calls: verified per-1M-token rates for 69 models, 17 providers.

    mcpmeasured growthOpen source ↗

    Install claude mcp add comparedge-llm-cost -- npx @comparedge/llm-cost-mcp

    2GitHub starsstable
  4. 4
    Active

    Cut vision-LLM token costs by snapping images to tile boundaries — Claude, GPT, Gemini, Llama, Qwen

    mcpmeasured growthOpen source ↗

    Install claude mcp add vision-squeezer -- npx vision-squeezer

    2GitHub starsstable
  5. 5
    Active

    Pay-per-call agent APIs over x402: web scraping, token compression, and semantic cache.

    mcpmeasured growthOpen source ↗

    Install claude mcp add gate402-mcp -- npx gate402-mcp

    2GitHub stars+1 (+100.0 %)
  6. 6

    Skill
    Active

    🤫 Token-lean sessions at the harness level. An easy to grasp output style, output-shrinking hooks, and log compression cut both input and output tokens.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add V-Songbird/hush

    47GitHub stars+4 (+9.3 %)
  7. 7

    Other
    Dormant

    Translate any language to English before Claude processes it. ~49% token savings combined with caveman.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  8. 8
    Dormant

    ⚡ Cut Claude token usage by 90%+ — free, open-source, local-first context compression for Claude Code. Hybrid RAG (BM25 + ONNX vectors), AST chunking, reranking. No API needed.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Madhan230205/token-reducer

    44GitHub starsstable
  9. 9
    Active

    Claude Code skill: caveman-style token efficiency, spoken with a real Spanish regional accent

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Mun1to/SpanishCaveMan

    2GitHub starsstable
  10. 10
    Active

    Persistent memory for any AI — zero token cost until recall

    mcpmeasured growthOpen source ↗

    44GitHub stars+3 (+7.3 %)
  11. 11

    Agent
    Active

    Agent context compressor and memory persistent orchestration

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  12. 12
    Active

    MCP server that cuts AI coding agent token usage via framework-aware context optimization

    mcpmeasured growthOpen source ↗

    Install claude mcp add ai-optimizer -- npx @ai-optimizer/core

    0GitHub starsstable
  13. 13
    Active

    Claude Code skill: reduce token usage 60-80% — targeted reads, model switching, context management. One-line install.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  14. 14

    MCP
    Dormant

    Audit MCP servers, skills, and CLAUDE.md bloat eating your Claude Code context window.

    githubmeasured growthOpen source ↗

    35GitHub starsstable
  15. 15
    Active

    Turn your strongest model into a conductor: route each step to the cheapest capable model, and escalate only load-bearing pieces to a pricier reviewer that catches pitfalls early. Subscription, API, or hybrid.

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  16. 16
    Dormant

    Convert long AI conversations into portable conversation state graphs for LLM handoffs.

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/Adityapal67/context-graph-compressor ~/.claude/skills/context-graph-compressor

    35GitHub starsstable
  17. 17
    Active

    A Claude Code plugin that shows exactly where your AI coding session wasted tokens — and how to fix it. Flags waste, never quality. Runs locally.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Jimmynycu/token-efficiency

    2GitHub starsstable
  18. 18
    Active

    Eco mode for Claude Code. /eco: -31% to -73% output tokens with critical findings intact; /eco-max: up to -75% with lowered effort. Measured hardest on Claude Fable 5 (fable5), deep-studied on Sonnet 5, works on Opus 4.8 too. We publish our…

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sup3x/claude-code-eco

    33GitHub starsstable
  19. 19
    Dormant

    Save tokens while coding — your AI agent gets structured code context, not file dumps.

    mcpmeasured growthOpen source ↗

    Install claude mcp add codeweave -- npx @codeweave/mcp

    4GitHub starsstable
  20. 20
    Active

    Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add NodeNestor/claude-rolling-context

    31GitHub stars+1 (+3.3 %)
  21. 21
    Active

    Reduces AI agent token usage by 90% via context compression and task checkpoint persistence.

    mcpmeasured growthOpen source ↗

    Install claude mcp add smart-context-mcp -- npx smart-context-mcp

    4GitHub starsstable
  22. 22
    Active

    Restore prior Claude Code AND Codex sessions with zero LLM calls. Cross-tool session handoff, cache-expiry prevention, real cost dashboard. One plugin, both hosts.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add ww-w-ai/super-token-saver

    31GitHub starsstable
  23. 23
    Active

    Compact, source-linked context packs that reduce token waste for coding agents.

    mcpmeasured growthOpen source ↗

    Install claude mcp add capsule -- npx capsulectx

    4GitHub starsstable
  24. 24
    Dormant

    Give your Claude Code Agent Teams a memory. Auto-injects role-specific context into every new teammate — your team never starts blind again.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add Gr122lyBr/claude-teams-brain

    26GitHub starsstable
  25. 25
    Active

    20 power-user skills for Claude Code: session memory and consolidation, context compression, multi-agent orchestration, adversarial bug hunting, security review, calibrated estimation and decision archaeology. Drop-in SKILL.md files. Built…

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/irfad7/claude-power-skills ~/.claude/skills/claude-power-skills

    4GitHub starsstable