Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

96 entries in this view.

Tool ranking

Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1

    Agent
    Active

    Agent context compressor and memory persistent orchestration

    githubmeasured growthOpen source ↗

    2GitHub starsstable
  2. 2
    Active

    Automatically detect and disable irrelevant Claude Code skills per project to save tokens and streamline your tech workflow.

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/mrdenox109-nyx/skill-manager ~/.claude/skills/skill-manager

    5GitHub starsstable
  3. 3
    Active

    Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add NodeNestor/claude-rolling-context

    31GitHub starsstable
  4. 4

    Skill
    Active

    mcp-recall compresses MCP tool outputs (94 KB → 3.5 KB · 96%) and stores full results in SQLite for retrieval — up to 30x more tool calls per session for heavy MCP workloads.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sakebomb/mcp-recall

    8GitHub starsstable
  5. 5
    Active

    Find the ghost tokens. Fix them. Survive compaction. Avoid context quality decay.

    githubmeasured growthOpen source ↗

    2 164GitHub stars+91 (+4.4 %)
  6. 6

    Skill
    Active

    Agent Dashboard: Visualization and analytics for Sessions and Quota Usage. Track, analyze, and optimize token usage across providers with heatmaps, cost tracking, token counting and quota resets..

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add JingbiaoMei/Tokdash

    67GitHub stars+2 (+3.1 %)
  7. 7

    Skill
    Active

    Content-aware output compression for AI coding assistants. 36 specialized processors cut CLI output tokens by 60-99% (git, pytest, npm, terraform, kubectl, docker, and more) without losing errors, diffs, or stack traces.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add ppgranger/token-saver

    142GitHub stars+2 (+1.4 %)
  8. 8

    CLI
    Active

    CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies

    githubmeasured growthOpen source ↗

    78 609GitHub stars+771 (+0.99 %)
  9. 9

    Skill
    Dormant

    Claude Code plugin that offloads large outputs to filesystem and retrieves when required.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sheeki03/Few-Word

    38GitHub starsstable
  10. 10

    MCP
    Active

    Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

    githubmeasured growthOpen source ↗

    68 952GitHub stars+925 (+1.4 %)
  11. 11
    Active

    Context compression plugin for Claude Code. Automatically trims large tool output—JSON, YAML, stack traces, and logs—before it enters the context window.

    githubmeasured growthOpen source ↗

    65GitHub stars+1 (+1.6 %)
  12. 12

    Skill
    Dormant

    Give Claude Code photographic memory in ONE portable file. No database, no SQLite, no ChromaDB - just a single .mv2 file you can git commit, scp, or share. Native Rust core with sub-ms operations.

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/memvid/claude-brain ~/.claude/skills/claude-brain

    575GitHub starsstable
  13. 13
    Active

    23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context optimization, research, and multi-LLM delegation. 186 skills, 128…

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/athola/claude-night-market ~/.claude/skills/claude-night-market

    335GitHub stars+3 (+0.90 %)
  14. 14

    MCP
    Active

    Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.

    githubmeasured growthOpen source ↗

    434GitHub stars+4 (+0.93 %)
  15. 15

    Skill
    Active

    Token-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add coalesce-labs/catalyst

    20GitHub starsstable
  16. 16
    UnknownLast commit: unavailable

    AI agent token-cost telemetry + 429 prediction. Per-agent attribution, anomaly + routing + quota.

    mcpmeasured growthOpen source ↗

    Install claude mcp add openclaw-cost-tracker-mcp -- uvx openclaw-cost-tracker-mcp

    129downloads over 7 daysstable
  17. 17
    UnknownLast commit: unavailable

    MCP server for Skim code transformation. Compresses code 60-95% for LLM context optimization.

    mcpestimated momentumOpen source ↗

    Install claude mcp add skim-mcp-server -- npx skim-mcp-server

    21downloads over 7 days

Listed without a public metric

Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    UnknownLast commit: unavailable

    Semantic context compression and cognitive memory layer for LLM swarms.

    mcpmetric not publishedOpen source ↗

    Remote server, to be declared in the MCP configuration https://recallmax-mcp.vercel.app/api/mcp

    No public metric provided
  2. 2

    MCP
    UnknownLast commit: unavailable

    Pre-injection context compression for coding agents. Zero LLM calls, zero telemetry, offline-safe.

    mcpmetric not publishedOpen source ↗

    No public metric provided
  3. 3
    UnknownLast commit: unavailable

    LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.

    mcpmetric not publishedOpen source ↗

    Remote server, to be declared in the MCP configuration https://cache.api.ainode.tech/mcp

    No public metric provided
  4. 4
    UnknownLast commit: unavailable

    Cut LLM token costs: count tokens, estimate cost, slim prompts, and pick the cheapest capable model.

    mcpmetric not publishedOpen source ↗

    Install claude mcp add mcp-token-optimizer -- npx mcp-token-optimizer

    No public metric provided