Token efficiency: tools for developers

Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.

Use case

Activity

Sort by

87 entries in this view.

Tool ranking

Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.

  1. 1
    Active

    Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add NodeNestor/claude-rolling-context

    31GitHub starsstable
  2. 2

    Skill
    Active

    mcp-recall compresses MCP tool outputs (94 KB → 3.5 KB · 96%) and stores full results in SQLite for retrieval — up to 30x more tool calls per session for heavy MCP workloads.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add sakebomb/mcp-recall

    8GitHub starsstable
  3. 3
    Active

    Find the ghost tokens. Fix them. Survive compaction. Avoid context quality decay.

    githubmeasured growthOpen source ↗

    2 170GitHub stars+86 (+4.1 %)
  4. 4

    Skill
    Active

    Agent Dashboard: Visualization and analytics for Sessions and Quota Usage. Track, analyze, and optimize token usage across providers with heatmaps, cost tracking, token counting and quota resets..

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add JingbiaoMei/Tokdash

    69GitHub stars+3 (+4.5 %)
  5. 5

    Skill
    Active

    Content-aware output compression for AI coding assistants. 36 specialized processors cut CLI output tokens by 60-99% (git, pytest, npm, terraform, kubectl, docker, and more) without losing errors, diffs, or stack traces.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add ppgranger/token-saver

    142GitHub stars+2 (+1.4 %)
  6. 6

    CLI
    Active

    CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies

    githubmeasured growthOpen source ↗

    78 957GitHub stars+981 (+1.3 %)
  7. 7

    MCP
    Active

    Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.

    githubmeasured growthOpen source ↗

    69 050GitHub stars+927 (+1.4 %)
  8. 8

    Library
    Active

    A high-performance, ergonomic web framework for Rust with native AI/LLM support.

    githubestimated momentumOpen source ↗

    61GitHub stars
  9. 9
    Active

    Context compression plugin for Claude Code. Automatically trims large tool output—JSON, YAML, stack traces, and logs—before it enters the context window.

    githubmeasured growthOpen source ↗

    65GitHub stars+1 (+1.6 %)
  10. 10
    Active

    23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context optimization, research, and multi-LLM delegation. 186 skills, 128…

    githubmeasured growthOpen source ↗

    Install git clone https://github.com/athola/claude-night-market ~/.claude/skills/claude-night-market

    335GitHub stars+3 (+0.90 %)
  11. 11

    MCP
    Active

    Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.

    githubmeasured growthOpen source ↗

    434GitHub stars+4 (+0.93 %)
  12. 12

    Skill
    Active

    Token-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.

    githubmeasured growthOpen source ↗

    Install /plugin marketplace add coalesce-labs/catalyst

    20GitHub starsstable