Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
87 entries in this view.
Tool ranking
Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add NodeNestor/claude-rolling-context31GitHub starsstable - 2Active
mcp-recall
Skillmcp-recall compresses MCP tool outputs (94 KB → 3.5 KB · 96%) and stores full results in SQLite for retrieval — up to 30x more tool calls per session for heavy MCP workloads.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add sakebomb/mcp-recall8GitHub starsstable - 3Active
token-optimizer
OtherFind the ghost tokens. Fix them. Survive compaction. Avoid context quality decay.
githubmeasured growthOpen source ↗
2 170GitHub stars+86 (+4.1 %) - 4Active
Tokdash
SkillAgent Dashboard: Visualization and analytics for Sessions and Quota Usage. Track, analyze, and optimize token usage across providers with heatmaps, cost tracking, token counting and quota resets..
githubmeasured growthOpen source ↗
Install
/plugin marketplace add JingbiaoMei/Tokdash69GitHub stars+3 (+4.5 %) - 5Active
token-saver
SkillContent-aware output compression for AI coding assistants. 36 specialized processors cut CLI output tokens by 60-99% (git, pytest, npm, terraform, kubectl, docker, and more) without losing errors, diffs, or stack traces.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add ppgranger/token-saver142GitHub stars+2 (+1.4 %) - 6Active
rtk
CLICLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
githubmeasured growthOpen source ↗
78 957GitHub stars+981 (+1.3 %) - 7Active
headroom
MCPCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
githubmeasured growthOpen source ↗
69 050GitHub stars+927 (+1.4 %) - 8Active
RustAPI
LibraryA high-performance, ergonomic web framework for Rust with native AI/LLM support.
githubestimated momentumOpen source ↗
61GitHub stars— - 9Active
toonify-mcp
MCPContext compression plugin for Claude Code. Automatically trims large tool output—JSON, YAML, stack traces, and logs—before it enters the context window.
githubmeasured growthOpen source ↗
65GitHub stars+1 (+1.6 %) - 10Active
claude-night-market
Skill23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context optimization, research, and multi-LLM delegation. 186 skills, 128…
githubmeasured growthOpen source ↗
Install
git clone https://github.com/athola/claude-night-market ~/.claude/skills/claude-night-market335GitHub stars+3 (+0.90 %) - 11Active
ratel
MCPContext engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
githubmeasured growthOpen source ↗
434GitHub stars+4 (+0.93 %) - 12Active
catalyst
SkillToken-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add coalesce-labs/catalyst20GitHub starsstable