Token efficiency: tools for developers
Tools that reduce token costs, compress context or cache LLM requests. An entry can appear under several use cases when its topics or description provide several explicit signals.
Use case
Activity
Sort by
96 entries in this view.
Tool ranking
Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1Active
cycgraph
AgentAgent context compressor and memory persistent orchestration
githubmeasured growthOpen source ↗
2GitHub starsstable - 2Active
skill-manager
SkillAutomatically detect and disable irrelevant Claude Code skills per project to save tokens and streamline your tech workflow.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/mrdenox109-nyx/skill-manager ~/.claude/skills/skill-manager5GitHub starsstable - 3Active
Rolling context compression for Claude Code — never hit the context wall. Auto-compresses old messages while keeping recent context verbatim. Zero config, zero latency. Works as a Claude Code plugin.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add NodeNestor/claude-rolling-context31GitHub starsstable - 4Active
mcp-recall
Skillmcp-recall compresses MCP tool outputs (94 KB → 3.5 KB · 96%) and stores full results in SQLite for retrieval — up to 30x more tool calls per session for heavy MCP workloads.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add sakebomb/mcp-recall8GitHub starsstable - 5Active
token-optimizer
OtherFind the ghost tokens. Fix them. Survive compaction. Avoid context quality decay.
githubmeasured growthOpen source ↗
2 164GitHub stars+91 (+4.4 %) - 6Active
Tokdash
SkillAgent Dashboard: Visualization and analytics for Sessions and Quota Usage. Track, analyze, and optimize token usage across providers with heatmaps, cost tracking, token counting and quota resets..
githubmeasured growthOpen source ↗
Install
/plugin marketplace add JingbiaoMei/Tokdash67GitHub stars+2 (+3.1 %) - 7Active
token-saver
SkillContent-aware output compression for AI coding assistants. 36 specialized processors cut CLI output tokens by 60-99% (git, pytest, npm, terraform, kubectl, docker, and more) without losing errors, diffs, or stack traces.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add ppgranger/token-saver142GitHub stars+2 (+1.4 %) - 8Active
rtk
CLICLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies
githubmeasured growthOpen source ↗
78 609GitHub stars+771 (+0.99 %) - 9Dormant
Few-Word
SkillClaude Code plugin that offloads large outputs to filesystem and retrieves when required.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add sheeki03/Few-Word38GitHub starsstable - 10Active
headroom
MCPCompress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.
githubmeasured growthOpen source ↗
68 952GitHub stars+925 (+1.4 %) - 11Active
toonify-mcp
MCPContext compression plugin for Claude Code. Automatically trims large tool output—JSON, YAML, stack traces, and logs—before it enters the context window.
githubmeasured growthOpen source ↗
65GitHub stars+1 (+1.6 %) - 12Dormant
claude-brain
SkillGive Claude Code photographic memory in ONE portable file. No database, no SQLite, no ChromaDB - just a single .mv2 file you can git commit, scp, or share. Native Rust core with sub-ms operations.
githubmeasured growthOpen source ↗
Install
git clone https://github.com/memvid/claude-brain ~/.claude/skills/claude-brain575GitHub starsstable - 13Active
claude-night-market
Skill23 Claude Code plugins: TDD enforcement hooks, git/PR workflows, spec-driven development, code review, project lifecycle, fix-from-error, maintenance automation, context optimization, research, and multi-LLM delegation. 186 skills, 128…
githubmeasured growthOpen source ↗
Install
git clone https://github.com/athola/claude-night-market ~/.claude/skills/claude-night-market335GitHub stars+3 (+0.90 %) - 14Active
ratel
MCPContext engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
githubmeasured growthOpen source ↗
434GitHub stars+4 (+0.93 %) - 15Active
catalyst
SkillToken-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.
githubmeasured growthOpen source ↗
Install
/plugin marketplace add coalesce-labs/catalyst20GitHub starsstable - 16UnknownLast commit: unavailable
AI agent token-cost telemetry + 429 prediction. Per-agent attribution, anomaly + routing + quota.
mcpmeasured growthOpen source ↗
Install
claude mcp add openclaw-cost-tracker-mcp -- uvx openclaw-cost-tracker-mcp129downloads over 7 daysstable - 17UnknownLast commit: unavailable
MCP server for Skim code transformation. Compresses code 60-95% for LLM context optimization.
mcpestimated momentumOpen source ↗
Install
claude mcp add skim-mcp-server -- npx skim-mcp-server21downloads over 7 days—
Listed without a public metric
Ranked by creation date, newest first; undated entries come last. Raw values keep their source-specific window, shown under each tool; a change requires at least two snapshots from the last 7 days.
- 1UnknownLast commit: unavailable
Semantic context compression and cognitive memory layer for LLM swarms.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://recallmax-mcp.vercel.app/api/mcp—No public metric provided - 2UnknownLast commit: unavailable
sqz
MCPPre-injection context compression for coding agents. Zero LLM calls, zero telemetry, offline-safe.
mcpmetric not publishedOpen source ↗
—No public metric provided - 3UnknownLast commit: unavailable
LLM caching proxy (x402 USDC on Base) - exact + semantic cache. Free health.
mcpmetric not publishedOpen source ↗
Remote server, to be declared in the MCP configuration
https://cache.api.ainode.tech/mcp—No public metric provided - 4UnknownLast commit: unavailable
Cut LLM token costs: count tokens, estimate cost, slim prompts, and pick the cheapest capable model.
mcpmetric not publishedOpen source ↗
Install
claude mcp add mcp-token-optimizer -- npx mcp-token-optimizer—No public metric provided