ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
123
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
A test runner for agentskills.io-style AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval713GitHubスター+12 (+1.7 %) - 225 035GitHubスター+26 (+0.10 %)
- 3活動中
untell
その他AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
github測定済み成長情報源を開く ↗
19GitHubスター+2 (+11.8 %) - 4活動中
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness72GitHubスター+4 (+5.9 %) - 5活動中
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
github測定済み成長情報源を開く ↗
22GitHubスター+2 (+10.0 %) - 6活動中
Skills, prompts, and instructions for building AI agents on top of Dynatrace production context
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai131GitHubスター+5 (+4.0 %) - 7活動中
ai-dev-stack
スキルProduction-grade AI coding rules for Cursor and Claude Code. 15 rules + 9 doc templates + skills + agents + MCP setup. Drop into any project.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/aiagentwithdhruv/ai-dev-stack ~/.claude/skills/ai-dev-stack10GitHubスター+1 (+11.1 %) - 8活動中
Audit which Claude Code skills you actually use — surface dead installs and hallucinated invocations from your session logs.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add sfrangulov/skill-graveyard10GitHubスター+1 (+11.1 %) - 9休止中
eval-layer
その他A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
github測定済み成長情報源を開く ↗
13GitHubスター+1 (+8.3 %) - 10活動中
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills96GitHubスター+3 (+3.2 %) - 11活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge807GitHubスター+7 (+0.88 %) - 12活動中
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
github測定済み成長情報源を開く ↗
49GitHubスター+2 (+4.3 %) - 13活動中
superlog
その他Open-source observability tool that uses AI agents to self-heal your software
github測定済み成長情報源を開く ↗
1 404GitHubスター+7 (+0.50 %) - 14活動中
SkillForge
スキルA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHubスター+5 (+0.57 %) - 15活動中
kubeshark
エージェントeBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
github測定済み成長情報源を開く ↗
12 065GitHubスター+7 (+0.06 %) - 16活動中
kubesphere
スキルThe container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 035GitHubスター+6 (+0.04 %) - 17活動中
Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add rennf93/opus-fable-playbook33GitHubスター+1 (+3.1 %) - 18活動中
claudestat
スキルReal-time execution trace and cost intelligence for Claude Code
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add DeibyGS/claudestat34GitHubスター+1 (+3.0 %) - 19活動中
cap-evolve
MCPOptimize any AI agent’s skills, tools/MCP, and prompts against your own evals.
github測定済み成長情報源を開く ↗
47GitHubスター+1 (+2.2 %) - 20活動中
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs237GitHubスター+2 (+0.85 %) - 21活動中
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
github測定済み成長情報源を開く ↗
258GitHubスター+2 (+0.78 %) - 22活動中
Dashboard for monitoring claude code sessions.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add JayantDevkar/claude-code-karma321GitHubスター+2 (+0.63 %) - 23活動中
skill-kit
スキルlocal-first analytics for AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHubスター+1 (+1.3 %)
学習・参考リソース
情報源間で正規化した測定済み成長順です。 これらは個別に参照でき、主要ランキングには含まれません。
- 1活動中情報源を開く ↗
trigger_tree
スキルDocumentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt
githubインストール
リソース測定済み成長/plugin marketplace add Hedde/trigger_tree14GitHubスター+1 (+7.7 %) - 2活動中情報源を開く ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubインストール
リソース測定済み成長git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability31GitHubスター+1 (+3.3 %)