ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
123
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
SkillCorpus
スキルOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHubスター+204 (+92.7 %) - 2活動中
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHubスター+63 (+19.3 %) - 3活動中
craft-skills
スキルResearch-backed, eval-driven skills for AI agents
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHubスター+21 (+17.4 %) - 4活動中
agent-kernel
エージェントThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
github測定済み成長情報源を開く ↗
156GitHubスター+22 (+16.4 %) - 5活動中
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHubスター+4 (+5.8 %) - 6活動中
Skills, prompts, and instructions for building AI agents on top of Dynatrace production context
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHubスター+6 (+4.8 %) - 7活動中
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHubスター+4 (+4.3 %) - 8活動中
databuff
エージェントDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
github測定済み成長情報源を開く ↗
634GitHubスター+26 (+4.3 %) - 9活動中
YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHubスター+54 (+2.1 %) - 10活動中
A test runner for agentskills.io-style AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHubスター+15 (+2.1 %) - 11活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 759GitHubスター+24 (+1.4 %) - 12活動中
skill-kit
スキルlocal-first analytics for AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHubスター+1 (+1.3 %) - 13活動中
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHubスター+30 (+1.3 %) - 14活動中
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHubスター+3 (+1.3 %) - 15活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge809GitHubスター+10 (+1.3 %) - 16105GitHubスター+1 (+0.96 %)
- 17活動中
Dashboard for monitoring claude code sessions.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add JayantDevkar/claude-code-karma323GitHubスター+3 (+0.94 %) - 18活動中
fable-method
スキルThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHubスター+19 (+0.84 %) - 19活動中
🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHubスター+1 (+0.76 %) - 20活動中
openobserve
その他Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
github測定済み成長情報源を開く ↗
21 608GitHubスター+115 (+0.54 %) - 21活動中
promptfoo
その他Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
github測定済み成長情報源を開く ↗
24 738GitHubスター+131 (+0.53 %) - 229 079GitHubスター+43 (+0.48 %)
- 23活動中
mastra
その他Mastra is the modern TypeScript framework for AI-powered applications and agents.
github測定済み成長情報源を開く ↗
27 624GitHubスター+121 (+0.44 %) - 24活動中
SkillForge
スキルA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHubスター+3 (+0.34 %) - 259 052GitHubスター+28 (+0.31 %)