ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
117
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
SkillCorpus
スキルOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHubスター+170 (+66.9 %) - 2活動中
VeriRun
その他Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
github測定済み成長情報源を開く ↗
140GitHubスター+24 (+20.7 %) - 3活動中
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHubスター+57 (+17.2 %) - 4活動中
craft-skills
スキルResearch-backed, eval-driven skills for AI agents
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHubスター+18 (+14.5 %) - 5活動中
agent-kernel
エージェントThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
github測定済み成長情報源を開く ↗
156GitHubスター+19 (+13.9 %) - 6活動中
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHubスター+4 (+5.8 %) - 7活動中
databuff
エージェントDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
github測定済み成長情報源を開く ↗
642GitHubスター+32 (+5.2 %) - 8活動中
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHubスター+4 (+4.3 %) - 9活動中
Skills, prompts, and instructions for building AI agents on top of Dynatrace production context
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHubスター+4 (+3.1 %) - 10活動中
A test runner for agentskills.io-style AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHubスター+13 (+1.8 %) - 11活動中
YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHubスター+35 (+1.4 %) - 12活動中
skill-kit
スキルlocal-first analytics for AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHubスター+1 (+1.3 %) - 13活動中
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHubスター+3 (+1.3 %) - 14活動中
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHubスター+28 (+1.2 %) - 15活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 759GitHubスター+18 (+1.0 %) - 16105GitHubスター+1 (+0.96 %)
- 17活動中
🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHubスター+1 (+0.76 %) - 18活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge809GitHubスター+6 (+0.75 %) - 19活動中
fable-method
スキルThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHubスター+16 (+0.71 %) - 20活動中
Dashboard for monitoring claude code sessions.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add JayantDevkar/claude-code-karma323GitHubスター+2 (+0.62 %) - 21活動中
promptfoo
その他Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
github測定済み成長情報源を開く ↗
24 768GitHubスター+142 (+0.58 %) - 22活動中
mastra
その他Mastra is the modern TypeScript framework for AI-powered applications and agents.
github測定済み成長情報源を開く ↗
27 658GitHubスター+128 (+0.46 %) - 23活動中
openobserve
その他Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
github測定済み成長情報源を開く ↗
21 615GitHubスター+96 (+0.45 %) - 249 056GitHubスター+31 (+0.34 %)
- 259 079GitHubスター+31 (+0.34 %)