ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
123
ツールランキング
Ranked by known GitHub stars, highest first; archived repositories come after maintained repositories. 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
netdata
その他The fastest path to AI-powered full stack observability, even for lean teams.
github測定済み成長情報源を開く ↗
80 402GitHubスター+91 (+0.11 %) - 2活動中
signoz
その他SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
github測定済み成長情報源を開く ↗
31 992GitHubスター+62 (+0.19 %) - 3活動中
mlflow
エージェントThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
github測定済み成長情報源を開く ↗
27 768GitHubスター+80 (+0.29 %) - 4活動中
mastra
その他Mastra is the modern TypeScript framework for AI-powered applications and agents.
github測定済み成長情報源を開く ↗
27 624GitHubスター+121 (+0.44 %) - 525 041GitHubスター+27 (+0.11 %)
- 6活動中
promptfoo
その他Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
github測定済み成長情報源を開く ↗
24 738GitHubスター+131 (+0.53 %) - 7活動中
prefect
ライブラリPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
github測定済み成長情報源を開く ↗
23 759GitHubスター+66 (+0.28 %) - 822 507GitHubスター+45 (+0.20 %)
- 9活動中
openobserve
その他Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
github測定済み成長情報源を開く ↗
21 608GitHubスター+115 (+0.54 %) - 10活動中
kubesphere
スキルThe container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 037GitHubスター+9 (+0.05 %) - 11活動中
kubeshark
エージェントeBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
github測定済み成長情報源を開く ↗
12 066GitHubスター+7 (+0.06 %) - 129 079GitHubスター+43 (+0.48 %)
- 139 052GitHubスター+28 (+0.31 %)
- 14活動中
YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHubスター+54 (+2.1 %) - 15活動中
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHubスター+30 (+1.3 %) - 16活動中
fable-method
スキルThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHubスター+19 (+0.84 %) - 17活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 759GitHubスター+24 (+1.4 %) - 18活動中
superlog
その他Open-source observability tool that uses AI agents to self-heal your software
github測定済み成長情報源を開く ↗
1 404GitHubスター+2 (+0.14 %) - 19活動中
SkillForge
スキルA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHubスター+3 (+0.34 %) - 20活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge809GitHubスター+10 (+1.3 %) - 21782GitHubスター+1 (+0.13 %)
- 22活動中
A test runner for agentskills.io-style AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHubスター+15 (+2.1 %) - 23活動中
databuff
エージェントDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
github測定済み成長情報源を開く ↗
634GitHubスター+26 (+4.3 %) - 24活動中
SkillCorpus
スキルOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHubスター+204 (+92.7 %) - 25活動中
Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHubスター+63 (+19.3 %)