ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
123
ツールランキング
Ranked by creation date, newest first; undated entries come last. 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHubスター+15 (+0.64 %) - 2活動中
green-agent
エージェントA2A green-agent orchestrator for evaluating agents on the AppWorld benchmark, built on the AgentBeats SDK
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 3休止中
MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge17GitHubスター安定 - 4活動中
adl-cli
その他A command-line tool to scaffold and manage enterprise-ready AI Agents powered by the A2A (Agent-to-Agent) protocol
github測定済み成長情報源を開く ↗
14GitHubスター+1 (+7.7 %) - 5活動中
agent-kernel
エージェントThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
github測定済み成長情報源を開く ↗
166GitHubスター+25 (+17.7 %) - 6活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge816GitHubスター+11 (+1.4 %) - 7活動中
🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform199GitHubスター安定 - 8活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 763GitHubスター+15 (+0.86 %) - 9活動中
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add langfuse -- npx langfuse-observability-mcp-server80GitHubスター+2 (+2.6 %) - 10活動中
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
github測定済み成長情報源を開く ↗
105GitHubスター安定 - 11休止中
Agents-eval
エージェントA Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.
github測定済み成長情報源を開く ↗
2GitHubスター安定 - 12活動中
mastra
その他Mastra is the modern TypeScript framework for AI-powered applications and agents.
github測定済み成長情報源を開く ↗
27 712GitHubスター+149 (+0.54 %) - 139 060GitHubスター+31 (+0.34 %)
- 149 110GitHubスター+37 (+0.41 %)
- 15活動中
promptfoo
その他Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
github測定済み成長情報源を開く ↗
24 825GitHubスター+159 (+0.64 %) - 16活動中
openobserve
その他Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
github測定済み成長情報源を開く ↗
21 643GitHubスター+87 (+0.40 %) - 17活動中
kubeshark
エージェントeBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
github測定済み成長情報源を開く ↗
12 068GitHubスター+7 (+0.06 %) - 18活動中
signoz
その他SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
github測定済み成長情報源を開く ↗
32 012GitHubスター+42 (+0.13 %) - 1922 511GitHubスター+24 (+0.11 %)
- 20活動中
prefect
ライブラリPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
github測定済み成長情報源を開く ↗
23 783GitHubスター+67 (+0.28 %) - 21活動中
mlflow
エージェントThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
github測定済み成長情報源を開く ↗
27 817GitHubスター+86 (+0.31 %) - 2225 070GitHubスター+42 (+0.17 %)
- 23活動中
netdata
その他The fastest path to AI-powered full stack observability, even for lean teams.
github測定済み成長情報源を開く ↗
80 432GitHubスター+78 (+0.10 %)