ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
75
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-server -- npx @spanlens/mcp-server12GitHubスター安定 - 2活動中
promptfoo
その他Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
github測定済み成長情報源を開く ↗
24 768GitHubスター+142 (+0.58 %) - 3782GitHubスター+1 (+0.13 %)
- 4活動中
mlflow
エージェントThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
github測定済み成長情報源を開く ↗
27 783GitHubスター+82 (+0.30 %) - 5活動中
prefect
ライブラリPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
github測定済み成長情報源を開く ↗
23 766GitHubスター+66 (+0.28 %) - 622 509GitHubスター+40 (+0.18 %)
- 7活動中
openobserve
その他Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
github測定済み成長情報源を開く ↗
21 615GitHubスター+96 (+0.45 %) - 8活動中
signoz
その他SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
github測定済み成長情報源を開く ↗
32 003GitHubスター+50 (+0.16 %) - 9活動中
mastra
その他Mastra is the modern TypeScript framework for AI-powered applications and agents.
github測定済み成長情報源を開く ↗
27 658GitHubスター+128 (+0.46 %) - 10活動中
netdata
その他The fastest path to AI-powered full stack observability, even for lean teams.
github測定済み成長情報源を開く ↗
80 412GitHubスター+85 (+0.11 %) - 119 056GitHubスター+31 (+0.34 %)
- 1225 047GitHubスター+29 (+0.12 %)
- 13活動中
kubeshark
エージェントeBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
github測定済み成長情報源を開く ↗
12 068GitHubスター+8 (+0.07 %) - 14活動中
databuff
エージェントDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
github測定済み成長情報源を開く ↗
642GitHubスター+32 (+5.2 %) - 15活動中
VeriRun
その他Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
github測定済み成長情報源を開く ↗
140GitHubスター+24 (+20.7 %) - 16活動中
agent-kernel
エージェントThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
github測定済み成長情報源を開く ↗
156GitHubスター+19 (+13.9 %) - 17活動中
Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…
github測定済み成長情報源を開く ↗
3GitHubスター安定 - 18活動中
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 19活動中
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
github測定済み成長情報源を開く ↗
254GitHubスター-3 (-1.2 %) - 20活動中
dsh-plugins
エージェントGeneric DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 21活動中
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
github測定済み成長情報源を開く ↗
2GitHubスター安定 - 22活動中
agent-stack
スキルProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2GitHubスター+1 (+100.0 %) - 23活動中
Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude2GitHubスター安定 - 24活動中
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHubスター+3 (+1.3 %) - 25活動中
boundary-bench
エージェントDeterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.
github測定済み成長情報源を開く ↗
1GitHubスター安定