ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

74

ツールランキング

情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1
    活動中

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcp測定済み成長情報源を開く ↗

    インストール claude mcp add langfuse -- npx langfuse-observability-mcp-server

    80GitHubスター+2 (+2.6 %)
  2. 2

    スキル
    活動中

    Axiom is a curated marketplace of shared plugins for Claude Code and Codex.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/netopsengineer/axiom ~/.claude/skills/axiom

    5GitHubスター安定
  3. 3
    活動中

    A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

    github測定済み成長情報源を開く ↗

    1 763GitHubスター+15 (+0.86 %)
  4. 4

    エージェント
    活動中

    面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  5. 5

    スキル
    活動中

    local-first analytics for AI agent skills

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77GitHubスター+1 (+1.3 %)
  6. 6

    MCP
    活動中

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    github測定済み成長情報源を開く ↗

    339GitHubスター+81 (+31.4 %)
  7. 7

    スキル
    活動中

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    73GitHubスター+2 (+2.8 %)
  8. 8

    エージェント
    活動中

    Provide clear documentation for AgentStack’s MCP protocol, plugins, and ecosystem API with usage examples and tool references.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  9. 9

    その他
    活動中

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    github測定済み成長情報源を開く ↗

    68GitHubスター安定
  10. 10
    活動中

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    github測定済み成長情報源を開く ↗

    3GitHubスター安定
  11. 11

    スキル
    活動中

    Real-time execution trace and cost intelligence for Claude Code

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add DeibyGS/claudestat

    34GitHubスター安定
  12. 12

    スキル
    活動中

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51GitHubスター-5 (-8.9 %)
  13. 13

    MCP
    活動中

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    github測定済み成長情報源を開く ↗

    50GitHubスター+1 (+2.0 %)
  14. 14

    スキル
    活動中

    YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill

    2 596GitHubスター+37 (+1.4 %)
  15. 15

    スキル
    活動中

    Dashboard for monitoring claude code sessions.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add JayantDevkar/claude-code-karma

    324GitHubスター+3 (+0.93 %)
  16. 16

    その他
    活動中

    A Go-native framework for LLM agents, with OpenTelemetry observability built in.

    github測定済み成長情報源を開く ↗

    10GitHubスター安定
  17. 17

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    166GitHubスター+25 (+17.7 %)
  18. 18

    エージェント
    活動中

    Deterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  19. 19

    エージェント
    活動中

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  20. 20

    スキル
    活動中

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    404GitHubスター+53 (+15.1 %)
  21. 21
    活動中

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2GitHubスター安定
  22. 22

    その他
    活動中

    Local-first Agentic AI Infrastructure Platform with LLM Gateway, Agent DAG Runtime, MCP Tool Hub, A2A Agent Mesh, Local RAG, Tool Sandbox and Observability.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  23. 23
    活動中

    Governed local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.

    mcp測定済み成長情報源を開く ↗

    インストール claude mcp add ai-guardian -- uvx ai-guardian-aiops

    0GitHubスター安定
  24. 24
    活動中

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github測定済み成長情報源を開く ↗

    6GitHubスター+1 (+20.0 %)
  25. 25
    活動中

    Self improving agents through iterations

    github測定済み成長情報源を開く ↗

    104GitHubスター安定