ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

119

ツールランキング

情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1
    休止中

    Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add prime-radiant-inc/parallel-adversarial-review

    17GitHubスター安定
  2. 2

    エージェント
    休止中

    AgentOps: Multi-agent infrastructure remediation platform. A2A protocol, agent coordination, HITL approval, auto-rollback.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  3. 3

    スキル
    活動中

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHubスター安定
  4. 4

    エージェント
    休止中

    Real-time debugging proxy for Agent2Agent (A2A) multi-agent systems

    github測定済み成長情報源を開く ↗

    3GitHubスター安定
  5. 5

    スキル
    活動中

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    405GitHubスター+42 (+11.6 %)
  6. 6
    活動中

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  7. 7

    その他
    活動中

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    github測定済み成長情報源を開く ↗

    68GitHubスター安定
  8. 8
    活動中

    Governed local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.

    mcp測定済み成長情報源を開く ↗

    インストール claude mcp add ai-guardian -- uvx ai-guardian-aiops

    0GitHubスター安定
  9. 9

    スキル
    休止中

    A self-improving harness router for Claude Code.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add SeongwoongCho/adaptive-harness

    8GitHubスター安定
  10. 10

    スキル
    活動中

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    885GitHubスター+1 (+0.11 %)
  11. 11

    スキル
    活動中

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51GitHubスター+1 (+2.0 %)
  12. 12

    その他
    休止中

    Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.

    github測定済み成長情報源を開く ↗

    2GitHubスター安定
  13. 13
    活動中

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    github測定済み成長情報源を開く ↗

    130GitHubスター+1 (+0.78 %)
  14. 14

    MCP
    活動中

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    github測定済み成長情報源を開く ↗

    50GitHubスター+1 (+2.0 %)
  15. 15

    エージェント
    活動中

    MemroOS / memroos: memory OS and governance layer for AI agents, agent workflows, dispatch, proof, and context continuity.

    github測定済み成長情報源を開く ↗

    7GitHubスター安定
  16. 16
    活動中

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  17. 17

    エージェント
    活動中

    The agent framework where the model never holds the trigger — every consequential action clears your policy first, waits for a human when it matters, and lands on a record you can verify. Build on it, or put it around the agent you already…

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  18. 18

    エージェント
    休止中

    Contract-based testing for LLM agents: hybrid evaluation, multi-agent + A2A, deterministic replay.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  19. 19
    活動中

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add rennf93/opus-fable-playbook

    34GitHubスター+1 (+3.0 %)
  20. 20

    スキル
    活動中

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 035GitHubスター安定
  21. 21

    スキル
    活動中

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    240GitHubスター+3 (+1.3 %)
  22. 22

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    166GitHubスター+20 (+13.7 %)
  23. 23

    スキル
    活動中

    Real-time execution trace and cost intelligence for Claude Code

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add DeibyGS/claudestat

    34GitHubスター安定
  24. 24

    エージェント
    活動中

    A lightweight Python library that decouples agentic runtime from applications it builds

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  25. 25

    スキル
    活動中

    Research-backed, eval-driven skills for AI agents

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    148GitHubスター+10 (+7.2 %)