工具说明为英文。

LLM 可观测性

LLM 应用的监控、追踪、评估和质量管理。

用途

活跃度

排序方式

122

工具排名

Ranked by creation date, newest first; undated entries come last. 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。

  1. 1

    MCP
    活跃

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    github实测增长打开来源 ↗

    50GitHub 星标+1 (+2.0 %)
  2. 2
    休眠

    A self-improving harness router for Claude Code.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add SeongwoongCho/adaptive-harness

    8GitHub 星标稳定
  3. 3
    休眠

    OpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  4. 4

    技能
    活跃

    Production-grade AI coding rules for Cursor and Claude Code. 15 rules + 9 doc templates + skills + agents + MCP setup. Drop into any project.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/aiagentwithdhruv/ai-dev-stack ~/.claude/skills/ai-dev-stack

    10GitHub 星标稳定
  5. 5
    活跃

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    97GitHub 星标+1 (+1.0 %)
  6. 6

    MCP
    活跃

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    github实测增长打开来源 ↗

    342GitHub 星标+84 (+32.6 %)
  7. 7

    MCP
    休眠

    Open-source Python framework to deploy AI agents via HTTP, A2A, and MCP with built-in observability

    github实测增长打开来源 ↗

    5GitHub 星标稳定
  8. 8

    智能体
    休眠

    AgentOps: Multi-agent infrastructure remediation platform. A2A protocol, agent coordination, HITL approval, auto-rollback.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  9. 9

    技能
    活跃

    local-first analytics for AI agent skills

    github实测增长打开来源 ↗

    安装 git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    78GitHub 星标+1 (+1.3 %)
  10. 10

    其他
    休眠

    Official Python SDK for GT8004 — AI agent observability with MCP, A2A, x402 payment tracking. FastAPI, Flask, FastMCP middleware included.

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  11. 11
    活跃

    Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.

    github实测增长打开来源 ↗

    22GitHub 星标+1 (+4.8 %)
  12. 12

    智能体
    休眠

    Policy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.

    github实测增长打开来源 ↗

    26GitHub 星标稳定
  13. 13
    休眠

    A Python proof-of-concept for tracing multi-turn Agent-to-Agent (A2A) conversations as a single unified MLflow trace for LLM observability and evaluation.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  14. 14

    技能
    活跃

    Measure prompt and skill improvements with blind A/B comparison.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add shinpr/rashomon

    18GitHub 星标稳定
  15. 15
    活跃

    Dashboard for monitoring claude code sessions.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add JayantDevkar/claude-code-karma

    325GitHub 星标+4 (+1.2 %)
  16. 16

    技能
    活跃

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    885GitHub 星标+1 (+0.11 %)
  17. 17
    休眠

    🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill

    27GitHub 星标稳定
  18. 18
    活跃

    Self improving agents through iterations

    github实测增长打开来源 ↗

    105GitHub 星标稳定
  19. 19
    休眠

    An AI-powered multi-agent system that demonstrates clinical triage, OTC medication recommendations, and e-pharmacy integration for respiratory conditions. Built with modular agents that collaborate to provide safe, intelligent healthcare…

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  20. 20
    休眠

    A safety-first multi-agent mental health companion with real-time distress tracking, triple-layer guardrails, and evidence-based grounding techniques. Built for Kaggle × Google Agents Intensive 2025 Capstone (Agents for Good Track)

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  21. 21

    其他
    休眠

    Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.

    github实测增长打开来源 ↗

    2GitHub 星标稳定
  22. 22
    休眠

    Local open-source dev tool to debug, secure, and evaluate LLM agents. Provides static analysis, dynamic security checks, and runtime monitoring - integrates with Cursor and Claude Code.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add cylestio/agent-inspector

    9GitHub 星标稳定
  23. 23
    活跃

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform

    2 379GitHub 星标+3 (+0.13 %)
  24. 24
    活跃

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator

    2 367GitHub 星标+9 (+0.38 %)
  25. 25

    智能体
    活跃

    A2A green-agent orchestrator for evaluating agents on the AppWorld benchmark, built on the AgentBeats SDK

    github实测增长打开来源 ↗

    0GitHub 星标稳定