工具说明为英文。

LLM 可观测性

LLM 应用的监控、追踪、评估和质量管理。

用途

活跃度

排序方式

122

工具排名

按实测增长排序,并在不同来源间归一化。 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。

  1. 1

    智能体
    活跃

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    github实测增长打开来源 ↗

    12 068GitHub 星标+7 (+0.06 %)
  2. 2
    活跃

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    github实测增长打开来源 ↗

    安装 git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    45GitHub 星标+6 (+15.4 %)
  3. 3

    技能
    活跃

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    github实测增长打开来源 ↗

    安装 git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    240GitHub 星标+4 (+1.7 %)
  4. 4
    活跃

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    github实测增长打开来源 ↗

    安装 git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    135GitHub 星标+4 (+3.1 %)
  5. 5
    活跃

    Dashboard for monitoring claude code sessions.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add JayantDevkar/claude-code-karma

    324GitHub 星标+3 (+0.93 %)
  6. 6

    其他
    活跃

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    github实测增长打开来源 ↗

    783GitHub 星标+2 (+0.26 %)
  7. 7
    活跃

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    github实测增长打开来源 ↗

    22GitHub 星标+2 (+10.0 %)
  8. 8
    活跃

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcp实测增长打开来源 ↗

    安装 claude mcp add langfuse -- npx langfuse-observability-mcp-server

    80GitHub 星标+2 (+2.6 %)
  9. 9
    活跃

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    github实测增长打开来源 ↗

    安装 git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    73GitHub 星标+2 (+2.8 %)
  10. 10
    活跃

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    97GitHub 星标+1 (+1.0 %)
  11. 11

    技能
    活跃

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub 星标+1 (+100.0 %)
  12. 12
    活跃

    Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.

    github实测增长打开来源 ↗

    22GitHub 星标+1 (+4.8 %)
  13. 13

    技能
    活跃

    local-first analytics for AI agent skills

    github实测增长打开来源 ↗

    安装 git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77GitHub 星标+1 (+1.3 %)
  14. 14
    活跃

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add rennf93/opus-fable-playbook

    34GitHub 星标+1 (+3.0 %)
  15. 15

    MCP
    活跃

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    github实测增长打开来源 ↗

    50GitHub 星标+1 (+2.0 %)
  16. 16
    活跃

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github实测增长打开来源 ↗

    6GitHub 星标+1 (+20.0 %)
  17. 17

    其他
    活跃

    A command-line tool to scaffold and manage enterprise-ready AI Agents powered by the A2A (Agent-to-Agent) protocol

    github实测增长打开来源 ↗

    14GitHub 星标+1 (+7.7 %)
  18. 18
    活跃

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    github实测增长打开来源 ↗

    安装 git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHub 星标+1 (+0.76 %)
  19. 19

    技能
    活跃

    Anchor — the production-grade AGENTS.md template for AI coding agents. 51 sections of battle-tested rules. Keep your agents grounded.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/peva3/anchor ~/.claude/skills/anchor

    8GitHub 星标稳定
  20. 20
    活跃

    🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform

    199GitHub 星标稳定
  21. 21
    活跃

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  22. 22
    活跃

    AgentGateway — independent third-party profile of a public API surface, by API Evangelist. AgentGateway is an open-source, AI-native proxy and gateway for routing, observing, and governing traffic to and from AI agents, LLM providers, and…

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  23. 23
    休眠

    OpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  24. 24

    技能
    休眠

    Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie

    89GitHub 星标稳定

学习与参考资源

按实测增长排序,并在不同来源间归一化。 这些资源可单独访问,不参与主要排名。

  1. 1
    活跃打开来源 ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    安装 git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    资源实测增长
    33GitHub 星标+2 (+6.5 %)