工具说明为英文。

LLM 可观测性

LLM 应用的监控、追踪、评估和质量管理。

用途

活跃度

排序方式

112

工具排名

按实测增长排序,并在不同来源间归一化。 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。

  1. 1

    技能
    活跃

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    360GitHub 星标+168 (+87.5 %)
  2. 2
    活跃

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    363GitHub 星标+58 (+19.0 %)
  3. 3

    技能
    活跃

    Research-backed, eval-driven skills for AI agents

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    138GitHub 星标+21 (+17.9 %)
  4. 4

    技能
    活跃

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    38GitHub 星标+10 (+35.7 %)
  5. 5

    其他
    活跃

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    github实测增长打开来源 ↗

    27 601GitHub 星标+128 (+0.47 %)
  6. 6
    活跃

    YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill

    2 565GitHub 星标+58 (+2.3 %)
  7. 7

    MCP
    活跃

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    github实测增长打开来源 ↗

    2GitHub 星标+1 (+100.0 %)
  8. 8

    技能
    活跃

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub 星标+1 (+100.0 %)
  9. 9

    智能体
    活跃

    A local A2A event and continuity service for agent applications, with structured history, provenance-safe transcripts, literal search, and natural-language queries.

    github实测增长打开来源 ↗

    1GitHub 星标+1
  10. 10

    智能体
    活跃

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github实测增长打开来源 ↗

    146GitHub 星标+13 (+9.8 %)
  11. 11

    智能体
    活跃

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    github实测增长打开来源 ↗

    627GitHub 星标+22 (+3.6 %)
  12. 12

    其他
    活跃

    The fastest path to AI-powered full stack observability, even for lean teams.

    github实测增长打开来源 ↗

    80 382GitHub 星标+80 (+0.10 %)
  13. 13
    活跃

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator

    2 358GitHub 星标+29 (+1.2 %)
  14. 14

    其他
    活跃

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    github实测增长打开来源 ↗

    31 983GitHub 星标+48 (+0.15 %)
  15. 15
    活跃

    A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

    github实测增长打开来源 ↗

    1 752GitHub 星标+22 (+1.3 %)
  16. 16

    技能
    活跃

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 269GitHub 星标+22 (+0.98 %)
  17. 17
    活跃

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github实测增长打开来源 ↗

    5GitHub 星标+1 (+25.0 %)
  18. 18

    MCP
    休眠

    Open-source Python framework to deploy AI agents via HTTP, A2A, and MCP with built-in observability

    github实测增长打开来源 ↗

    5GitHub 星标+1 (+25.0 %)
  19. 19
    活跃

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    github实测增长打开来源 ↗

    9 045GitHub 星标+28 (+0.31 %)
  20. 20
    活跃

    A test runner for agentskills.io-style AI agent skills

    github实测增长打开来源 ↗

    安装 git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval

    713GitHub 星标+12 (+1.7 %)
  21. 21

    其他
    活跃

    eBPF-based Networking, Security, and Observability

    github实测增长打开来源 ↗

    25 035GitHub 星标+26 (+0.10 %)
  22. 22

    其他
    活跃

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    github实测增长打开来源 ↗

    19GitHub 星标+2 (+11.8 %)
  23. 23
    活跃

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    github实测增长打开来源 ↗

    安装 git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    72GitHub 星标+4 (+5.9 %)
  24. 24
    活跃

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    github实测增长打开来源 ↗

    22GitHub 星标+2 (+10.0 %)
  25. 25
    活跃

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    github实测增长打开来源 ↗

    安装 git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    131GitHub 星标+5 (+4.0 %)