工具说明为英文。

LLM 可观测性

LLM 应用的监控、追踪、评估和质量管理。

用途

活跃度

排序方式

77

工具排名

Ranked by creation date, newest first; undated entries come last. 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。

  1. 1
    活跃

    Run many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…

    github估算动量打开来源 ↗

    安装 git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases

    1GitHub 星标
  2. 2

    智能体
    活跃

    面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  3. 3

    技能
    活跃

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    76GitHub 星标+48 (+171.4 %)
  4. 4

    智能体
    活跃

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  5. 5
    活跃

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    github实测增长打开来源 ↗

    3GitHub 星标稳定
  6. 6
    活跃

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2GitHub 星标稳定
  7. 7
    活跃

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    github实测增长打开来源 ↗

    安装 git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    43GitHub 星标+4 (+10.3 %)
  8. 8
    活跃

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    github实测增长打开来源 ↗

    22GitHub 星标+2 (+10.0 %)
  9. 9
    活跃

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github实测增长打开来源 ↗

    6GitHub 星标+2 (+50.0 %)
  10. 10

    技能
    活跃

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    493GitHub 星标+205 (+71.2 %)
  11. 11

    智能体
    活跃

    Local-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  12. 12

    技能
    活跃

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub 星标+1 (+100.0 %)
  13. 13

    技能
    活跃

    Research-backed, eval-driven skills for AI agents

    github实测增长打开来源 ↗

    安装 git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    145GitHub 星标+14 (+10.7 %)
  14. 14
    活跃

    Make Claude write clearly, for everyone.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add stefanobaghino/simple-output-styles

    16GitHub 星标稳定
  15. 15

    其他
    活跃

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    github实测增长打开来源 ↗

    163GitHub 星标+47 (+40.5 %)
  16. 16

    智能体
    活跃

    Deterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  17. 17
    活跃

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    github实测增长打开来源 ↗

    安装 git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHub 星标+1 (+0.76 %)
  18. 18

    MCP
    活跃

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    github实测增长打开来源 ↗

    2GitHub 星标稳定
  19. 19
    活跃

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    github实测增长打开来源 ↗

    1GitHub 星标稳定
  20. 20
    活跃

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  21. 21
    活跃

    Governed local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.

    mcp实测增长打开来源 ↗

    安装 claude mcp add ai-guardian -- uvx ai-guardian-aiops

    0GitHub 星标稳定
  22. 22

    其他
    活跃

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    github实测增长打开来源 ↗

    782GitHub 星标+1 (+0.13 %)

学习与参考资源

Ranked by creation date, newest first; undated entries come last. 这些资源可单独访问,不参与主要排名。

  1. 1
    活跃打开来源 ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    github资源实测增长
    18GitHub 星标稳定
  2. 2
    活跃打开来源 ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    安装 /plugin marketplace add Hedde/trigger_tree

    资源实测增长
    14GitHub 星标+1 (+7.7 %)
  3. 3
    活跃打开来源 ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    安装 git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    资源实测增长
    33GitHub 星标+3 (+10.0 %)