工具说明为英文。
LLM 可观测性
LLM 应用的监控、追踪、评估和质量管理。
用途
活跃度
排序方式
117
工具排名
按实测增长排序,并在不同来源间归一化。 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。
- 1活跃
kubeshark
智能体eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
github实测增长打开来源 ↗
12 068GitHub 星标+7 (+0.06 %) - 2活跃
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
github实测增长打开来源 ↗
安装
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite45GitHub 星标+6 (+15.4 %) - 3活跃
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github实测增长打开来源 ↗
安装
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs240GitHub 星标+4 (+1.7 %) - 4活跃
Skills, prompts, and instructions for building AI agents on top of Dynatrace production context
github实测增长打开来源 ↗
安装
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai135GitHub 星标+4 (+3.1 %) - 5活跃
Dashboard for monitoring claude code sessions.
github实测增长打开来源 ↗
安装
/plugin marketplace add JayantDevkar/claude-code-karma324GitHub 星标+3 (+0.93 %) - 6783GitHub 星标+2 (+0.26 %)
- 7活跃
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
github实测增长打开来源 ↗
22GitHub 星标+2 (+10.0 %) - 8活跃
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp实测增长打开来源 ↗
安装
claude mcp add langfuse -- npx langfuse-observability-mcp-server80GitHub 星标+2 (+2.6 %) - 9活跃
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github实测增长打开来源 ↗
安装
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHub 星标+2 (+2.8 %) - 10活跃
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github实测增长打开来源 ↗
安装
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHub 星标+1 (+1.0 %) - 11活跃
Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
github实测增长打开来源 ↗
安装
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2GitHub 星标+1 (+100.0 %) - 12活跃
Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.
github实测增长打开来源 ↗
22GitHub 星标+1 (+4.8 %) - 13活跃
local-first analytics for AI agent skills
github实测增长打开来源 ↗
安装
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHub 星标+1 (+1.3 %) - 14活跃
Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
github实测增长打开来源 ↗
安装
/plugin marketplace add rennf93/opus-fable-playbook34GitHub 星标+1 (+3.0 %) - 15活跃
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
github实测增长打开来源 ↗
50GitHub 星标+1 (+2.0 %) - 16活跃
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
github实测增长打开来源 ↗
6GitHub 星标+1 (+20.0 %) - 17活跃
adl-cli
其他A command-line tool to scaffold and manage enterprise-ready AI Agents powered by the A2A (Agent-to-Agent) protocol
github实测增长打开来源 ↗
14GitHub 星标+1 (+7.7 %) - 18活跃
🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
github实测增长打开来源 ↗
安装
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHub 星标+1 (+0.76 %) - 19活跃
anchor
技能Anchor — the production-grade AGENTS.md template for AI coding agents. 51 sections of battle-tested rules. Keep your agents grounded.
github实测增长打开来源 ↗
安装
git clone https://github.com/peva3/anchor ~/.claude/skills/anchor8GitHub 星标稳定 - 20活跃
🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
github实测增长打开来源 ↗
安装
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform199GitHub 星标稳定 - 21活跃
a2a-otel-kit
MCPVendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.
github实测增长打开来源 ↗
1GitHub 星标稳定 - 22活跃
agentgateway
MCPAgentGateway — independent third-party profile of a public API surface, by API Evangelist. AgentGateway is an open-source, AI-native proxy and gateway for routing, observing, and governing traffic to and from AI agents, LLM providers, and…
github实测增长打开来源 ↗
0GitHub 星标稳定 - 23休眠
OpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.
github实测增长打开来源 ↗
0GitHub 星标稳定 - 24休眠
anti-lie
技能Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github实测增长打开来源 ↗
安装
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHub 星标稳定
学习与参考资源
按实测增长排序,并在不同来源间归一化。 这些资源可单独访问,不参与主要排名。
- 1活跃打开来源 ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
github安装
资源实测增长git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability33GitHub 星标+2 (+6.5 %)