工具说明为英文。
LLM 可观测性
LLM 应用的监控、追踪、评估和质量管理。
用途
活跃度
排序方式
122
工具排名
Ranked by normalized popularity across sources. 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。
- 1活跃
Dashboard for monitoring claude code sessions.
github实测增长打开来源 ↗
安装
/plugin marketplace add JayantDevkar/claude-code-karma323GitHub 星标+2 (+0.62 %) - 2活跃
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
github实测增长打开来源 ↗
254GitHub 星标-3 (-1.2 %) - 3活跃
🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
github实测增长打开来源 ↗
安装
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHub 星标+3 (+1.3 %) - 4活跃
🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
github实测增长打开来源 ↗
安装
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform198GitHub 星标-1 (-0.50 %) - 5活跃
agent-kernel
智能体The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
github实测增长打开来源 ↗
156GitHub 星标+19 (+13.9 %) - 6活跃
Research-backed, eval-driven skills for AI agents
github实测增长打开来源 ↗
安装
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub 星标+18 (+14.5 %) - 7活跃
VeriRun
其他Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
github实测增长打开来源 ↗
140GitHub 星标+24 (+20.7 %) - 8活跃
🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
github实测增长打开来源 ↗
安装
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHub 星标+1 (+0.76 %) - 9活跃
Skills, prompts, and instructions for building AI agents on top of Dynatrace production context
github实测增长打开来源 ↗
安装
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHub 星标+4 (+3.1 %) - 10活跃
Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
github实测增长打开来源 ↗
131GitHub 星标-4 (-3.0 %) - 11活跃
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
github实测增长打开来源 ↗
129GitHub 星标-4 (-3.0 %) - 12活跃
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
github实测增长打开来源 ↗
105GitHub 星标稳定 - 13105GitHub 星标+1 (+0.96 %)
- 14活跃
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github实测增长打开来源 ↗
安装
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHub 星标+4 (+4.3 %) - 15休眠
anti-lie
技能Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github实测增长打开来源 ↗
安装
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHub 星标稳定 - 16活跃
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp实测增长打开来源 ↗
安装
claude mcp add langfuse -- npx langfuse-observability-mcp-server78GitHub 星标稳定 - 17活跃
local-first analytics for AI agent skills
github实测增长打开来源 ↗
安装
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHub 星标+1 (+1.3 %) - 18活跃
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github实测增长打开来源 ↗
安装
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHub 星标+4 (+5.8 %) - 19活跃
AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
github实测增长打开来源 ↗
68GitHub 星标稳定 - 20活跃
Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
github实测增长打开来源 ↗
安装
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55GitHub 星标+27 (+96.4 %) - 21活跃
Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
github实测增长打开来源 ↗
安装
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHub 星标-5 (-8.9 %) - 22活跃
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
github实测增长打开来源 ↗
50GitHub 星标+2 (+4.2 %) - 23活跃
Agent skills for Arize — datasets, experiments, and traces via the ax CLI
github实测增长打开来源 ↗
安装
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47GitHub 星标稳定 - 24活跃
cap-evolve
MCPOptimize any AI agent’s skills, tools/MCP, and prompts against your own evals.
github实测增长打开来源 ↗
47GitHub 星标+1 (+2.2 %)
学习与参考资源
Ranked by normalized popularity across sources. 这些资源可单独访问,不参与主要排名。
- 1活跃打开来源 ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
github资源实测增长77GitHub 星标稳定