工具说明为英文。

LLM 可观测性

LLM 应用的监控、追踪、评估和质量管理。

用途

活跃度

排序方式

117

工具排名

混合排名:实测增长优先。 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。

  1. 1

    智能体
    休眠

    A Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.

    github实测增长打开来源 ↗

    2GitHub 星标稳定
  2. 2
    活跃

    A Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability

    github实测增长打开来源 ↗

    105GitHub 星标稳定
  3. 3

    其他
    休眠

    Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.

    github实测增长打开来源 ↗

    2GitHub 星标稳定
  4. 4
    活跃

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    github实测增长打开来源 ↗

    0GitHub 星标稳定
  5. 5

    技能
    活跃

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    github实测增长打开来源 ↗

    安装 git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 035GitHub 星标稳定
  6. 6
    休眠

    Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.

    github实测增长打开来源 ↗

    安装 /plugin marketplace add prime-radiant-inc/parallel-adversarial-review

    17GitHub 星标稳定
  7. 7
    活跃

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github实测增长打开来源 ↗

    安装 git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform

    2 376GitHub 星标稳定
  8. 8
    活跃

    Run many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…

    github实测增长打开来源 ↗

    安装 git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases

    1GitHub 星标稳定
  9. 9

    技能
    活跃

    Glanceable Claude Code and Codex state in your terminal tabs: white=idle, blue=working, orange=waiting. Multi-terminal (iTerm2, WezTerm, AI Power Term), a live status bar with Anthropic usage limits, /sfl and /nil window save-and-restore,…

    github实测增长打开来源 ↗

    安装 git clone https://github.com/wasulajr/headsup ~/.claude/skills/headsup

    1GitHub 星标稳定
  10. 10

    其他
    活跃

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    github实测增长打开来源 ↗

    131GitHub 星标-4 (-3.0 %)
  11. 11
    活跃

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    github实测增长打开来源 ↗

    129GitHub 星标-4 (-3.0 %)
  12. 12

    其他
    活跃

    Open-source observability tool that uses AI agents to self-heal your software

    github估算动量打开来源 ↗

    1 404GitHub 星标

学习与参考资源

按实测增长排序,并在不同来源间归一化。 这些资源可单独访问,不参与主要排名。

  1. 1
    活跃打开来源 ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    github资源实测增长
    18GitHub 星标稳定
  2. 2
    活跃打开来源 ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    github资源实测增长
    77GitHub 星标稳定
  3. 3
    活跃打开来源 ↗

    Handbook técnico aberto sobre engenharia de IA em produção: ML tradicional, LLMs, RAG, agentes, segurança, observabilidade, FinOps e deployment.

    github资源实测增长
    1GitHub 星标稳定
  4. 4
    活跃打开来源 ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    安装 /plugin marketplace add Hedde/trigger_tree

    资源实测增长
    14GitHub 星标稳定
  5. 5

    tunelab

    技能
    活跃打开来源 ↗

    Claude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.

    github

    安装 /plugin marketplace add rchaz/tunelab

    资源实测增长
    6GitHub 星标稳定