工具说明为英文。
LLM 可观测性
LLM 应用的监控、追踪、评估和质量管理。
用途
活跃度
排序方式
122
工具排名
混合排名:实测增长优先。 原始值保留各自的数据窗口;计算变化需要 7 天内至少两次测量。
- 1活跃
Make Claude write clearly, for everyone.
github实测增长打开来源 ↗
安装
/plugin marketplace add stefanobaghino/simple-output-styles16GitHub 星标稳定 - 2活跃
Marketing Mix Model expert agent plugin for Claude Code - pymc-marketing v0.18.2+
github实测增长打开来源 ↗
安装
/plugin marketplace add Yakoub-ai/agent-mmm4GitHub 星标稳定 - 3活跃
tulip-agents
智能体The agent framework where the model never holds the trigger — every consequential action clears your policy first, waits for a human when it matters, and lands on a record you can verify. Build on it, or put it around the agent you already…
github实测增长打开来源 ↗
1GitHub 星标稳定 - 4休眠
Agents-eval
智能体A Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.
github实测增长打开来源 ↗
2GitHub 星标稳定 - 5休眠
MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations
mcp实测增长打开来源 ↗
安装
claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge17GitHub 星标稳定 - 6休眠
An AI-powered multi-agent system that demonstrates clinical triage, OTC medication recommendations, and e-pharmacy integration for respiratory conditions. Built with modular agents that collaborate to provide safe, intelligent healthcare…
github实测增长打开来源 ↗
0GitHub 星标稳定 - 7活跃
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
github实测增长打开来源 ↗
105GitHub 星标稳定 - 8休眠
Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.
github实测增长打开来源 ↗
2GitHub 星标稳定 - 9活跃
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
github实测增长打开来源 ↗
0GitHub 星标稳定 - 10活跃
The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
github实测增长打开来源 ↗
安装
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 035GitHub 星标稳定 - 11休眠
Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.
github实测增长打开来源 ↗
安装
/plugin marketplace add prime-radiant-inc/parallel-adversarial-review17GitHub 星标稳定 - 12活跃
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github实测增长打开来源 ↗
安装
git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform2 376GitHub 星标稳定 - 13活跃
Run many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…
github实测增长打开来源 ↗
安装
git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases1GitHub 星标稳定 - 14活跃
headsup
技能Glanceable Claude Code and Codex state in your terminal tabs: white=idle, blue=working, orange=waiting. Multi-terminal (iTerm2, WezTerm, AI Power Term), a live status bar with Anthropic usage limits, /sfl and /nil window save-and-restore,…
github实测增长打开来源 ↗
安装
git clone https://github.com/wasulajr/headsup ~/.claude/skills/headsup1GitHub 星标稳定 - 15活跃
Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
github实测增长打开来源 ↗
131GitHub 星标-4 (-3.0 %) - 16活跃
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
github实测增长打开来源 ↗
129GitHub 星标-4 (-3.0 %) - 17活跃
superlog
其他Open-source observability tool that uses AI agents to self-heal your software
github估算动量打开来源 ↗
1 404GitHub 星标—
学习与参考资源
按实测增长排序,并在不同来源间归一化。 这些资源可单独访问,不参与主要排名。
- 1活跃打开来源 ↗
My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
github资源实测增长18GitHub 星标稳定 - 2活跃打开来源 ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
github资源实测增长77GitHub 星标稳定 - 3活跃打开来源 ↗
Handbook técnico aberto sobre engenharia de IA em produção: ML tradicional, LLMs, RAG, agentes, segurança, observabilidade, FinOps e deployment.
github资源实测增长1GitHub 星标稳定 - 4活跃打开来源 ↗
Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt
github安装
资源实测增长/plugin marketplace add Hedde/trigger_tree14GitHub 星标稳定 - 5活跃打开来源 ↗
tunelab
技能Claude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.
github安装
资源实测增长/plugin marketplace add rchaz/tunelab6GitHub 星标稳定