Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
75
Werkzeug-Rangliste
Ranked by creation date, newest first; undated entries come last. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
CodeFlow
Agent面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 2Aktiv
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55GitHub-Sterne+27 (+96.4 %) - 3Aktiv
dsh-plugins
AgentGeneric DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 4Aktiv
Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil - 5Aktiv
Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude2GitHub-Sternestabil - 6Aktiv
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubgeschätztes MomentumQuelle öffnen ↗
Installieren
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite39GitHub-Sterne— - 7Aktiv
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
githubgemessenes WachstumQuelle öffnen ↗
22GitHub-Sterne+2 (+10.0 %) - 8Aktiv
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
githubgemessenes WachstumQuelle öffnen ↗
5GitHub-Sterne+1 (+25.0 %) - 9Aktiv
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHub-Sterne+170 (+66.9 %) - 10Aktiv
evoagent-os
AgentLocal-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 11Aktiv
agent-stack
SkillProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2GitHub-Sterne+1 (+100.0 %) - 12Aktiv
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub-Sterne+18 (+14.5 %) - 13Aktiv
simple-output-styles
SkillMake Claude write clearly, for everyone.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add stefanobaghino/simple-output-styles16GitHub-Sternestabil - 14Aktiv
VeriRun
SonstigeEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubgemessenes WachstumQuelle öffnen ↗
140GitHub-Sterne+24 (+20.7 %) - 15Aktiv
boundary-bench
AgentDeterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 16Aktiv
adlc-team-skills
Skill🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHub-Sterne+1 (+0.76 %) - 17Aktiv
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubgemessenes WachstumQuelle öffnen ↗
2GitHub-Sternestabil - 18Aktiv
a2a-otel-kit
MCPVendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 19Aktiv
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 20Aktiv
AI Guardian
MCPGoverned local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add ai-guardian -- uvx ai-guardian-aiops0GitHub-Sternestabil - 21Aktiv
Flawless
SonstigeAI SRE AgenticOps for Kubernetes and cloud infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
782GitHub-Sterne+1 (+0.13 %) - 22Aktiv
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHub-Sterne-5 (-8.9 %)
Lern- und Referenzressourcen
Ranked by creation date, newest first; undated entries come last. Diese Ressourcen bleiben getrennt zugänglich und fließen nicht in die Hauptwertung ein.
- 1AktivQuelle öffnen ↗
Agentic_AI_Engineer
AgentMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubRessourcegemessenes Wachstum18GitHub-Sternestabil - 2AktivQuelle öffnen ↗
trigger_tree
SkillDocumentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt
githubInstallieren
Ressourcegemessenes Wachstum/plugin marketplace add Hedde/trigger_tree14GitHub-Sterne+1 (+7.7 %) - 3AktivQuelle öffnen ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubInstallieren
Ressourcegemessenes Wachstumgit clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability32GitHub-Sterne+2 (+6.7 %)