Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
115
Werkzeug-Rangliste
Ranked by normalized popularity across sources. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
langfuse-docs
Skill🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHub-Sterne+3 (+1.3 %) - 2Aktiv
idun-agent-platform
Skill🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform198GitHub-Sterne-1 (-0.50 %) - 3Aktiv
agent-kernel
AgentThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubgemessenes WachstumQuelle öffnen ↗
156GitHub-Sterne+19 (+13.9 %) - 4Aktiv
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub-Sterne+18 (+14.5 %) - 5Aktiv
VeriRun
SonstigeEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubgemessenes WachstumQuelle öffnen ↗
140GitHub-Sterne+24 (+20.7 %) - 6Aktiv
adlc-team-skills
Skill🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHub-Sterne+1 (+0.76 %) - 7Aktiv
dynatrace-for-ai
SkillSkills, prompts, and instructions for building AI agents on top of Dynatrace production context
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHub-Sterne+4 (+3.1 %) - 8Aktiv
rocketplaneIO
SonstigeSelf-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
githubgemessenes WachstumQuelle öffnen ↗
131GitHub-Sterne-4 (-3.0 %) - 9Aktiv
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
githubgemessenes WachstumQuelle öffnen ↗
129GitHub-Sterne-4 (-3.0 %) - 10Aktiv
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
githubgemessenes WachstumQuelle öffnen ↗
105GitHub-Sternestabil - 11105GitHub-Sterne+1 (+0.96 %)
- 12Aktiv
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHub-Sterne+4 (+4.3 %) - 13Ruhend
anti-lie
SkillDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHub-Sternestabil - 14Aktiv
MCP server for Langfuse LLM observability — trace and observation analysis.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add langfuse -- npx langfuse-observability-mcp-server78GitHub-Sternestabil - 15Aktiv
skill-kit
Skilllocal-first analytics for AI agent skills
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHub-Sterne+1 (+1.3 %) - 16Aktiv
skill-eval-harness
SkillAgent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHub-Sterne+4 (+5.8 %) - 17Aktiv
AgentX-Python
SonstigeAgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
githubgemessenes WachstumQuelle öffnen ↗
68GitHub-Sternestabil - 18Aktiv
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55GitHub-Sterne+27 (+96.4 %) - 19Aktiv
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHub-Sterne-5 (-8.9 %) - 20Aktiv
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
githubgemessenes WachstumQuelle öffnen ↗
50GitHub-Sterne+2 (+4.2 %) - 21Aktiv
arize-skills
SkillAgent skills for Arize — datasets, experiments, and traces via the ax CLI
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47GitHub-Sternestabil - 22Aktiv
cap-evolve
MCPOptimize any AI agent’s skills, tools/MCP, and prompts against your own evals.
githubgemessenes WachstumQuelle öffnen ↗
47GitHub-Sterne+1 (+2.2 %) - 23Aktiv
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubgeschätztes MomentumQuelle öffnen ↗
Installieren
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite39GitHub-Sterne— - 24Aktiv
claudestat
SkillReal-time execution trace and cost intelligence for Claude Code
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add DeibyGS/claudestat34GitHub-Sterne+1 (+3.0 %)
Lern- und Referenzressourcen
Ranked by normalized popularity across sources. Diese Ressourcen bleiben getrennt zugänglich und fließen nicht in die Hauptwertung ein.
- 1AktivQuelle öffnen ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRessourcegemessenes Wachstum77GitHub-Sternestabil