Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
116
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHub-Sterne+170 (+66.9 %) - 2Aktiv
VeriRun
SonstigeEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubgemessenes WachstumQuelle öffnen ↗
140GitHub-Sterne+24 (+20.7 %) - 3Aktiv
SkillEvaluator
SkillMulti-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHub-Sterne+57 (+17.2 %) - 4Aktiv
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub-Sterne+18 (+14.5 %) - 5Aktiv
agent-kernel
AgentThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubgemessenes WachstumQuelle öffnen ↗
156GitHub-Sterne+19 (+13.9 %) - 6Aktiv
skill-eval-harness
SkillAgent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHub-Sterne+4 (+5.8 %) - 7Aktiv
databuff
AgentDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
githubgemessenes WachstumQuelle öffnen ↗
642GitHub-Sterne+32 (+5.2 %) - 8Aktiv
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHub-Sterne+4 (+4.3 %) - 9Aktiv
dynatrace-for-ai
SkillSkills, prompts, and instructions for building AI agents on top of Dynatrace production context
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHub-Sterne+4 (+3.1 %) - 10Aktiv
agent-skills-eval
SkillA test runner for agentskills.io-style AI agent skills
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHub-Sterne+13 (+1.8 %) - 11Aktiv
yao-meta-skill
SkillYAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHub-Sterne+35 (+1.4 %) - 12Aktiv
skill-kit
Skilllocal-first analytics for AI agent skills
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHub-Sterne+1 (+1.3 %) - 13Aktiv
langfuse-docs
Skill🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs239GitHub-Sterne+3 (+1.3 %) - 14Aktiv
agent-skill-creator
SkillBuild tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHub-Sterne+28 (+1.2 %) - 15Aktiv
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
githubgemessenes WachstumQuelle öffnen ↗
1 759GitHub-Sterne+18 (+1.0 %) - 16105GitHub-Sterne+1 (+0.96 %)
- 17Aktiv
adlc-team-skills
Skill🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHub-Sterne+1 (+0.76 %) - 18Aktiv
OpenJudge
SkillOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge809GitHub-Sterne+6 (+0.75 %) - 19Aktiv
fable-method
SkillThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHub-Sterne+16 (+0.71 %) - 20Aktiv
claude-code-karma
SkillDashboard for monitoring claude code sessions.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add JayantDevkar/claude-code-karma323GitHub-Sterne+2 (+0.62 %) - 21Aktiv
mastra
SonstigeMastra is the modern TypeScript framework for AI-powered applications and agents.
githubgemessenes WachstumQuelle öffnen ↗
27 658GitHub-Sterne+128 (+0.46 %) - 22Aktiv
openstatus
MCP🫖 Status page with uptime monitoring & API monitoring as code 🫖
githubgemessenes WachstumQuelle öffnen ↗
9 056GitHub-Sterne+31 (+0.34 %) - 23Aktiv
signoz
SonstigeSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubgemessenes WachstumQuelle öffnen ↗
32 003GitHub-Sterne+50 (+0.16 %) - 24Aktiv
superlog
SonstigeOpen-source observability tool that uses AI agents to self-heal your software
githubgemessenes WachstumQuelle öffnen ↗
1 404GitHub-Sterne+2 (+0.14 %) - 25Aktiv
Flawless
SonstigeAI SRE AgenticOps for Kubernetes and cloud infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
782GitHub-Sterne+1 (+0.13 %)