Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
74
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHub-Sterne+204 (+92.7 %) - 2Aktiv
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55GitHub-Sterne+27 (+96.4 %) - 3Aktiv
SkillEvaluator
SkillMulti-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHub-Sterne+63 (+19.3 %) - 4Aktiv
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub-Sterne+21 (+17.4 %) - 5Aktiv
agent-kernel
AgentThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubgemessenes WachstumQuelle öffnen ↗
156GitHub-Sterne+22 (+16.4 %) - 6Aktiv
promptfoo
SonstigeTest your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
githubgemessenes WachstumQuelle öffnen ↗
24 738GitHub-Sterne+131 (+0.53 %) - 7Aktiv
openobserve
SonstigeOpen source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
githubgemessenes WachstumQuelle öffnen ↗
21 608GitHub-Sterne+115 (+0.54 %) - 8Aktiv
mastra
SonstigeMastra is the modern TypeScript framework for AI-powered applications and agents.
githubgemessenes WachstumQuelle öffnen ↗
27 624GitHub-Sterne+121 (+0.44 %) - 9Aktiv
yao-meta-skill
SkillYAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHub-Sterne+54 (+2.1 %) - 10Aktiv
agent-stack
SkillProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2GitHub-Sterne+1 (+100.0 %) - 11Aktiv
databuff
AgentDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
githubgemessenes WachstumQuelle öffnen ↗
634GitHub-Sterne+26 (+4.3 %) - 12Aktiv
mlflow
AgentThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
githubgemessenes WachstumQuelle öffnen ↗
27 768GitHub-Sterne+80 (+0.29 %) - 13Aktiv
netdata
SonstigeThe fastest path to AI-powered full stack observability, even for lean teams.
githubgemessenes WachstumQuelle öffnen ↗
80 402GitHub-Sterne+91 (+0.11 %) - 14Aktiv
prefect
BibliothekPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
githubgemessenes WachstumQuelle öffnen ↗
23 759GitHub-Sterne+66 (+0.28 %) - 15Aktiv
signoz
SonstigeSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubgemessenes WachstumQuelle öffnen ↗
31 992GitHub-Sterne+62 (+0.19 %) - 16Aktiv
agent-skill-creator
SkillBuild tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHub-Sterne+30 (+1.3 %) - 179 079GitHub-Sterne+43 (+0.48 %)
- 18Aktiv
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
githubgemessenes WachstumQuelle öffnen ↗
1 759GitHub-Sterne+24 (+1.4 %) - 19Aktiv
vector
SonstigeA high-performance observability data pipeline.
githubgemessenes WachstumQuelle öffnen ↗
22 507GitHub-Sterne+45 (+0.20 %) - 20Aktiv
agent-skills-eval
SkillA test runner for agentskills.io-style AI agent skills
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHub-Sterne+15 (+2.1 %) - 21Aktiv
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
githubgemessenes WachstumQuelle öffnen ↗
5GitHub-Sterne+1 (+25.0 %) - 22Aktiv
openstatus
MCP🫖 Status page with uptime monitoring & API monitoring as code 🫖
githubgemessenes WachstumQuelle öffnen ↗
9 052GitHub-Sterne+28 (+0.31 %) - 23Aktiv
dynatrace-for-ai
SkillSkills, prompts, and instructions for building AI agents on top of Dynatrace production context
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai132GitHub-Sterne+6 (+4.8 %) - 24Aktiv
cilium
SonstigeeBPF-based Networking, Security, and Observability
githubgemessenes WachstumQuelle öffnen ↗
25 041GitHub-Sterne+27 (+0.11 %) - 25Aktiv
skill-eval-harness
SkillAgent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHub-Sterne+4 (+5.8 %)