Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
122
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424GitHub-Sterne+170 (+66.9 %) - 2Aktiv
promptfoo
SonstigeTest your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
githubgemessenes WachstumQuelle öffnen ↗
24 768GitHub-Sterne+142 (+0.58 %) - 3Aktiv
mastra
SonstigeMastra is the modern TypeScript framework for AI-powered applications and agents.
githubgemessenes WachstumQuelle öffnen ↗
27 658GitHub-Sterne+128 (+0.46 %) - 4Aktiv
openobserve
SonstigeOpen source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
githubgemessenes WachstumQuelle öffnen ↗
21 615GitHub-Sterne+96 (+0.45 %) - 5Aktiv
netdata
SonstigeThe fastest path to AI-powered full stack observability, even for lean teams.
githubgemessenes WachstumQuelle öffnen ↗
80 412GitHub-Sterne+85 (+0.11 %) - 6Aktiv
mlflow
AgentThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
githubgemessenes WachstumQuelle öffnen ↗
27 783GitHub-Sterne+82 (+0.30 %) - 7Aktiv
prefect
BibliothekPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
githubgemessenes WachstumQuelle öffnen ↗
23 766GitHub-Sterne+66 (+0.28 %) - 8Aktiv
SkillEvaluator
SkillMulti-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator389GitHub-Sterne+57 (+17.2 %) - 9Aktiv
signoz
SonstigeSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubgemessenes WachstumQuelle öffnen ↗
32 003GitHub-Sterne+50 (+0.16 %) - 10Aktiv
vector
SonstigeA high-performance observability data pipeline.
githubgemessenes WachstumQuelle öffnen ↗
22 509GitHub-Sterne+40 (+0.18 %) - 11Aktiv
yao-meta-skill
SkillYAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 577GitHub-Sterne+35 (+1.4 %) - 12Aktiv
databuff
AgentDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
githubgemessenes WachstumQuelle öffnen ↗
642GitHub-Sterne+32 (+5.2 %) - 13Aktiv
openstatus
MCP🫖 Status page with uptime monitoring & API monitoring as code 🫖
githubgemessenes WachstumQuelle öffnen ↗
9 056GitHub-Sterne+31 (+0.34 %) - 149 079GitHub-Sterne+31 (+0.34 %)
- 15Aktiv
cilium
SonstigeeBPF-based Networking, Security, and Observability
githubgemessenes WachstumQuelle öffnen ↗
25 047GitHub-Sterne+29 (+0.12 %) - 16Aktiv
agent-skill-creator
SkillBuild tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHub-Sterne+28 (+1.2 %) - 17Aktiv
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55GitHub-Sterne+27 (+96.4 %) - 18Aktiv
VeriRun
SonstigeEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubgemessenes WachstumQuelle öffnen ↗
140GitHub-Sterne+24 (+20.7 %) - 19Aktiv
agent-kernel
AgentThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubgemessenes WachstumQuelle öffnen ↗
156GitHub-Sterne+19 (+13.9 %) - 20Aktiv
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142GitHub-Sterne+18 (+14.5 %) - 21Aktiv
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
githubgemessenes WachstumQuelle öffnen ↗
1 759GitHub-Sterne+18 (+1.0 %) - 22Aktiv
fable-method
SkillThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHub-Sterne+16 (+0.71 %) - 23Aktiv
agent-skills-eval
SkillA test runner for agentskills.io-style AI agent skills
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval719GitHub-Sterne+13 (+1.8 %) - 24Aktiv
kubeshark
AgenteBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
githubgemessenes WachstumQuelle öffnen ↗
12 068GitHub-Sterne+8 (+0.07 %) - 25Aktiv
kubesphere
SkillThe container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 037GitHub-Sterne+8 (+0.05 %)