Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
122
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
mlflow
AgentThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
githubgemessenes WachstumQuelle öffnen ↗
27 783GitHub-Sterne+82 (+0.30 %) - 2Aktiv
prefect
BibliothekPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
githubgemessenes WachstumQuelle öffnen ↗
23 766GitHub-Sterne+66 (+0.28 %) - 3Aktiv
vector
SonstigeA high-performance observability data pipeline.
githubgemessenes WachstumQuelle öffnen ↗
22 509GitHub-Sterne+40 (+0.18 %) - 4Aktiv
signoz
SonstigeSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubgemessenes WachstumQuelle öffnen ↗
32 003GitHub-Sterne+50 (+0.16 %) - 5Aktiv
superlog
SonstigeOpen-source observability tool that uses AI agents to self-heal your software
githubgemessenes WachstumQuelle öffnen ↗
1 404GitHub-Sterne+2 (+0.14 %) - 6Aktiv
Flawless
SonstigeAI SRE AgenticOps for Kubernetes and cloud infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
782GitHub-Sterne+1 (+0.13 %) - 7Aktiv
cilium
SonstigeeBPF-based Networking, Security, and Observability
githubgemessenes WachstumQuelle öffnen ↗
25 047GitHub-Sterne+29 (+0.12 %) - 8Aktiv
SkillForge
SkillA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHub-Sterne+1 (+0.11 %) - 9Aktiv
netdata
SonstigeThe fastest path to AI-powered full stack observability, even for lean teams.
githubgemessenes WachstumQuelle öffnen ↗
80 412GitHub-Sterne+85 (+0.11 %) - 10Aktiv
kubeshark
AgenteBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
githubgemessenes WachstumQuelle öffnen ↗
12 068GitHub-Sterne+8 (+0.07 %) - 11Aktiv
kubesphere
SkillThe container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 037GitHub-Sterne+8 (+0.05 %) - 12Aktiv
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
githubgemessenes WachstumQuelle öffnen ↗
105GitHub-Sternestabil - 13Ruhend
anti-lie
SkillDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHub-Sternestabil - 14Aktiv
MCP server for Langfuse LLM observability — trace and observation analysis.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add langfuse -- npx langfuse-observability-mcp-server78GitHub-Sternestabil - 15Aktiv
AgentX-Python
SonstigeAgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
githubgemessenes WachstumQuelle öffnen ↗
68GitHub-Sternestabil - 16Aktiv
idun-agent-platform
Skill🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform198GitHub-Sterne-1 (-0.50 %) - 17Aktiv
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
254GitHub-Sterne-3 (-1.2 %) - 18Aktiv
rocketplaneIO
SonstigeSelf-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
githubgemessenes WachstumQuelle öffnen ↗
131GitHub-Sterne-4 (-3.0 %) - 19Aktiv
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
githubgemessenes WachstumQuelle öffnen ↗
129GitHub-Sterne-4 (-3.0 %) - 20Aktiv
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHub-Sterne-5 (-8.9 %) - 21Aktiv
Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil - 22Ruhend
forge-skills
SkillAn assembly line for AI software development. 35 skills, 11 agent personas, 29 commands. From raw idea to shipped code.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add aneja5/forge-skills3GitHub-Sternestabil - 23Aktiv
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 24Aktiv
sre-on-call
AgentMulti-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil
Lern- und Referenzressourcen
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Diese Ressourcen bleiben getrennt zugänglich und fließen nicht in die Hauptwertung ein.
- 1AktivQuelle öffnen ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRessourcegemessenes Wachstum77GitHub-Sternestabil