Die Werkzeugbeschreibungen sind auf Englisch.

LLM-Beobachtbarkeit

Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.

Anwendungsfall

Aktivität

Sortieren nach

122

Werkzeug-Rangliste

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.

  1. 1

    Agent
    Aktiv

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    githubgemessenes WachstumQuelle öffnen ↗

    27 783GitHub-Sterne+82 (+0.30 %)
  2. 2

    Bibliothek
    Aktiv

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    githubgemessenes WachstumQuelle öffnen ↗

    23 766GitHub-Sterne+66 (+0.28 %)
  3. 3

    Sonstige
    Aktiv

    A high-performance observability data pipeline.

    githubgemessenes WachstumQuelle öffnen ↗

    22 509GitHub-Sterne+40 (+0.18 %)
  4. 4

    Sonstige
    Aktiv

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    githubgemessenes WachstumQuelle öffnen ↗

    32 003GitHub-Sterne+50 (+0.16 %)
  5. 5

    Sonstige
    Aktiv

    Open-source observability tool that uses AI agents to self-heal your software

    githubgemessenes WachstumQuelle öffnen ↗

    1 404GitHub-Sterne+2 (+0.14 %)
  6. 6

    Sonstige
    Aktiv

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    782GitHub-Sterne+1 (+0.13 %)
  7. 7

    Sonstige
    Aktiv

    eBPF-based Networking, Security, and Observability

    githubgemessenes WachstumQuelle öffnen ↗

    25 047GitHub-Sterne+29 (+0.12 %)
  8. 8

    Skill
    Aktiv

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    884GitHub-Sterne+1 (+0.11 %)
  9. 9

    Sonstige
    Aktiv

    The fastest path to AI-powered full stack observability, even for lean teams.

    githubgemessenes WachstumQuelle öffnen ↗

    80 412GitHub-Sterne+85 (+0.11 %)
  10. 10

    Agent
    Aktiv

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubgemessenes WachstumQuelle öffnen ↗

    12 068GitHub-Sterne+8 (+0.07 %)
  11. 11

    Skill
    Aktiv

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 037GitHub-Sterne+8 (+0.05 %)
  12. 12
    Aktiv

    A Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability

    githubgemessenes WachstumQuelle öffnen ↗

    105GitHub-Sternestabil
  13. 13

    Skill
    Ruhend

    Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie

    89GitHub-Sternestabil
  14. 14
    Aktiv

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcpgemessenes WachstumQuelle öffnen ↗

    Installieren claude mcp add langfuse -- npx langfuse-observability-mcp-server

    78GitHub-Sternestabil
  15. 15

    Sonstige
    Aktiv

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    githubgemessenes WachstumQuelle öffnen ↗

    68GitHub-Sternestabil
  16. 16
    Aktiv

    🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform

    198GitHub-Sterne-1 (-0.50 %)
  17. 17

    MCP
    Aktiv

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    254GitHub-Sterne-3 (-1.2 %)
  18. 18

    Sonstige
    Aktiv

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    githubgemessenes WachstumQuelle öffnen ↗

    131GitHub-Sterne-4 (-3.0 %)
  19. 19
    Aktiv

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    githubgemessenes WachstumQuelle öffnen ↗

    129GitHub-Sterne-4 (-3.0 %)
  20. 20
    Aktiv

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51GitHub-Sterne-5 (-8.9 %)
  21. 21
    Aktiv

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    githubgemessenes WachstumQuelle öffnen ↗

    3GitHub-Sternestabil
  22. 22

    Skill
    Ruhend

    An assembly line for AI software development. 35 skills, 11 agent personas, 29 commands. From raw idea to shipped code.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren /plugin marketplace add aneja5/forge-skills

    3GitHub-Sternestabil
  23. 23
    Aktiv

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  24. 24

    Agent
    Aktiv

    Multi-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.

    githubgemessenes WachstumQuelle öffnen ↗

    3GitHub-Sternestabil

Lern- und Referenzressourcen

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Diese Ressourcen bleiben getrennt zugänglich und fließen nicht in die Hauptwertung ein.

  1. 1
    AktivQuelle öffnen ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRessourcegemessenes Wachstum
    77GitHub-Sternestabil