Die Werkzeugbeschreibungen sind auf Englisch.

LLM-Beobachtbarkeit

Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.

Anwendungsfall

Aktivität

Sortieren nach

123

Werkzeug-Rangliste

Gemischte Rangfolge: Gemessenes Wachstum hat Vorrang. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.

  1. 1
    Aktiv

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcpgemessenes WachstumQuelle öffnen ↗

    Installieren claude mcp add mcp-server -- npx @spanlens/mcp-server

    12GitHub-Sternestabil
  2. 2

    Sonstige
    Aktiv

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    githubgemessenes WachstumQuelle öffnen ↗

    24 738GitHub-Sterne+131 (+0.53 %)
  3. 3

    Sonstige
    Aktiv

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    782GitHub-Sterne+1 (+0.13 %)
  4. 4

    Sonstige
    Aktiv

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    githubgemessenes WachstumQuelle öffnen ↗

    31 992GitHub-Sterne+62 (+0.19 %)
  5. 5

    Agent
    Aktiv

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    githubgemessenes WachstumQuelle öffnen ↗

    27 768GitHub-Sterne+80 (+0.29 %)
  6. 6

    Bibliothek
    Aktiv

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    githubgemessenes WachstumQuelle öffnen ↗

    23 759GitHub-Sterne+66 (+0.28 %)
  7. 7

    Sonstige
    Aktiv

    A high-performance observability data pipeline.

    githubgemessenes WachstumQuelle öffnen ↗

    22 507GitHub-Sterne+45 (+0.20 %)
  8. 8

    Sonstige
    Aktiv

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    githubgemessenes WachstumQuelle öffnen ↗

    21 608GitHub-Sterne+115 (+0.54 %)
  9. 9

    Sonstige
    Aktiv

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    githubgemessenes WachstumQuelle öffnen ↗

    27 624GitHub-Sterne+121 (+0.44 %)
  10. 10

    Sonstige
    Aktiv

    The fastest path to AI-powered full stack observability, even for lean teams.

    githubgemessenes WachstumQuelle öffnen ↗

    80 402GitHub-Sterne+91 (+0.11 %)
  11. 11
    Aktiv

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    githubgemessenes WachstumQuelle öffnen ↗

    9 052GitHub-Sterne+28 (+0.31 %)
  12. 12

    Sonstige
    Aktiv

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    githubgemessenes WachstumQuelle öffnen ↗

    131GitHub-Sterne-11 (-7.7 %)
  13. 13

    Sonstige
    Aktiv

    eBPF-based Networking, Security, and Observability

    githubgemessenes WachstumQuelle öffnen ↗

    25 041GitHub-Sterne+27 (+0.11 %)
  14. 14

    Agent
    Aktiv

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubgemessenes WachstumQuelle öffnen ↗

    12 066GitHub-Sterne+7 (+0.06 %)
  15. 15

    Agent
    Aktiv

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    githubgemessenes WachstumQuelle öffnen ↗

    634GitHub-Sterne+26 (+4.3 %)
  16. 16

    Sonstige
    Aktiv

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    githubgeschätztes MomentumQuelle öffnen ↗

    116GitHub-Sterne
  17. 17

    Skill
    Aktiv

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 272GitHub-Sterne+19 (+0.84 %)
  18. 18
    Aktiv

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    githubgemessenes WachstumQuelle öffnen ↗

    129GitHub-Sterne-10 (-7.2 %)
  19. 19

    Agent
    Ruhend

    Policy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.

    githubgemessenes WachstumQuelle öffnen ↗

    26GitHub-Sternestabil
  20. 20
    Ruhend

    OpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  21. 21
    Aktiv

    Mide si tu agente de IA cumple las reglas que le escribiste. Lee el historial local de Claude Code y devuelve un porcentaje por regla. Sin instalar nada, sin red, biblioteca estándar.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/jleonceo/adherencia-reglas ~/.claude/skills/adherencia-reglas

    1GitHub-Sternestabil
  22. 22

    Skill
    Aktiv

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    884GitHub-Sterne+3 (+0.34 %)
  23. 23
    Aktiv

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    132GitHub-Sterne+6 (+4.8 %)
  24. 24

    Sonstige
    Aktiv

    🔑 HomeStream · 家园·流 — 零成本自托管多Agent协作框架,通往AI世界的那把钥匙 | Zero-cost self-hosted multi-agent framework — The key to AI world

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  25. 25

    Skill
    Aktiv

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    424GitHub-Sterne+204 (+92.7 %)