Die Werkzeugbeschreibungen sind auf Englisch.

LLM-Beobachtbarkeit

Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.

Anwendungsfall

Aktivität

Sortieren nach

75

Werkzeug-Rangliste

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.

  1. 1
    Aktiv

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcpgemessenes WachstumQuelle öffnen ↗

    Installieren claude mcp add mcp-server -- npx @spanlens/mcp-server

    12GitHub-Sternestabil
  2. 2

    Sonstige
    Aktiv

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    githubgemessenes WachstumQuelle öffnen ↗

    24 768GitHub-Sterne+142 (+0.58 %)
  3. 3

    Sonstige
    Aktiv

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    782GitHub-Sterne+1 (+0.13 %)
  4. 4

    Agent
    Aktiv

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    githubgemessenes WachstumQuelle öffnen ↗

    27 783GitHub-Sterne+82 (+0.30 %)
  5. 5

    Bibliothek
    Aktiv

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    githubgemessenes WachstumQuelle öffnen ↗

    23 766GitHub-Sterne+66 (+0.28 %)
  6. 6

    Sonstige
    Aktiv

    A high-performance observability data pipeline.

    githubgemessenes WachstumQuelle öffnen ↗

    22 509GitHub-Sterne+40 (+0.18 %)
  7. 7

    Sonstige
    Aktiv

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    githubgemessenes WachstumQuelle öffnen ↗

    21 615GitHub-Sterne+96 (+0.45 %)
  8. 8

    Sonstige
    Aktiv

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    githubgemessenes WachstumQuelle öffnen ↗

    32 003GitHub-Sterne+50 (+0.16 %)
  9. 9

    Sonstige
    Aktiv

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    githubgemessenes WachstumQuelle öffnen ↗

    27 658GitHub-Sterne+128 (+0.46 %)
  10. 10

    Sonstige
    Aktiv

    The fastest path to AI-powered full stack observability, even for lean teams.

    githubgemessenes WachstumQuelle öffnen ↗

    80 412GitHub-Sterne+85 (+0.11 %)
  11. 11
    Aktiv

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    githubgemessenes WachstumQuelle öffnen ↗

    9 056GitHub-Sterne+31 (+0.34 %)
  12. 12

    Sonstige
    Aktiv

    eBPF-based Networking, Security, and Observability

    githubgemessenes WachstumQuelle öffnen ↗

    25 047GitHub-Sterne+29 (+0.12 %)
  13. 13

    Agent
    Aktiv

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubgemessenes WachstumQuelle öffnen ↗

    12 068GitHub-Sterne+8 (+0.07 %)
  14. 14

    Agent
    Aktiv

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    githubgemessenes WachstumQuelle öffnen ↗

    642GitHub-Sterne+32 (+5.2 %)
  15. 15

    Sonstige
    Aktiv

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    githubgemessenes WachstumQuelle öffnen ↗

    140GitHub-Sterne+24 (+20.7 %)
  16. 16

    Agent
    Aktiv

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    githubgemessenes WachstumQuelle öffnen ↗

    156GitHub-Sterne+19 (+13.9 %)
  17. 17
    Aktiv

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    githubgemessenes WachstumQuelle öffnen ↗

    3GitHub-Sternestabil
  18. 18
    Aktiv

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  19. 19

    MCP
    Aktiv

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    254GitHub-Sterne-3 (-1.2 %)
  20. 20

    Agent
    Aktiv

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    githubgemessenes WachstumQuelle öffnen ↗

    1GitHub-Sternestabil
  21. 21

    MCP
    Aktiv

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    githubgemessenes WachstumQuelle öffnen ↗

    2GitHub-Sternestabil
  22. 22

    Skill
    Aktiv

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub-Sterne+1 (+100.0 %)
  23. 23
    Aktiv

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2GitHub-Sternestabil
  24. 24
    Aktiv

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    239GitHub-Sterne+3 (+1.3 %)
  25. 25
    Aktiv

    Deterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.

    githubgemessenes WachstumQuelle öffnen ↗

    1GitHub-Sternestabil