Die Werkzeugbeschreibungen sind auf Englisch.

LLM-Beobachtbarkeit

Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.

Anwendungsfall

Aktivität

Sortieren nach

114

Werkzeug-Rangliste

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.

  1. 1
    Aktiv

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcpgemessenes WachstumQuelle öffnen ↗

    Installieren claude mcp add mcp-server -- npx @spanlens/mcp-server

    12GitHub-Sternestabil
  2. 2

    Sonstige
    Aktiv

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    782GitHub-Sterne+1 (+0.13 %)
  3. 3

    Sonstige
    Aktiv

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    githubgemessenes WachstumQuelle öffnen ↗

    32 003GitHub-Sterne+50 (+0.16 %)
  4. 4

    Sonstige
    Aktiv

    The fastest path to AI-powered full stack observability, even for lean teams.

    githubgemessenes WachstumQuelle öffnen ↗

    80 412GitHub-Sterne+85 (+0.11 %)
  5. 5
    Aktiv

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    githubgemessenes WachstumQuelle öffnen ↗

    9 056GitHub-Sterne+31 (+0.34 %)
  6. 6

    Sonstige
    Aktiv

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    githubgemessenes WachstumQuelle öffnen ↗

    131GitHub-Sterne-4 (-3.0 %)
  7. 7

    Sonstige
    Aktiv

    eBPF-based Networking, Security, and Observability

    githubgemessenes WachstumQuelle öffnen ↗

    25 047GitHub-Sterne+29 (+0.12 %)
  8. 8

    Agent
    Aktiv

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubgemessenes WachstumQuelle öffnen ↗

    12 068GitHub-Sterne+8 (+0.07 %)
  9. 9

    Agent
    Aktiv

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    githubgemessenes WachstumQuelle öffnen ↗

    642GitHub-Sterne+32 (+5.2 %)
  10. 10

    Sonstige
    Aktiv

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    githubgemessenes WachstumQuelle öffnen ↗

    140GitHub-Sterne+24 (+20.7 %)
  11. 11

    Skill
    Aktiv

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 272GitHub-Sterne+16 (+0.71 %)
  12. 12

    Agent
    Aktiv

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    githubgemessenes WachstumQuelle öffnen ↗

    156GitHub-Sterne+19 (+13.9 %)
  13. 13
    Aktiv

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    githubgemessenes WachstumQuelle öffnen ↗

    3GitHub-Sternestabil
  14. 14

    Skill
    Ruhend

    An assembly line for AI software development. 35 skills, 11 agent personas, 29 commands. From raw idea to shipped code.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren /plugin marketplace add aneja5/forge-skills

    3GitHub-Sternestabil
  15. 15
    Aktiv

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  16. 16

    Agent
    Aktiv

    Multi-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.

    githubgemessenes WachstumQuelle öffnen ↗

    3GitHub-Sternestabil
  17. 17

    MCP
    Aktiv

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    254GitHub-Sterne-3 (-1.2 %)
  18. 18

    Agent
    Aktiv

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    githubgemessenes WachstumQuelle öffnen ↗

    1GitHub-Sternestabil
  19. 19

    MCP
    Aktiv

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    githubgemessenes WachstumQuelle öffnen ↗

    2GitHub-Sternestabil
  20. 20
    Aktiv

    Claude Code skills where every entry ships receipts — accuracy-gated benchmarks against baseline and placebo, rejects published

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/sjh9714/skill-receipts ~/.claude/skills/skill-receipts

    2GitHub-Sternestabil
  21. 21

    Skill
    Aktiv

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub-Sterne+1 (+100.0 %)
  22. 22

    Agent
    Aktiv

    A lightweight Python library that decouples agentic runtime from applications it builds

    githubgemessenes WachstumQuelle öffnen ↗

    1GitHub-Sternestabil
  23. 23
    Aktiv

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2GitHub-Sternestabil
  24. 24

    Sonstige
    Ruhend

    Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.

    githubgemessenes WachstumQuelle öffnen ↗

    2GitHub-Sternestabil
  25. 25
    Aktiv

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    239GitHub-Sterne+3 (+1.3 %)