Die Werkzeugbeschreibungen sind auf Englisch.

LLM-Beobachtbarkeit

Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.

Anwendungsfall

Aktivität

Sortieren nach

117

Werkzeug-Rangliste

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.

  1. 1

    Agent
    Aktiv

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubgemessenes WachstumQuelle öffnen ↗

    12 068GitHub-Sterne+7 (+0.06 %)
  2. 2
    Aktiv

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    45GitHub-Sterne+6 (+15.4 %)
  3. 3
    Aktiv

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    240GitHub-Sterne+4 (+1.7 %)
  4. 4
    Aktiv

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    135GitHub-Sterne+4 (+3.1 %)
  5. 5
    Aktiv

    Dashboard for monitoring claude code sessions.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren /plugin marketplace add JayantDevkar/claude-code-karma

    324GitHub-Sterne+3 (+0.93 %)
  6. 6

    Sonstige
    Aktiv

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubgemessenes WachstumQuelle öffnen ↗

    783GitHub-Sterne+2 (+0.26 %)
  7. 7
    Aktiv

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    githubgemessenes WachstumQuelle öffnen ↗

    22GitHub-Sterne+2 (+10.0 %)
  8. 8
    Aktiv

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcpgemessenes WachstumQuelle öffnen ↗

    Installieren claude mcp add langfuse -- npx langfuse-observability-mcp-server

    80GitHub-Sterne+2 (+2.6 %)
  9. 9
    Aktiv

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    73GitHub-Sterne+2 (+2.8 %)
  10. 10
    Aktiv

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    97GitHub-Sterne+1 (+1.0 %)
  11. 11

    Skill
    Aktiv

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHub-Sterne+1 (+100.0 %)
  12. 12
    Aktiv

    Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.

    githubgemessenes WachstumQuelle öffnen ↗

    22GitHub-Sterne+1 (+4.8 %)
  13. 13

    Skill
    Aktiv

    local-first analytics for AI agent skills

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77GitHub-Sterne+1 (+1.3 %)
  14. 14
    Aktiv

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren /plugin marketplace add rennf93/opus-fable-playbook

    34GitHub-Sterne+1 (+3.0 %)
  15. 15

    MCP
    Aktiv

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    githubgemessenes WachstumQuelle öffnen ↗

    50GitHub-Sterne+1 (+2.0 %)
  16. 16
    Aktiv

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    githubgemessenes WachstumQuelle öffnen ↗

    6GitHub-Sterne+1 (+20.0 %)
  17. 17

    Sonstige
    Aktiv

    A command-line tool to scaffold and manage enterprise-ready AI Agents powered by the A2A (Agent-to-Agent) protocol

    githubgemessenes WachstumQuelle öffnen ↗

    14GitHub-Sterne+1 (+7.7 %)
  18. 18
    Aktiv

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHub-Sterne+1 (+0.76 %)
  19. 19

    Skill
    Aktiv

    Anchor — the production-grade AGENTS.md template for AI coding agents. 51 sections of battle-tested rules. Keep your agents grounded.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/peva3/anchor ~/.claude/skills/anchor

    8GitHub-Sternestabil
  20. 20
    Aktiv

    🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform

    199GitHub-Sternestabil
  21. 21
    Aktiv

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    githubgemessenes WachstumQuelle öffnen ↗

    1GitHub-Sternestabil
  22. 22
    Aktiv

    AgentGateway — independent third-party profile of a public API surface, by API Evangelist. AgentGateway is an open-source, AI-native proxy and gateway for routing, observing, and governing traffic to and from AI agents, LLM providers, and…

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  23. 23
    Ruhend

    OpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.

    githubgemessenes WachstumQuelle öffnen ↗

    0GitHub-Sternestabil
  24. 24

    Skill
    Ruhend

    Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.

    githubgemessenes WachstumQuelle öffnen ↗

    Installieren git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie

    89GitHub-Sternestabil

Lern- und Referenzressourcen

Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Diese Ressourcen bleiben getrennt zugänglich und fließen nicht in die Hauptwertung ein.

  1. 1
    AktivQuelle öffnen ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    Installieren git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    Ressourcegemessenes Wachstum
    33GitHub-Sterne+2 (+6.5 %)