Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

123

Classifica degli strumenti

Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1

    Altro
    Attivo

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    githubcrescita misurataApri fonte ↗

    782stelle GitHub+1 (+0.13 %)
  2. 2

    MCP
    Attivo

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    githubcrescita misurataApri fonte ↗

    254stelle GitHub-3 (-1.2 %)
  3. 3
    Attivo

    🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform

    198stelle GitHubstabile
  4. 4
    Attivo

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    githubcrescita misurataApri fonte ↗

    131stelle GitHub-11 (-7.7 %)
  5. 5
    Attivo

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    githubcrescita misurataApri fonte ↗

    129stelle GitHub-10 (-7.2 %)
  6. 6
    Attivo

    A Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability

    githubcrescita misurataApri fonte ↗

    105stelle GitHubstabile
  7. 7

    Skill
    Dormiente

    Don't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie

    89stelle GitHubstabile
  8. 8
    Attivo

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add langfuse -- npx langfuse-observability-mcp-server

    78stelle GitHubstabile
  9. 9
    Attivo

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    githubcrescita misurataApri fonte ↗

    68stelle GitHubstabile
  10. 10
    Attivo

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51stelle GitHub-5 (-8.9 %)
  11. 11

    Skill
    Attivo

    Agent skills for Arize — datasets, experiments, and traces via the ax CLI

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills

    47stelle GitHubstabile
  12. 12
    Attivo

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add rennf93/opus-fable-playbook

    33stelle GitHubstabile
  13. 13
    Dormiente

    🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill

    27stelle GitHubstabile
  14. 14

    Agente
    Dormiente

    Policy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.

    githubcrescita misurataApri fonte ↗

    26stelle GitHubstabile
  15. 15
    Attivo

    Sentry instrumentation skill for system-behavior tracking

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation

    24stelle GitHubstabile
  16. 16

    Skill
    Attivo

    Measure prompt and skill improvements with blind A/B comparison.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add shinpr/rashomon

    18stelle GitHubstabile
  17. 17
    Attivo

    Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add prime-radiant-inc/parallel-adversarial-review

    17stelle GitHubstabile
  18. 18
    Dormiente

    MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge

    17stelle GitHubstabile
  19. 19
    Attivo

    Make Claude write clearly, for everyone.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add stefanobaghino/simple-output-styles

    16stelle GitHubstabile
  20. 20

    Altro
    Dormiente

    A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.

    githubcrescita misurataApri fonte ↗

    13stelle GitHubstabile
  21. 21
    Attivo

    Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace

    13stelle GitHubstabile
  22. 22
    Attivo

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add mcp-server -- npx @spanlens/mcp-server

    12stelle GitHubstabile
  23. 23

    Skill
    Attivo

    Turn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add CoriChui/bakeoff

    10stelle GitHubstabile

Risorse didattiche e di riferimento

Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRisorsacrescita misurata
    77stelle GitHubstabile
  2. 2
    AttivoApri fonte ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubRisorsacrescita misurata
    18stelle GitHubstabile