Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

111

Classifica degli strumenti

Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1
    Attivo

    MCP server for Langfuse LLM observability — trace and observation analysis.

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add langfuse -- npx langfuse-observability-mcp-server

    78stelle GitHubstabile
  2. 2
    Attivo

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    githubcrescita misurataApri fonte ↗

    68stelle GitHubstabile
  3. 3
    Attivo

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51stelle GitHub-5 (-8.9 %)
  4. 4

    Skill
    Attivo

    Agent skills for Arize — datasets, experiments, and traces via the ax CLI

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills

    47stelle GitHubstabile
  5. 5
    Attivo

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add rennf93/opus-fable-playbook

    33stelle GitHubstabile
  6. 6
    Dormiente

    🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill

    27stelle GitHubstabile
  7. 7

    Agente
    Dormiente

    Policy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.

    githubcrescita misurataApri fonte ↗

    26stelle GitHubstabile
  8. 8
    Attivo

    Sentry instrumentation skill for system-behavior tracking

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation

    24stelle GitHubstabile
  9. 9

    Altro
    Attivo

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    githubcrescita misurataApri fonte ↗

    18stelle GitHubstabile
  10. 10

    Skill
    Attivo

    Measure prompt and skill improvements with blind A/B comparison.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add shinpr/rashomon

    18stelle GitHubstabile
  11. 11
    Attivo

    Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add prime-radiant-inc/parallel-adversarial-review

    17stelle GitHubstabile
  12. 12
    Attivo

    Make Claude write clearly, for everyone.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add stefanobaghino/simple-output-styles

    16stelle GitHubstabile
  13. 13
    Attivo

    Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace

    13stelle GitHubstabile
  14. 14

    Altro
    Dormiente

    A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.

    githubcrescita misurataApri fonte ↗

    13stelle GitHubstabile
  15. 15
    Attivo

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add mcp-server -- npx @spanlens/mcp-server

    12stelle GitHubstabile
  16. 16

    Skill
    Attivo

    Turn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add CoriChui/bakeoff

    10stelle GitHubstabile
  17. 17

    Altro
    Attivo

    A Go-native framework for LLM agents, with OpenTelemetry observability built in.

    githubcrescita misurataApri fonte ↗

    10stelle GitHubstabile
  18. 18

    Skill
    Attivo

    Anchor — the production-grade AGENTS.md template for AI coding agents. 51 sections of battle-tested rules. Keep your agents grounded.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/peva3/anchor ~/.claude/skills/anchor

    8stelle GitHubstabile
  19. 19
    Dormiente

    A self-improving harness router for Claude Code.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add SeongwoongCho/adaptive-harness

    8stelle GitHubstabile
  20. 20

    Agente
    Attivo

    MemroOS / memroos: memory OS and governance layer for AI agents, agent workflows, dispatch, proof, and context continuity.

    githubcrescita misurataApri fonte ↗

    7stelle GitHubstabile
  21. 21

    MCP
    Dormiente

    Open-source Python framework to deploy AI agents via HTTP, A2A, and MCP with built-in observability

    githubcrescita misurataApri fonte ↗

    5stelle GitHubstabile
  22. 22

    Skill
    Attivo

    Axiom is a curated marketplace of shared plugins for Claude Code and Codex.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/netopsengineer/axiom ~/.claude/skills/axiom

    5stelle GitHubstabile

Risorse didattiche e di riferimento

Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRisorsacrescita misurata
    77stelle GitHubstabile
  2. 2
    AttivoApri fonte ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubRisorsacrescita misurata
    18stelle GitHubstabile
  3. 3

    tunelab

    Skill
    AttivoApri fonte ↗

    Claude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.

    github

    Installa /plugin marketplace add rchaz/tunelab

    Risorsacrescita misurata
    6stelle GitHubstabile