Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

123

Classifica degli strumenti

Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1
    Attivo

    A test runner for agentskills.io-style AI agent skills

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval

    713stelle GitHub+12 (+1.7 %)
  2. 2

    Altro
    Attivo

    eBPF-based Networking, Security, and Observability

    githubcrescita misurataApri fonte ↗

    25 035stelle GitHub+26 (+0.10 %)
  3. 3

    Altro
    Attivo

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    githubcrescita misurataApri fonte ↗

    19stelle GitHub+2 (+11.8 %)
  4. 4
    Attivo

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    72stelle GitHub+4 (+5.9 %)
  5. 5
    Attivo

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    githubcrescita misurataApri fonte ↗

    22stelle GitHub+2 (+10.0 %)
  6. 6
    Attivo

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    131stelle GitHub+5 (+4.0 %)
  7. 7

    Skill
    Attivo

    Production-grade AI coding rules for Cursor and Claude Code. 15 rules + 9 doc templates + skills + agents + MCP setup. Drop into any project.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/aiagentwithdhruv/ai-dev-stack ~/.claude/skills/ai-dev-stack

    10stelle GitHub+1 (+11.1 %)
  8. 8
    Attivo

    Audit which Claude Code skills you actually use — surface dead installs and hallucinated invocations from your session logs.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add sfrangulov/skill-graveyard

    10stelle GitHub+1 (+11.1 %)
  9. 9

    Altro
    Dormiente

    A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.

    githubcrescita misurataApri fonte ↗

    13stelle GitHub+1 (+8.3 %)
  10. 10
    Attivo

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    96stelle GitHub+3 (+3.2 %)
  11. 11

    Skill
    Attivo

    OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge

    807stelle GitHub+7 (+0.88 %)
  12. 12

    MCP
    Attivo

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    githubcrescita misurataApri fonte ↗

    49stelle GitHub+2 (+4.3 %)
  13. 13

    Altro
    Attivo

    Open-source observability tool that uses AI agents to self-heal your software

    githubcrescita misurataApri fonte ↗

    1 404stelle GitHub+7 (+0.50 %)
  14. 14

    Skill
    Attivo

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    884stelle GitHub+5 (+0.57 %)
  15. 15

    Agente
    Attivo

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubcrescita misurataApri fonte ↗

    12 065stelle GitHub+7 (+0.06 %)
  16. 16

    Skill
    Attivo

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 035stelle GitHub+6 (+0.04 %)
  17. 17
    Attivo

    Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add rennf93/opus-fable-playbook

    33stelle GitHub+1 (+3.1 %)
  18. 18

    Skill
    Attivo

    Real-time execution trace and cost intelligence for Claude Code

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add DeibyGS/claudestat

    34stelle GitHub+1 (+3.0 %)
  19. 19
    Attivo

    Optimize any AI agent’s skills, tools/MCP, and prompts against your own evals.

    githubcrescita misurataApri fonte ↗

    47stelle GitHub+1 (+2.2 %)
  20. 20
    Attivo

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    237stelle GitHub+2 (+0.85 %)
  21. 21

    MCP
    Attivo

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    githubcrescita misurataApri fonte ↗

    258stelle GitHub+2 (+0.78 %)
  22. 22
    Attivo

    Dashboard for monitoring claude code sessions.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add JayantDevkar/claude-code-karma

    321stelle GitHub+2 (+0.63 %)
  23. 23

    Skill
    Attivo

    local-first analytics for AI agent skills

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77stelle GitHub+1 (+1.3 %)

Risorse didattiche e di riferimento

Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    Installa /plugin marketplace add Hedde/trigger_tree

    Risorsacrescita misurata
    14stelle GitHub+1 (+7.7 %)
  2. 2
    AttivoApri fonte ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    Installa git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    Risorsacrescita misurata
    31stelle GitHub+1 (+3.3 %)