Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
76
Classifica degli strumenti
Ranked by normalized popularity across sources. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1Attivo
agent-kernel
AgenteThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubcrescita misurataApri fonte ↗
166stelle GitHub+29 (+21.2 %) - 2Attivo
VeriRun
AltroEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubcrescita misurataApri fonte ↗
163stelle GitHub+47 (+40.5 %) - 3Attivo
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills145stelle GitHub+14 (+10.7 %) - 4Attivo
dynatrace-for-ai
SkillSkills, prompts, and instructions for building AI agents on top of Dynatrace production context
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai135stelle GitHub+4 (+3.1 %) - 5Attivo
adlc-team-skills
Skill🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132stelle GitHub+1 (+0.76 %) - 6Attivo
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
githubcrescita misurataApri fonte ↗
129stelle GitHub-4 (-3.0 %) - 7Attivo
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
githubcrescita misurataApri fonte ↗
105stelle GitHubstabile - 8105stelle GitHub+1 (+0.96 %)
- 9Attivo
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97stelle GitHub+1 (+1.0 %) - 10Attivo
MCP server for Langfuse LLM observability — trace and observation analysis.
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add langfuse -- npx langfuse-observability-mcp-server80stelle GitHub+2 (+2.6 %) - 11Attivo
skill-kit
Skilllocal-first analytics for AI agent skills
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77stelle GitHub+1 (+1.3 %) - 12Attivo
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT76stelle GitHub+48 (+171.4 %) - 13Attivo
skill-eval-harness
SkillAgent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73stelle GitHub+4 (+5.8 %) - 14Attivo
AgentX-Python
AltroAgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
githubcrescita misurataApri fonte ↗
68stelle GitHubstabile - 15Attivo
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51stelle GitHub-5 (-8.9 %) - 16Attivo
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
githubcrescita misurataApri fonte ↗
50stelle GitHub+2 (+4.2 %) - 17Attivo
arize-skills
SkillAgent skills for Arize — datasets, experiments, and traces via the ax CLI
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47stelle GitHubstabile - 18Attivo
cap-evolve
MCPOptimize any AI agent’s skills, tools/MCP, and prompts against your own evals.
githubmomentum stimatoApri fonte ↗
47stelle GitHub— - 19Attivo
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite43stelle GitHub+4 (+10.3 %) - 20Attivo
claudestat
SkillReal-time execution trace and cost intelligence for Claude Code
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add DeibyGS/claudestat34stelle GitHubstabile - 21Attivo
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
githubcrescita misurataApri fonte ↗
22stelle GitHub+2 (+10.0 %) - 22Attivo
untell
AltroAI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
githubcrescita misurataApri fonte ↗
18stelle GitHubstabile - 23Attivo
rashomon
SkillMeasure prompt and skill improvements with blind A/B comparison.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add shinpr/rashomon18stelle GitHubstabile
Risorse didattiche e di riferimento
Ranked by normalized popularity across sources. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.
- 1AttivoApri fonte ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRisorsacrescita misurata77stelle GitHubstabile - 2AttivoApri fonte ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubInstalla
Risorsacrescita misuratagit clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability33stelle GitHub+3 (+10.0 %)