Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
123
Classifica degli strumenti
Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1Attivo
Flawless
AltroAI SRE AgenticOps for Kubernetes and cloud infrastructure.
githubcrescita misurataApri fonte ↗
782stelle GitHub+1 (+0.13 %) - 2Attivo
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
githubcrescita misurataApri fonte ↗
254stelle GitHub-3 (-1.2 %) - 3Attivo
idun-agent-platform
Skill🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform198stelle GitHubstabile - 4Attivo
rocketplaneIO
AltroSelf-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
githubcrescita misurataApri fonte ↗
131stelle GitHub-11 (-7.7 %) - 5Attivo
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
githubcrescita misurataApri fonte ↗
129stelle GitHub-10 (-7.2 %) - 6Attivo
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
githubcrescita misurataApri fonte ↗
105stelle GitHubstabile - 7Dormiente
anti-lie
SkillDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89stelle GitHubstabile - 8Attivo
MCP server for Langfuse LLM observability — trace and observation analysis.
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add langfuse -- npx langfuse-observability-mcp-server78stelle GitHubstabile - 9Attivo
AgentX-Python
AltroAgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
githubcrescita misurataApri fonte ↗
68stelle GitHubstabile - 10Attivo
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51stelle GitHub-5 (-8.9 %) - 11Attivo
arize-skills
SkillAgent skills for Arize — datasets, experiments, and traces via the ax CLI
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47stelle GitHubstabile - 12Attivo
opus-fable-playbook
SkillMake Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add rennf93/opus-fable-playbook33stelle GitHubstabile - 13Dormiente
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27stelle GitHubstabile - 14Dormiente
astragraph
AgentePolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
githubcrescita misurataApri fonte ↗
26stelle GitHubstabile - 15Attivo
Sentry instrumentation skill for system-behavior tracking
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24stelle GitHubstabile - 16Attivo
rashomon
SkillMeasure prompt and skill improvements with blind A/B comparison.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add shinpr/rashomon18stelle GitHubstabile - 17Attivo
Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add prime-radiant-inc/parallel-adversarial-review17stelle GitHubstabile - 18Dormiente
MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge17stelle GitHubstabile - 19Attivo
simple-output-styles
SkillMake Claude write clearly, for everyone.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add stefanobaghino/simple-output-styles16stelle GitHubstabile - 20Dormiente
eval-layer
AltroA Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
githubcrescita misurataApri fonte ↗
13stelle GitHubstabile - 21Attivo
Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace13stelle GitHubstabile - 22Attivo
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add mcp-server -- npx @spanlens/mcp-server12stelle GitHubstabile - 23Attivo
bakeoff
SkillTurn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add CoriChui/bakeoff10stelle GitHubstabile
Risorse didattiche e di riferimento
Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.
- 1AttivoApri fonte ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRisorsacrescita misurata77stelle GitHubstabile - 2AttivoApri fonte ↗
Agentic_AI_Engineer
AgenteMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubRisorsacrescita misurata18stelle GitHubstabile