Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
105
Classifica degli strumenti
Ranked by normalized popularity across sources. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1105stelle GitHub+1 (+0.96 %)
- 2Attivo
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97stelle GitHub+4 (+4.3 %) - 3Dormiente
anti-lie
SkillDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89stelle GitHubstabile - 4Attivo
MCP server for Langfuse LLM observability — trace and observation analysis.
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add langfuse -- npx langfuse-observability-mcp-server78stelle GitHubstabile - 5Attivo
skill-kit
Skilllocal-first analytics for AI agent skills
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77stelle GitHub+1 (+1.3 %) - 6Attivo
skill-eval-harness
SkillAgent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73stelle GitHub+4 (+5.8 %) - 7Attivo
AgentX-Python
AltroAgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
githubcrescita misurataApri fonte ↗
68stelle GitHubstabile - 8Attivo
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55stelle GitHub+27 (+96.4 %) - 9Attivo
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51stelle GitHub-5 (-8.9 %) - 10Attivo
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
githubcrescita misurataApri fonte ↗
50stelle GitHub+2 (+4.2 %) - 11Attivo
arize-skills
SkillAgent skills for Arize — datasets, experiments, and traces via the ax CLI
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47stelle GitHubstabile - 12Attivo
cap-evolve
MCPOptimize any AI agent’s skills, tools/MCP, and prompts against your own evals.
githubcrescita misurataApri fonte ↗
47stelle GitHub+1 (+2.2 %) - 13Attivo
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubmomentum stimatoApri fonte ↗
Installa
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite39stelle GitHub— - 14Attivo
claudestat
SkillReal-time execution trace and cost intelligence for Claude Code
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add DeibyGS/claudestat34stelle GitHub+1 (+3.0 %) - 15Attivo
opus-fable-playbook
SkillMake Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add rennf93/opus-fable-playbook33stelle GitHubstabile - 16Dormiente
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27stelle GitHubstabile - 17Dormiente
astragraph
AgentePolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
githubcrescita misurataApri fonte ↗
26stelle GitHubstabile - 18Attivo
Sentry instrumentation skill for system-behavior tracking
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24stelle GitHubstabile - 19Attivo
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
githubcrescita misurataApri fonte ↗
22stelle GitHub+2 (+10.0 %) - 20Attivo
Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.
githubcrescita misurataApri fonte ↗
22stelle GitHub+1 (+4.8 %) - 21Attivo
untell
AltroAI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
githubcrescita misurataApri fonte ↗
18stelle GitHubstabile - 22Attivo
rashomon
SkillMeasure prompt and skill improvements with blind A/B comparison.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add shinpr/rashomon18stelle GitHubstabile
Risorse didattiche e di riferimento
Ranked by normalized popularity across sources. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.
- 1AttivoApri fonte ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRisorsacrescita misurata77stelle GitHubstabile - 2AttivoApri fonte ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubInstalla
Risorsacrescita misuratagit clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability32stelle GitHub+2 (+6.7 %) - 3AttivoApri fonte ↗
Agentic_AI_Engineer
AgenteMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubRisorsacrescita misurata18stelle GitHubstabile