Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
75
Classifica degli strumenti
Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1Attivo
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus360stelle GitHub+168 (+87.5 %) - 2Attivo
SkillEvaluator
SkillMulti-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator363stelle GitHub+58 (+19.0 %) - 3Attivo
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills138stelle GitHub+21 (+17.9 %) - 4Attivo
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT38stelle GitHub+10 (+35.7 %) - 5Attivo
promptfoo
AltroTest your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
githubcrescita misurataApri fonte ↗
24 710stelle GitHub+137 (+0.56 %) - 6Attivo
mastra
AltroMastra is the modern TypeScript framework for AI-powered applications and agents.
githubcrescita misurataApri fonte ↗
27 601stelle GitHub+128 (+0.47 %) - 7Attivo
openobserve
AltroOpen source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
githubcrescita misurataApri fonte ↗
21 594stelle GitHub+119 (+0.55 %) - 8Attivo
yao-meta-skill
SkillYAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill2 565stelle GitHub+58 (+2.3 %) - 9Attivo
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubcrescita misurataApri fonte ↗
2stelle GitHub+1 (+100.0 %) - 10Attivo
agent-stack
SkillProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2stelle GitHub+1 (+100.0 %) - 11Attivo
agent-kernel
AgenteThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubcrescita misurataApri fonte ↗
146stelle GitHub+13 (+9.8 %) - 12Attivo
mlflow
AgenteThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
githubcrescita misurataApri fonte ↗
27 759stelle GitHub+78 (+0.28 %) - 13Attivo
databuff
AgenteDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
githubcrescita misurataApri fonte ↗
627stelle GitHub+22 (+3.6 %) - 14Attivo
netdata
AltroThe fastest path to AI-powered full stack observability, even for lean teams.
githubcrescita misurataApri fonte ↗
80 382stelle GitHub+80 (+0.10 %) - 15Attivo
prefect
LibreriaPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
githubcrescita misurataApri fonte ↗
23 741stelle GitHub+55 (+0.23 %) - 169 079stelle GitHub+46 (+0.51 %)
- 17Attivo
agent-skill-creator
SkillBuild tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 358stelle GitHub+29 (+1.2 %) - 18Attivo
signoz
AltroSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubcrescita misurataApri fonte ↗
31 983stelle GitHub+48 (+0.15 %) - 1922 497stelle GitHub+41 (+0.18 %)
- 20Attivo
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
githubcrescita misurataApri fonte ↗
1 752stelle GitHub+22 (+1.3 %) - 21Attivo
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
githubcrescita misurataApri fonte ↗
5stelle GitHub+1 (+25.0 %) - 22Attivo
openstatus
MCP🫖 Status page with uptime monitoring & API monitoring as code 🫖
githubcrescita misurataApri fonte ↗
9 045stelle GitHub+28 (+0.31 %) - 23Attivo
agent-skills-eval
SkillA test runner for agentskills.io-style AI agent skills
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval713stelle GitHub+12 (+1.7 %) - 24Attivo
cilium
AltroeBPF-based Networking, Security, and Observability
githubcrescita misurataApri fonte ↗
25 035stelle GitHub+26 (+0.10 %) - 25Attivo
untell
AltroAI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
githubcrescita misurataApri fonte ↗
19stelle GitHub+2 (+11.8 %)