Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

117

Classifica degli strumenti

Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1

    Skill
    Attivo

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    424stelle GitHub+170 (+66.9 %)
  2. 2

    Altro
    Attivo

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    githubcrescita misurataApri fonte ↗

    24 768stelle GitHub+142 (+0.58 %)
  3. 3

    Altro
    Attivo

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    githubcrescita misurataApri fonte ↗

    27 658stelle GitHub+128 (+0.46 %)
  4. 4

    Altro
    Attivo

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    githubcrescita misurataApri fonte ↗

    21 615stelle GitHub+96 (+0.45 %)
  5. 5

    Altro
    Attivo

    The fastest path to AI-powered full stack observability, even for lean teams.

    githubcrescita misurataApri fonte ↗

    80 412stelle GitHub+85 (+0.11 %)
  6. 6

    Agente
    Attivo

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    githubcrescita misurataApri fonte ↗

    27 783stelle GitHub+82 (+0.30 %)
  7. 7

    Libreria
    Attivo

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    githubcrescita misurataApri fonte ↗

    23 766stelle GitHub+66 (+0.28 %)
  8. 8
    Attivo

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    389stelle GitHub+57 (+17.2 %)
  9. 9

    Altro
    Attivo

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    githubcrescita misurataApri fonte ↗

    32 003stelle GitHub+50 (+0.16 %)
  10. 10

    Altro
    Attivo

    A high-performance observability data pipeline.

    githubcrescita misurataApri fonte ↗

    22 509stelle GitHub+40 (+0.18 %)
  11. 11
    Attivo

    YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill

    2 577stelle GitHub+35 (+1.4 %)
  12. 12

    Agente
    Attivo

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    githubcrescita misurataApri fonte ↗

    642stelle GitHub+32 (+5.2 %)
  13. 13
    Attivo

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    githubcrescita misurataApri fonte ↗

    9 056stelle GitHub+31 (+0.34 %)
  14. 14

    Altro
    Attivo

    the LLM vulnerability scanner

    githubcrescita misurataApri fonte ↗

    9 079stelle GitHub+31 (+0.34 %)
  15. 15

    Altro
    Attivo

    eBPF-based Networking, Security, and Observability

    githubcrescita misurataApri fonte ↗

    25 047stelle GitHub+29 (+0.12 %)
  16. 16
    Attivo

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator

    2 367stelle GitHub+28 (+1.2 %)
  17. 17

    Skill
    Attivo

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    55stelle GitHub+27 (+96.4 %)
  18. 18

    Altro
    Attivo

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    githubcrescita misurataApri fonte ↗

    140stelle GitHub+24 (+20.7 %)
  19. 19

    Agente
    Attivo

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    githubcrescita misurataApri fonte ↗

    156stelle GitHub+19 (+13.9 %)
  20. 20

    Skill
    Attivo

    Research-backed, eval-driven skills for AI agents

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    142stelle GitHub+18 (+14.5 %)
  21. 21
    Attivo

    A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

    githubcrescita misurataApri fonte ↗

    1 759stelle GitHub+18 (+1.0 %)
  22. 22

    Skill
    Attivo

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 272stelle GitHub+16 (+0.71 %)
  23. 23
    Attivo

    A test runner for agentskills.io-style AI agent skills

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval

    719stelle GitHub+13 (+1.8 %)
  24. 24

    Agente
    Attivo

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    githubcrescita misurataApri fonte ↗

    12 068stelle GitHub+8 (+0.07 %)
  25. 25

    Skill
    Attivo

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 037stelle GitHub+8 (+0.05 %)