Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

120

Classifica degli strumenti

Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1

    Agente
    Attivo

    面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。

    githubcrescita misurataApri fonte ↗

    0stelle GitHubstabile
  2. 2
    Attivo

    Claude Code skills where every entry ships receipts — accuracy-gated benchmarks against baseline and placebo, rejects published

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/sjh9714/skill-receipts ~/.claude/skills/skill-receipts

    2stelle GitHubstabile
  3. 3
    Attivo

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    githubcrescita misurataApri fonte ↗

    22stelle GitHubstabile
  4. 4
    Attivo

    A test runner for agentskills.io-style AI agent skills

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval

    724stelle GitHub+11 (+1.5 %)
  5. 5

    MCP
    Attivo

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    githubcrescita misurataApri fonte ↗

    2stelle GitHubstabile
  6. 6

    Altro
    Attivo

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    githubcrescita misurataApri fonte ↗

    18stelle GitHub-1 (-5.3 %)
  7. 7

    Agente
    Attivo

    Multi-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.

    githubcrescita misurataApri fonte ↗

    3stelle GitHubstabile
  8. 8

    Skill
    Attivo

    Measure prompt and skill improvements with blind A/B comparison.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add shinpr/rashomon

    18stelle GitHubstabile
  9. 9

    Skill
    Attivo

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 036stelle GitHub+1 (+0.006 %)
  10. 10

    Agente
    Dormiente

    Real-time debugging proxy for Agent2Agent (A2A) multi-agent systems

    githubcrescita misurataApri fonte ↗

    3stelle GitHubstabile
  11. 11

    Skill
    Attivo

    Glanceable Claude Code and Codex state in your terminal tabs: white=idle, blue=working, orange=waiting. Multi-terminal (iTerm2, WezTerm, AI Power Term), a live status bar with Anthropic usage limits, /sfl and /nil window save-and-restore,…

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/wasulajr/headsup ~/.claude/skills/headsup

    1stelle GitHubstabile
  12. 12
    Attivo

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    47stelle GitHub+8 (+20.5 %)
  13. 13
    Attivo

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform

    2 379stelle GitHub+3 (+0.13 %)
  14. 14
    Attivo

    Run many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases

    1stelle GitHubstabile

Risorse didattiche e di riferimento

Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    Installa /plugin marketplace add Hedde/trigger_tree

    Risorsacrescita misurata
    14stelle GitHubstabile
  2. 2
    AttivoApri fonte ↗

    Handbook técnico aberto sobre engenharia de IA em produção: ML tradicional, LLMs, RAG, agentes, segurança, observabilidade, FinOps e deployment.

    githubRisorsacrescita misurata
    1stelle GitHubstabile
  3. 3
    AttivoApri fonte ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRisorsacrescita misurata
    78stelle GitHub+1 (+1.3 %)
  4. 4
    AttivoApri fonte ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    Installa git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    Risorsacrescita misurata
    33stelle GitHub+2 (+6.5 %)
  5. 5
    AttivoApri fonte ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubRisorsacrescita misurata
    18stelle GitHubstabile
  6. 6

    tunelab

    Skill
    AttivoApri fonte ↗

    Claude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.

    github

    Installa /plugin marketplace add rchaz/tunelab

    Risorsacrescita misurata
    6stelle GitHubstabile