Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

65

Classifica degli strumenti

Ranked by creation date, newest first; undated entries come last. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1

    Agente
    Attivo

    面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。

    githubcrescita misurataApri fonte ↗

    0stelle GitHubstabile
  2. 2

    Skill
    Attivo

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    55stelle GitHub+27 (+96.4 %)
  3. 3

    Agente
    Attivo

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    githubcrescita misurataApri fonte ↗

    1stelle GitHubstabile
  4. 4
    Attivo

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    githubcrescita misurataApri fonte ↗

    3stelle GitHubstabile
  5. 5
    Attivo

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2stelle GitHubstabile
  6. 6
    Attivo

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    githubmomentum stimatoApri fonte ↗

    Installa git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    39stelle GitHub
  7. 7
    Attivo

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    githubcrescita misurataApri fonte ↗

    22stelle GitHub+2 (+10.0 %)
  8. 8
    Attivo

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    githubcrescita misurataApri fonte ↗

    5stelle GitHub+1 (+25.0 %)
  9. 9

    Skill
    Attivo

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    424stelle GitHub+170 (+66.9 %)
  10. 10

    Agente
    Attivo

    Local-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.

    githubcrescita misurataApri fonte ↗

    0stelle GitHubstabile
  11. 11

    Skill
    Attivo

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2stelle GitHub+1 (+100.0 %)
  12. 12

    Skill
    Attivo

    Research-backed, eval-driven skills for AI agents

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    142stelle GitHub+18 (+14.5 %)
  13. 13
    Attivo

    Make Claude write clearly, for everyone.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add stefanobaghino/simple-output-styles

    16stelle GitHubstabile
  14. 14
    Attivo

    Deterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.

    githubcrescita misurataApri fonte ↗

    1stelle GitHubstabile
  15. 15
    Attivo

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132stelle GitHub+1 (+0.76 %)
  16. 16

    MCP
    Attivo

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    githubcrescita misurataApri fonte ↗

    2stelle GitHubstabile
  17. 17
    Attivo

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    githubcrescita misurataApri fonte ↗

    1stelle GitHubstabile
  18. 18
    Attivo

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    githubcrescita misurataApri fonte ↗

    0stelle GitHubstabile
  19. 19
    Attivo

    Governed local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.

    mcpcrescita misurataApri fonte ↗

    Installa claude mcp add ai-guardian -- uvx ai-guardian-aiops

    0stelle GitHubstabile
  20. 20
    Attivo

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51stelle GitHub-5 (-8.9 %)
  21. 21

    Altro
    Attivo

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    githubcrescita misurataApri fonte ↗

    18stelle GitHubstabile

Risorse didattiche e di riferimento

Ranked by creation date, newest first; undated entries come last. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubRisorsacrescita misurata
    18stelle GitHubstabile
  2. 2
    AttivoApri fonte ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    Installa /plugin marketplace add Hedde/trigger_tree

    Risorsacrescita misurata
    14stelle GitHub+1 (+7.7 %)
  3. 3
    AttivoApri fonte ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    Installa git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    Risorsacrescita misurata
    32stelle GitHub+2 (+6.7 %)
  4. 4
    AttivoApri fonte ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRisorsacrescita misurata
    77stelle GitHubstabile