Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
65
Classifica degli strumenti
Ranked by creation date, newest first; undated entries come last. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1Attivo
CodeFlow
Agente面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。
githubcrescita misurataApri fonte ↗
0stelle GitHubstabile - 2Attivo
deslop-GPT
SkillDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT55stelle GitHub+27 (+96.4 %) - 3Attivo
dsh-plugins
AgenteGeneric DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.
githubcrescita misurataApri fonte ↗
1stelle GitHubstabile - 4Attivo
Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…
githubcrescita misurataApri fonte ↗
3stelle GitHubstabile - 5Attivo
Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude2stelle GitHubstabile - 6Attivo
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubmomentum stimatoApri fonte ↗
Installa
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite39stelle GitHub— - 7Attivo
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
githubcrescita misurataApri fonte ↗
22stelle GitHub+2 (+10.0 %) - 8Attivo
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
githubcrescita misurataApri fonte ↗
5stelle GitHub+1 (+25.0 %) - 9Attivo
SkillCorpus
SkillOpen-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus424stelle GitHub+170 (+66.9 %) - 10Attivo
evoagent-os
AgenteLocal-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.
githubcrescita misurataApri fonte ↗
0stelle GitHubstabile - 11Attivo
agent-stack
SkillProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2stelle GitHub+1 (+100.0 %) - 12Attivo
craft-skills
SkillResearch-backed, eval-driven skills for AI agents
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills142stelle GitHub+18 (+14.5 %) - 13Attivo
simple-output-styles
SkillMake Claude write clearly, for everyone.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add stefanobaghino/simple-output-styles16stelle GitHubstabile - 14Attivo
boundary-bench
AgenteDeterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.
githubcrescita misurataApri fonte ↗
1stelle GitHubstabile - 15Attivo
adlc-team-skills
Skill🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132stelle GitHub+1 (+0.76 %) - 16Attivo
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubcrescita misurataApri fonte ↗
2stelle GitHubstabile - 17Attivo
a2a-otel-kit
MCPVendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.
githubcrescita misurataApri fonte ↗
1stelle GitHubstabile - 18Attivo
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
githubcrescita misurataApri fonte ↗
0stelle GitHubstabile - 19Attivo
AI Guardian
MCPGoverned local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.
mcpcrescita misurataApri fonte ↗
Installa
claude mcp add ai-guardian -- uvx ai-guardian-aiops0stelle GitHubstabile - 20Attivo
seo-skill-bench
SkillOpen benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51stelle GitHub-5 (-8.9 %) - 21Attivo
untell
AltroAI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
githubcrescita misurataApri fonte ↗
18stelle GitHubstabile
Risorse didattiche e di riferimento
Ranked by creation date, newest first; undated entries come last. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.
- 1AttivoApri fonte ↗
Agentic_AI_Engineer
AgenteMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubRisorsacrescita misurata18stelle GitHubstabile - 2AttivoApri fonte ↗
trigger_tree
SkillDocumentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt
githubInstalla
Risorsacrescita misurata/plugin marketplace add Hedde/trigger_tree14stelle GitHub+1 (+7.7 %) - 3AttivoApri fonte ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubInstalla
Risorsacrescita misuratagit clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability32stelle GitHub+2 (+6.7 %) - 4AttivoApri fonte ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRisorsacrescita misurata77stelle GitHubstabile