Le descrizioni degli strumenti sono in inglese.
Osservabilità LLM
Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.
Utilizzo
Attività
Ordina per
120
Classifica degli strumenti
Classifica per crescita misurata e normalizzata tra le fonti. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.
- 1Attivo
CodeFlow
Agente面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。
githubcrescita misurataApri fonte ↗
0stelle GitHubstabile - 2Attivo
skill-receipts
SkillClaude Code skills where every entry ships receipts — accuracy-gated benchmarks against baseline and placebo, rejects published
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/sjh9714/skill-receipts ~/.claude/skills/skill-receipts2stelle GitHubstabile - 3Attivo
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
githubcrescita misurataApri fonte ↗
22stelle GitHubstabile - 4Attivo
agent-skills-eval
SkillA test runner for agentskills.io-style AI agent skills
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval724stelle GitHub+11 (+1.5 %) - 5Attivo
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubcrescita misurataApri fonte ↗
2stelle GitHubstabile - 6Attivo
untell
AltroAI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
githubcrescita misurataApri fonte ↗
18stelle GitHub-1 (-5.3 %) - 7Attivo
sre-on-call
AgenteMulti-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.
githubcrescita misurataApri fonte ↗
3stelle GitHubstabile - 8Attivo
rashomon
SkillMeasure prompt and skill improvements with blind A/B comparison.
githubcrescita misurataApri fonte ↗
Installa
/plugin marketplace add shinpr/rashomon18stelle GitHubstabile - 9Attivo
kubesphere
SkillThe container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere17 036stelle GitHub+1 (+0.006 %) - 10Dormiente
agenttap
AgenteReal-time debugging proxy for Agent2Agent (A2A) multi-agent systems
githubcrescita misurataApri fonte ↗
3stelle GitHubstabile - 11Attivo
headsup
SkillGlanceable Claude Code and Codex state in your terminal tabs: white=idle, blue=working, orange=waiting. Multi-terminal (iTerm2, WezTerm, AI Power Term), a live status bar with Anthropic usage limits, /sfl and /nil window save-and-restore,…
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/wasulajr/headsup ~/.claude/skills/headsup1stelle GitHubstabile - 12Attivo
Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite47stelle GitHub+8 (+20.5 %) - 13Attivo
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform2 379stelle GitHub+3 (+0.13 %) - 14Attivo
sc-prism-releases
SkillRun many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…
githubcrescita misurataApri fonte ↗
Installa
git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases1stelle GitHubstabile
Risorse didattiche e di riferimento
Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.
- 1AttivoApri fonte ↗
trigger_tree
SkillDocumentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt
githubInstalla
Risorsacrescita misurata/plugin marketplace add Hedde/trigger_tree14stelle GitHubstabile - 2AttivoApri fonte ↗
Handbook técnico aberto sobre engenharia de IA em produção: ML tradicional, LLMs, RAG, agentes, segurança, observabilidade, FinOps e deployment.
githubRisorsacrescita misurata1stelle GitHubstabile - 3AttivoApri fonte ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubRisorsacrescita misurata78stelle GitHub+1 (+1.3 %) - 4AttivoApri fonte ↗
50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.
githubInstalla
Risorsacrescita misuratagit clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability33stelle GitHub+2 (+6.5 %) - 5AttivoApri fonte ↗
Agentic_AI_Engineer
AgenteMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubRisorsacrescita misurata18stelle GitHubstabile - 6AttivoApri fonte ↗
tunelab
SkillClaude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.
githubInstalla
Risorsacrescita misurata/plugin marketplace add rchaz/tunelab6stelle GitHubstabile