Le descrizioni degli strumenti sono in inglese.

Osservabilità LLM

Monitoraggio, tracce, valutazione e qualità delle applicazioni LLM.

Utilizzo

Attività

Ordina per

116

Classifica degli strumenti

Classifica mista: la crescita misurata ha la precedenza. I valori grezzi mantengono la propria finestra; una variazione richiede almeno due rilevazioni in 7 giorni.

  1. 1

    Agente
    Dormiente

    A Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.

    githubcrescita misurataApri fonte ↗

    2stelle GitHubstabile
  2. 2
    Attivo

    A Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability

    githubcrescita misurataApri fonte ↗

    105stelle GitHubstabile
  3. 3

    Altro
    Dormiente

    Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.

    githubcrescita misurataApri fonte ↗

    2stelle GitHubstabile
  4. 4
    Attivo

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    githubcrescita misurataApri fonte ↗

    0stelle GitHubstabile
  5. 5

    Skill
    Attivo

    The container platform tailored for Kubernetes multi-cloud, datacenter, and edge management ⎈ 🖥 ☁️

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/kubesphere/kubesphere ~/.claude/skills/kubesphere

    17 035stelle GitHubstabile
  6. 6
    Dormiente

    Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.

    githubcrescita misurataApri fonte ↗

    Installa /plugin marketplace add prime-radiant-inc/parallel-adversarial-review

    17stelle GitHubstabile
  7. 7
    Attivo

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    45stelle GitHub+6 (+15.4 %)
  8. 8
    Attivo

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/FrancyJGLisboa/agent-skills-platform ~/.claude/skills/agent-skills-platform

    2 376stelle GitHubstabile
  9. 9
    Attivo

    Run many AIs on one board and keep control of all of it. Deterministic code decides who acts — never a model. A privacy floor keeps sensitive work on your machine, your own tests decide what counts as done, and every action lands on a…

    githubcrescita misurataApri fonte ↗

    Installa git clone https://github.com/sandhusukhdeep2/sc-prism-releases ~/.claude/skills/sc-prism-releases

    1stelle GitHubstabile
  10. 10

    Altro
    Attivo

    Open-source observability tool that uses AI agents to self-heal your software

    githubmomentum stimatoApri fonte ↗

    1 404stelle GitHub

Risorse didattiche e di riferimento

Classifica per crescita misurata e normalizzata tra le fonti. Queste risorse restano accessibili separatamente e non partecipano alla classifica principale.

  1. 1
    AttivoApri fonte ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubRisorsacrescita misurata
    18stelle GitHubstabile
  2. 2
    AttivoApri fonte ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubRisorsacrescita misurata
    77stelle GitHubstabile
  3. 3
    AttivoApri fonte ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    Installa git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    Risorsacrescita misurata
    33stelle GitHub+2 (+6.5 %)
  4. 4
    AttivoApri fonte ↗

    Handbook técnico aberto sobre engenharia de IA em produção: ML tradicional, LLMs, RAG, agentes, segurança, observabilidade, FinOps e deployment.

    githubRisorsacrescita misurata
    1stelle GitHubstabile
  5. 5
    AttivoApri fonte ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    Installa /plugin marketplace add Hedde/trigger_tree

    Risorsacrescita misurata
    14stelle GitHubstabile
  6. 6

    tunelab

    Skill
    AttivoApri fonte ↗

    Claude Code plugin for LLM fine-tuning, distillation, and evaluation — decide whether you need fine-tuning at all, distill your LLM logs into small local models (MLX/LoRA), evaluate with held-out discipline, and learn the why at every step.

    github

    Installa /plugin marketplace add rchaz/tunelab

    Risorsacrescita misurata
    6stelle GitHubstabile