ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

75

ツールランキング

Ranked by creation date, newest first; undated entries come last. 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1

    エージェント
    活動中

    面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  2. 2

    スキル
    活動中

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    55GitHubスター+27 (+96.4 %)
  3. 3

    エージェント
    活動中

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  4. 4
    活動中

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    github測定済み成長情報源を開く ↗

    3GitHubスター安定
  5. 5
    活動中

    Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude

    2GitHubスター安定
  6. 6
    活動中

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    github推定モメンタム情報源を開く ↗

    インストール git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    39GitHubスター
  7. 7
    活動中

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    github測定済み成長情報源を開く ↗

    22GitHubスター+2 (+10.0 %)
  8. 8
    活動中

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github測定済み成長情報源を開く ↗

    5GitHubスター+1 (+25.0 %)
  9. 9

    スキル
    活動中

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    424GitHubスター+170 (+66.9 %)
  10. 10

    エージェント
    活動中

    Local-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  11. 11

    スキル
    活動中

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHubスター+1 (+100.0 %)
  12. 12

    スキル
    活動中

    Research-backed, eval-driven skills for AI agents

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    142GitHubスター+18 (+14.5 %)
  13. 13
    活動中

    Make Claude write clearly, for everyone.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add stefanobaghino/simple-output-styles

    16GitHubスター安定
  14. 14

    その他
    活動中

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    github測定済み成長情報源を開く ↗

    140GitHubスター+24 (+20.7 %)
  15. 15

    エージェント
    活動中

    Deterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  16. 16

    スキル
    活動中

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHubスター+1 (+0.76 %)
  17. 17

    MCP
    活動中

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    github測定済み成長情報源を開く ↗

    2GitHubスター安定
  18. 18
    活動中

    Vendor-neutral OpenTelemetry tracing for A2A agents and MCP services, with W3C context propagation and privacy-safe telemetry.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  19. 19
    活動中

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  20. 20
    活動中

    Governed local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.

    mcp測定済み成長情報源を開く ↗

    インストール claude mcp add ai-guardian -- uvx ai-guardian-aiops

    0GitHubスター安定
  21. 21

    その他
    活動中

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    github測定済み成長情報源を開く ↗

    782GitHubスター+1 (+0.13 %)
  22. 22

    スキル
    活動中

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51GitHubスター-5 (-8.9 %)

学習・参考リソース

Ranked by creation date, newest first; undated entries come last. これらは個別に参照でき、主要ランキングには含まれません。

  1. 1

    Agentic_AI_Engineer

    エージェント
    活動中情報源を開く ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubリソース測定済み成長
    18GitHubスター安定
  2. 2

    trigger_tree

    スキル
    活動中情報源を開く ↗

    Documentation-discovery telemetry for Claude Code — heat/cold maps, health grade, evidence-backed router fixes. 100% local, zero tokens. /tt

    github

    インストール /plugin marketplace add Hedde/trigger_tree

    リソース測定済み成長
    14GitHubスター+1 (+7.7 %)
  3. 3
    活動中情報源を開く ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    インストール git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    リソース測定済み成長
    32GitHubスター+2 (+6.7 %)