ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

74

ツールランキング

Ranked by normalized popularity across sources. 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    156GitHubスター+19 (+13.9 %)
  2. 2

    スキル
    活動中

    Research-backed, eval-driven skills for AI agents

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    142GitHubスター+18 (+14.5 %)
  3. 3

    その他
    活動中

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    github測定済み成長情報源を開く ↗

    140GitHubスター+24 (+20.7 %)
  4. 4

    スキル
    活動中

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHubスター+1 (+0.76 %)
  5. 5

    スキル
    活動中

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    132GitHubスター+4 (+3.1 %)
  6. 6
    活動中

    74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…

    github測定済み成長情報源を開く ↗

    129GitHubスター-4 (-3.0 %)
  7. 7
    活動中

    A Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability

    github測定済み成長情報源を開く ↗

    105GitHubスター安定
  8. 8
    活動中

    Self improving agents through iterations

    github測定済み成長情報源を開く ↗

    105GitHubスター+1 (+0.96 %)
  9. 9
    活動中

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    97GitHubスター+4 (+4.3 %)
  10. 10

    スキル
    活動中

    local-first analytics for AI agent skills

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77GitHubスター+1 (+1.3 %)
  11. 11

    スキル
    活動中

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    73GitHubスター+4 (+5.8 %)
  12. 12

    その他
    活動中

    AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.

    github測定済み成長情報源を開く ↗

    68GitHubスター安定
  13. 13

    スキル
    活動中

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    55GitHubスター+27 (+96.4 %)
  14. 14

    スキル
    活動中

    Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench

    51GitHubスター-5 (-8.9 %)
  15. 15

    MCP
    活動中

    Open-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.

    github測定済み成長情報源を開く ↗

    50GitHubスター+2 (+4.2 %)
  16. 16

    スキル
    活動中

    Agent skills for Arize — datasets, experiments, and traces via the ax CLI

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills

    47GitHubスター安定
  17. 17
    活動中

    Optimize any AI agent’s skills, tools/MCP, and prompts against your own evals.

    github測定済み成長情報源を開く ↗

    47GitHubスター+1 (+2.2 %)
  18. 18
    活動中

    Mini Program Engineering Skill Suite is an Agent Skill suite for evidence‑first mini‑program development. It helps agents bring WeChat or other mini‑program projects from vague intent to reliable engineering work: project intake, product…

    github推定モメンタム情報源を開く ↗

    インストール git clone https://github.com/NocodeMrLi/mini-program-engineering-skill-suite ~/.claude/skills/mini-program-engineering-skill-suite

    39GitHubスター
  19. 19

    スキル
    活動中

    Real-time execution trace and cost intelligence for Claude Code

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add DeibyGS/claudestat

    34GitHubスター+1 (+3.0 %)
  20. 20
    活動中

    Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.

    github測定済み成長情報源を開く ↗

    22GitHubスター+2 (+10.0 %)
  21. 21

    その他
    活動中

    AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…

    github測定済み成長情報源を開く ↗

    18GitHubスター安定
  22. 22

    スキル
    活動中

    Measure prompt and skill improvements with blind A/B comparison.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add shinpr/rashomon

    18GitHubスター安定

学習・参考リソース

Ranked by normalized popularity across sources. これらは個別に参照でき、主要ランキングには含まれません。

  1. 1
    活動中情報源を開く ↗

    Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.

    githubリソース測定済み成長
    77GitHubスター安定
  2. 2
    活動中情報源を開く ↗

    50+ curated LLM observability tools PLUS 26 Agent Skills (several with runnable, unit-tested scripts) to build, evaluate, debug, secure & monitor reliable LLM apps. Tracing, evals, guardrails, LLMOps.

    github

    インストール git clone https://github.com/ContextJet-ai/awesome-llm-observability ~/.claude/skills/awesome-llm-observability

    リソース測定済み成長
    32GitHubスター+2 (+6.7 %)
  3. 3

    Agentic_AI_Engineer

    エージェント
    活動中情報源を開く ↗

    My complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.

    githubリソース測定済み成長
    18GitHubスター安定