ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

123

ツールランキング

情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1

    スキル
    活動中

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    424GitHubスター+204 (+92.7 %)
  2. 2

    スキル
    活動中

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    389GitHubスター+63 (+19.3 %)
  3. 3

    スキル
    活動中

    Research-backed, eval-driven skills for AI agents

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    142GitHubスター+21 (+17.4 %)
  4. 4

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    156GitHubスター+22 (+16.4 %)
  5. 5

    スキル
    活動中

    Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness

    73GitHubスター+4 (+5.8 %)
  6. 6

    スキル
    活動中

    Skills, prompts, and instructions for building AI agents on top of Dynatrace production context

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Dynatrace/dynatrace-for-ai ~/.claude/skills/dynatrace-for-ai

    132GitHubスター+6 (+4.8 %)
  7. 7
    活動中

    Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills

    97GitHubスター+4 (+4.3 %)
  8. 8

    エージェント
    活動中

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    github測定済み成長情報源を開く ↗

    634GitHubスター+26 (+4.3 %)
  9. 9

    スキル
    活動中

    YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill

    2 577GitHubスター+54 (+2.1 %)
  10. 10

    スキル
    活動中

    A test runner for agentskills.io-style AI agent skills

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval

    719GitHubスター+15 (+2.1 %)
  11. 11
    活動中

    A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

    github測定済み成長情報源を開く ↗

    1 759GitHubスター+24 (+1.4 %)
  12. 12

    スキル
    活動中

    local-first analytics for AI agent skills

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit

    77GitHubスター+1 (+1.3 %)
  13. 13
    活動中

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator

    2 367GitHubスター+30 (+1.3 %)
  14. 14

    スキル
    活動中

    🪢 Langfuse documentation -- Langfuse is the open source LLM Engineering Platform. Observability, evals, prompt management, playground and metrics to debug and improve LLM apps

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/langfuse/langfuse-docs ~/.claude/skills/langfuse-docs

    239GitHubスター+3 (+1.3 %)
  15. 15

    スキル
    活動中

    OpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge

    809GitHubスター+10 (+1.3 %)
  16. 16
    活動中

    Self improving agents through iterations

    github測定済み成長情報源を開く ↗

    105GitHubスター+1 (+0.96 %)
  17. 17

    スキル
    活動中

    Dashboard for monitoring claude code sessions.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add JayantDevkar/claude-code-karma

    323GitHubスター+3 (+0.94 %)
  18. 18

    スキル
    活動中

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 272GitHubスター+19 (+0.84 %)
  19. 19

    スキル
    活動中

    🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills

    132GitHubスター+1 (+0.76 %)
  20. 20

    その他
    活動中

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    github測定済み成長情報源を開く ↗

    21 608GitHubスター+115 (+0.54 %)
  21. 21

    その他
    活動中

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    github測定済み成長情報源を開く ↗

    24 738GitHubスター+131 (+0.53 %)
  22. 22

    その他
    活動中

    the LLM vulnerability scanner

    github測定済み成長情報源を開く ↗

    9 079GitHubスター+43 (+0.48 %)
  23. 23

    その他
    活動中

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    github測定済み成長情報源を開く ↗

    27 624GitHubスター+121 (+0.44 %)
  24. 24

    スキル
    活動中

    A skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge

    884GitHubスター+3 (+0.34 %)
  25. 25
    活動中

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    github測定済み成長情報源を開く ↗

    9 052GitHubスター+28 (+0.31 %)