ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

118

ツールランキング

情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1

    スキル
    活動中

    Open-source infrastructure that turns scattered SKILL.md files into curated, retrieval-ready agent-skill corpora—with retrieval and evaluation tooling included.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/EverMind-AI/SkillCorpus ~/.claude/skills/SkillCorpus

    360GitHubスター+168 (+87.5 %)
  2. 2

    スキル
    活動中

    Multi-tier framework for evaluating AI agent skills with quality gates, semantic overlap detection, synthetic evaluation dataset generation, and live agent evaluation that measures how skills affect agent behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/NVIDIA/SkillEvaluator ~/.claude/skills/SkillEvaluator

    363GitHubスター+58 (+19.0 %)
  3. 3

    スキル
    活動中

    Research-backed, eval-driven skills for AI agents

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ZSeven-W/craft-skills ~/.claude/skills/craft-skills

    138GitHubスター+21 (+17.9 %)
  4. 4

    スキル
    活動中

    Deletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT

    38GitHubスター+10 (+35.7 %)
  5. 5

    その他
    活動中

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    github測定済み成長情報源を開く ↗

    24 710GitHubスター+137 (+0.56 %)
  6. 6

    その他
    活動中

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    github測定済み成長情報源を開く ↗

    27 601GitHubスター+128 (+0.47 %)
  7. 7

    その他
    活動中

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    github測定済み成長情報源を開く ↗

    21 594GitHubスター+119 (+0.55 %)
  8. 8

    スキル
    活動中

    YAO = Yielding AI Outcomes. A rigorous engineering, evaluation, governance, and portability system for reusable agent skills.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/yaojingang/yao-meta-skill ~/.claude/skills/yao-meta-skill

    2 565GitHubスター+58 (+2.3 %)
  9. 9

    MCP
    活動中

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    github測定済み成長情報源を開く ↗

    2GitHubスター+1 (+100.0 %)
  10. 10

    スキル
    活動中

    Production patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack

    2GitHubスター+1 (+100.0 %)
  11. 11

    エージェント
    活動中

    A local A2A event and continuity service for agent applications, with structured history, provenance-safe transcripts, literal search, and natural-language queries.

    github測定済み成長情報源を開く ↗

    1GitHubスター+1
  12. 12

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    146GitHubスター+13 (+9.8 %)
  13. 13

    エージェント
    活動中

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    github測定済み成長情報源を開く ↗

    27 759GitHubスター+78 (+0.28 %)
  14. 14

    エージェント
    活動中

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    github測定済み成長情報源を開く ↗

    627GitHubスター+22 (+3.6 %)
  15. 15

    その他
    活動中

    The fastest path to AI-powered full stack observability, even for lean teams.

    github測定済み成長情報源を開く ↗

    80 382GitHubスター+80 (+0.10 %)
  16. 16

    ライブラリ
    活動中

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    github測定済み成長情報源を開く ↗

    23 741GitHubスター+55 (+0.23 %)
  17. 17

    その他
    活動中

    the LLM vulnerability scanner

    github測定済み成長情報源を開く ↗

    9 079GitHubスター+46 (+0.51 %)
  18. 18
    活動中

    Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator

    2 358GitHubスター+29 (+1.2 %)
  19. 19

    その他
    活動中

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    github測定済み成長情報源を開く ↗

    31 983GitHubスター+48 (+0.15 %)
  20. 20

    その他
    活動中

    A high-performance observability data pipeline.

    github測定済み成長情報源を開く ↗

    22 497GitHubスター+41 (+0.18 %)
  21. 21
    活動中

    A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.

    github測定済み成長情報源を開く ↗

    1 752GitHubスター+22 (+1.3 %)
  22. 22

    スキル
    活動中

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 269GitHubスター+22 (+0.98 %)
  23. 23
    活動中

    Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.

    github測定済み成長情報源を開く ↗

    5GitHubスター+1 (+25.0 %)
  24. 24

    MCP
    休止中

    Open-source Python framework to deploy AI agents via HTTP, A2A, and MCP with built-in observability

    github測定済み成長情報源を開く ↗

    5GitHubスター+1 (+25.0 %)
  25. 25
    活動中

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    github測定済み成長情報源を開く ↗

    9 045GitHubスター+28 (+0.31 %)