ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
121
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1休止中
anti-lie
スキルDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHubスター安定 - 2活動中
agent-stack
スキルProduction patterns for agent orchestrators, harnesses, evals, MCP/A2A interoperability, memory, provider routing, and LLM usage metering.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ssheleg/agent-stack ~/.claude/skills/agent-stack2GitHubスター+1 (+100.0 %) - 3活動中
untell
その他AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
github測定済み成長情報源を開く ↗
18GitHubスター安定 - 4活動中
Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.
github測定済み成長情報源を開く ↗
22GitHubスター+1 (+4.8 %) - 5活動中
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
github測定済み成長情報源を開く ↗
22GitHubスター+2 (+10.0 %) - 6活動中
deslop-GPT
スキルDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT87GitHubスター+59 (+210.7 %) - 7休止中
agenttap
エージェントReal-time debugging proxy for Agent2Agent (A2A) multi-agent systems
github測定済み成長情報源を開く ↗
3GitHubスター安定 - 8休止中
agentanvil
エージェントContract-based testing for LLM agents: hybrid evaluation, multi-agent + A2A, deterministic replay.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 9活動中
Sentry instrumentation skill for system-behavior tracking
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24GitHubスター安定 - 10休止中
A self-improving harness router for Claude Code.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add SeongwoongCho/adaptive-harness8GitHubスター安定 - 11活動中
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add langfuse -- npx langfuse-observability-mcp-server80GitHubスター+2 (+2.6 %) - 12活動中
axiom
スキルAxiom is a curated marketplace of shared plugins for Claude Code and Codex.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/netopsengineer/axiom ~/.claude/skills/axiom5GitHubスター安定 - 13休止中
astragraph
エージェントPolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
github測定済み成長情報源を開く ↗
26GitHubスター安定 - 14活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 763GitHubスター+15 (+0.86 %) - 15活動中
CodeFlow
エージェント面向长程研发任务的 Java Agent Harness,支持 A2A 跨语言协作、可恢复执行、上下文工程与 Eval 驱动开发。
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 16休止中
kaggle-capstone-ai-agent
エージェントA safety-first multi-agent mental health companion with real-time distress tracking, triple-layer guardrails, and evidence-based grounding techniques. Built for Kaggle × Google Agents Intensive 2025 Capstone (Agents for Good Track)
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 17活動中
OpenJudge
スキルOpenJudge: A Unified Framework for Holistic Evaluation and Quality Rewards
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/agentscope-ai/OpenJudge ~/.claude/skills/OpenJudge816GitHubスター+11 (+1.4 %) - 18休止中
gt8004-sdk
その他Official Python SDK for GT8004 — AI agent observability with MCP, A2A, x402 payment tracking. FastAPI, Flask, FastMCP middleware included.
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 19休止中
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27GitHubスター安定 - 20活動中
skill-kit
スキルlocal-first analytics for AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHubスター+1 (+1.3 %) - 21活動中
8 MCP SMB products — standalone AI servers for SMBs: Guardrails, FinOps, Observability, Router, Trust Score, Memory, ThinkSecure, A2A Lite. 37 tools, TNC credits billing.
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 22休止中
agentops
エージェントAgentOps: Multi-agent infrastructure remediation platform. A2A protocol, agent coordination, HITL approval, auto-rollback.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 23活動中
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
github測定済み成長情報源を開く ↗
339GitHubスター+81 (+31.4 %) - 24活動中
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHubスター+2 (+2.8 %) - 25休止中
Local open-source dev tool to debug, secure, and evaluate LLM agents. Provides static analysis, dynamic security checks, and runtime monitoring - integrates with Cursor and Claude Code.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add cylestio/agent-inspector9GitHubスター安定