ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
119
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
Claude Code skills where every entry ships receipts — accuracy-gated benchmarks against baseline and placebo, rejects published
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/sjh9714/skill-receipts ~/.claude/skills/skill-receipts2GitHubスター安定 - 2休止中
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27GitHubスター安定 - 3活動中
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
github測定済み成長情報源を開く ↗
133GitHubスター+4 (+3.1 %) - 4活動中
galdor
その他A Go-native framework for LLM agents, with OpenTelemetry observability built in.
github測定済み成長情報源を開く ↗
10GitHubスター安定 - 5休止中
CustoFlow
その他Multi-agent customer support system with Google ADK & Gemini 2.5 Flash Lite. Kaggle capstone demonstrating 11+ concepts. Automates 80%+ queries, <10s response time.
github測定済み成長情報源を開く ↗
2GitHubスター安定 - 6休止中
astragraph
エージェントPolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
github測定済み成長情報源を開く ↗
26GitHubスター安定 - 7活動中
Amazon Bedrock AgentCore enterprise platform accelerator with AWS CDK, Terraform organization guardrails, MCP/A2A agents, security, memory, and observability.
github測定済み成長情報源を開く ↗
6GitHubスター+1 (+20.0 %) - 8休止中
agentanvil
エージェントContract-based testing for LLM agents: hybrid evaluation, multi-agent + A2A, deterministic replay.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 9活動中
🐙 ADLC Team Skills — Agentic SDLC for Engineering Teams
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tikalk/adlc-team-skills ~/.claude/skills/adlc-team-skills132GitHubスター安定 - 10休止中
Agents-eval
エージェントA Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.
github測定済み成長情報源を開く ↗
2GitHubスター安定 - 11活動中
Sentry instrumentation skill for system-behavior tracking
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24GitHubスター安定 - 12活動中
8 MCP SMB products — standalone AI servers for SMBs: Guardrails, FinOps, Observability, Router, Trust Score, Memory, ThinkSecure, A2A Lite. 37 tools, TNC credits billing.
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 13活動中
ai-dev-stack
スキルProduction-grade AI coding rules for Cursor and Claude Code. 15 rules + 9 doc templates + skills + agents + MCP setup. Drop into any project.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/aiagentwithdhruv/ai-dev-stack ~/.claude/skills/ai-dev-stack10GitHubスター安定 - 14休止中
Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add prime-radiant-inc/parallel-adversarial-review17GitHubスター安定 - 15105GitHubスター安定
- 16活動中
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 17活動中
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
github測定済み成長情報源を開く ↗
2GitHubスター安定 - 18活動中
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
github測定済み成長情報源を開く ↗
22GitHubスター安定 - 19活動中
adl-cli
その他A command-line tool to scaffold and manage enterprise-ready AI Agents powered by the A2A (Agent-to-Agent) protocol
github測定済み成長情報源を開く ↗
14GitHubスター安定 - 20活動中
A test runner for agentskills.io-style AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/darkrishabh/agent-skills-eval ~/.claude/skills/agent-skills-eval724GitHubスター+5 (+0.70 %) - 21活動中
pyxen
エージェントA lightweight Python library that decouples agentic runtime from applications it builds
github測定済み成長情報源を開く ↗
1GitHubスター安定 - 22活動中
Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude2GitHubスター安定 - 23活動中
Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.
github測定済み成長情報源を開く ↗
22GitHubスター安定 - 24活動中
deslop-GPT
スキルDeletion-first Agent Skill for removing test bloat, verification theater, and speculative fallbacks while preserving behavior.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/MrZoyo/deslop-GPT ~/.claude/skills/deslop-GPT109GitHubスター+54 (+98.2 %) - 25活動中
bakeoff
スキルTurn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add CoriChui/bakeoff10GitHubスター安定