ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
122
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
SkillForge
スキルA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHubスター-1 (-0.11 %) - 2活動中
🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform199GitHubスター安定 - 3活動中
Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
github測定済み成長情報源を開く ↗
131GitHubスター-4 (-3.0 %) - 4活動中
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
github測定済み成長情報源を開く ↗
129GitHubスター-4 (-3.0 %) - 5活動中
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
github測定済み成長情報源を開く ↗
105GitHubスター安定 - 6休止中
anti-lie
スキルDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHubスター安定 - 7活動中
AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
github測定済み成長情報源を開く ↗
68GitHubスター安定 - 8活動中
Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHubスター-5 (-8.9 %) - 9活動中
arize-skills
スキルAgent skills for Arize — datasets, experiments, and traces via the ax CLI
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47GitHubスター安定 - 10活動中
claudestat
スキルReal-time execution trace and cost intelligence for Claude Code
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add DeibyGS/claudestat34GitHubスター安定 - 11休止中
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27GitHubスター安定 - 12休止中
astragraph
エージェントPolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
github測定済み成長情報源を開く ↗
26GitHubスター安定 - 13活動中
Sentry instrumentation skill for system-behavior tracking
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24GitHubスター安定 - 14活動中
untell
その他AI-detector auditing toolkit: measures how often a detector flags genuine human writing, how stable its verdict is across seeds, and whether that verdict survives meaning-preserving edits. Detector-in-the-loop measurement harness. Claude…
github測定済み成長情報源を開く ↗
18GitHubスター安定 - 15活動中
rashomon
スキルMeasure prompt and skill improvements with blind A/B comparison.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add shinpr/rashomon18GitHubスター安定 - 16活動中
Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add prime-radiant-inc/parallel-adversarial-review17GitHubスター安定 - 17休止中
MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge17GitHubスター安定 - 18活動中
Make Claude write clearly, for everyone.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add stefanobaghino/simple-output-styles16GitHubスター安定 - 19活動中
Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace13GitHubスター安定 - 20休止中
eval-layer
その他A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
github測定済み成長情報源を開く ↗
13GitHubスター安定 - 21活動中
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-server -- npx @spanlens/mcp-server12GitHubスター安定 - 22活動中
bakeoff
スキルTurn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add CoriChui/bakeoff10GitHubスター安定 - 23活動中
galdor
その他A Go-native framework for LLM agents, with OpenTelemetry observability built in.
github測定済み成長情報源を開く ↗
10GitHubスター安定
学習・参考リソース
情報源間で正規化した測定済み成長順です。 これらは個別に参照でき、主要ランキングには含まれません。
- 1活動中情報源を開く ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubリソース測定済み成長77GitHubスター安定 - 2活動中情報源を開く ↗
Agentic_AI_Engineer
エージェントMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubリソース測定済み成長18GitHubスター安定