ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
123
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1782GitHubスター+1 (+0.13 %)
- 2活動中
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
github測定済み成長情報源を開く ↗
254GitHubスター-3 (-1.2 %) - 3活動中
🟪 Open-source runtime that ships any LangGraph or Google ADK agent as a production-ready FastAPI service. Bundled , AG-UI copilotkit API, chat UI, 15+ guardrails, MCP, OpenTelemetry, OIDC. One pip install. Self-hosted, no vendor lock-in.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Idun-Group/idun-agent-platform ~/.claude/skills/idun-agent-platform198GitHubスター安定 - 4活動中
Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
github測定済み成長情報源を開く ↗
131GitHubスター-11 (-7.7 %) - 5活動中
74 open-source Agent Skills for Claude Code and Codex: AI SEO, AEO and GEO, code review with an A-F ship grade, CI gates, AI evals, design systems, conversion copy, Instagram growth, iOS and Android app shipping, creator rights, and…
github測定済み成長情報源を開く ↗
129GitHubスター-10 (-7.2 %) - 6活動中
langfuse-mcp
MCPA Model Context Protocol (MCP) server for Langfuse, enabling AI agents to query Langfuse trace data for enhanced debugging and observability
github測定済み成長情報源を開く ↗
105GitHubスター安定 - 7休止中
anti-lie
スキルDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHubスター安定 - 8活動中
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add langfuse -- npx langfuse-observability-mcp-server78GitHubスター安定 - 9活動中
AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
github測定済み成長情報源を開く ↗
68GitHubスター安定 - 10活動中
Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHubスター-5 (-8.9 %) - 11活動中
arize-skills
スキルAgent skills for Arize — datasets, experiments, and traces via the ax CLI
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/Arize-ai/arize-skills ~/.claude/skills/arize-skills47GitHubスター安定 - 12活動中
Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add rennf93/opus-fable-playbook33GitHubスター安定 - 13休止中
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27GitHubスター安定 - 14休止中
astragraph
エージェントPolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
github測定済み成長情報源を開く ↗
26GitHubスター安定 - 15活動中
Sentry instrumentation skill for system-behavior tracking
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24GitHubスター安定 - 16活動中
rashomon
スキルMeasure prompt and skill improvements with blind A/B comparison.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add shinpr/rashomon18GitHubスター安定 - 17活動中
Two Claude Code skills for adversarial code review (single-model PAR and multi-model MMAR with cross-critique to catch hallucinations), plus a fixture-based eval suite.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add prime-radiant-inc/parallel-adversarial-review17GitHubスター安定 - 18休止中
MCP as a Judge: a behavioral MCP that strengthens AI coding assistants via explicit LLM evaluations
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-as-a-judge -- uvx mcp-as-a-judge17GitHubスター安定 - 19活動中
Make Claude write clearly, for everyone.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add stefanobaghino/simple-output-styles16GitHubスター安定 - 20休止中
eval-layer
その他A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
github測定済み成長情報源を開く ↗
13GitHubスター安定 - 21活動中
Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace13GitHubスター安定 - 22活動中
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add mcp-server -- npx @spanlens/mcp-server12GitHubスター安定 - 23活動中
bakeoff
スキルTurn one decision into a judged tournament of solutions, then pick the best — a Claude Code skill that generates candidates, auto-derives the rubric, judges independently, and returns a defensible winner.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add CoriChui/bakeoff10GitHubスター安定
学習・参考リソース
情報源間で正規化した測定済み成長順です。 これらは個別に参照でき、主要ランキングには含まれません。
- 1活動中情報源を開く ↗
Curated + scored map of official MCP servers and agents for DevOps, Cloud, SRE, and Platform Engineering — every entry rated on production access, approval gates, and audit evidence.
githubリソース測定済み成長77GitHubスター安定 - 2活動中情報源を開く ↗
Agentic_AI_Engineer
エージェントMy complete journey to becoming an Agentic AI Engineer through structured learning, projects, experiments, and production-ready implementations of modern AI systems.
githubリソース測定済み成長18GitHubスター安定