ツールの説明は英語です。
LLM可観測性
LLMアプリの監視、トレース、評価、品質管理。
用途
アクティビティ
並び順
117
ツールランキング
情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。
- 1活動中
evoagent-os
エージェントLocal-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 2105GitHubスター+1 (+0.96 %)
- 3活動中
Vendor-neutral OpenTelemetry skills for AI coding agents, grounded in upstream sources
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/ollygarden/opentelemetry-agent-skills ~/.claude/skills/opentelemetry-agent-skills97GitHubスター+4 (+4.3 %) - 4休止中
anti-lie
スキルDon't make LLMs honest. Make every factual claim auditable. — An LLM Claim Auditing Layer with T1-T7 truth gradients. 98.1% business effectiveness on LiarBench v0.2.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/lc198707/anti-lie ~/.claude/skills/anti-lie89GitHubスター安定 - 5活動中
green-agent
エージェントA2A green-agent orchestrator for evaluating agents on the AppWorld benchmark, built on the AgentBeats SDK
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 6活動中
MCP server for Langfuse LLM observability — trace and observation analysis.
mcp測定済み成長情報源を開く ↗
インストール
claude mcp add langfuse -- npx langfuse-observability-mcp-server78GitHubスター安定 - 7活動中
Build tested agent skills and govern their lifecycle through a user-defined marketplace: evidence, discovery, updates, rollback, quarantine, and 17-platform distribution.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/FrancyJGLisboa/agent-skill-creator ~/.claude/skills/agent-skill-creator2 367GitHubスター+28 (+1.2 %) - 8活動中
skill-kit
スキルlocal-first analytics for AI agent skills
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/crafter-station/skill-kit ~/.claude/skills/skill-kit77GitHubスター+1 (+1.3 %) - 9活動中
homestream
その他🔑 HomeStream · 家园·流 — 零成本自托管多Agent协作框架,通往AI世界的那把钥匙 | Zero-cost self-hosted multi-agent framework — The key to AI world
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 10活動中
Agent Skill evaluation harness for paired variants, trace artifacts, and runner adapters
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/adewale/skill-eval-harness ~/.claude/skills/skill-eval-harness73GitHubスター+4 (+5.8 %) - 11活動中
AgentX python SDK. Build multi-agent AI workforce. Run evaluation. Trace your agent. Full Observability.
github測定済み成長情報源を開く ↗
68GitHubスター安定 - 12活動中
Open benchmark for Claude Code SEO skills — real headless execution against fixture sites with planted-defect answer keys. Deterministic scoring, pre-registered rubric.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/aleclindz/seo-skill-bench ~/.claude/skills/seo-skill-bench51GitHubスター-5 (-8.9 %) - 13休止中
A Python proof-of-concept for tracing multi-turn Agent-to-Agent (A2A) conversations as a single unified MLflow trace for LLM observability and evaluation.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 14活動中
nora
MCPOpen-source, self-hosted control plane for OpenClaw and Hermes AI-agent fleets on Docker/Kubernetes — REST, CLI, and MCP.
github測定済み成長情報源を開く ↗
50GitHubスター+2 (+4.2 %) - 15活動中
A Go framework for building production agent systems with graph workflows, tools, memory, A2A, AG-UI, MCP, evaluation, and observability.
github測定済み成長情報源を開く ↗
1 759GitHubスター+18 (+1.0 %) - 16活動中
claudestat
スキルReal-time execution trace and cost intelligence for Claude Code
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add DeibyGS/claudestat34GitHubスター+1 (+3.0 %) - 17活動中
Make Claude Opus 4.8 behave like Claude Fable 5 — doctrine output style, drift-catching hooks, and an eval loop against golden Fable transcripts. Claude Code plugin.
github測定済み成長情報源を開く ↗
インストール
/plugin marketplace add rennf93/opus-fable-playbook33GitHubスター安定 - 18活動中
agentgateway
MCPAgentGateway — independent third-party profile of a public API surface, by API Evangelist. AgentGateway is an open-source, AI-native proxy and gateway for routing, observing, and governing traffic to and from AI agents, LLM providers, and…
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 19休止中
🔍 AI observability skill for Claude Code. Debug LangChain/LangGraph agents by fetching execution traces from LangSmith Studio directly in your terminal.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/OthmanAdi/langsmith-fetch-skill ~/.claude/skills/langsmith-fetch-skill27GitHubスター安定 - 20休止中
astragraph
エージェントPolicy-enforced observability and fail-closed guardrails for MCP/A2A multi-agent systems.
github測定済み成長情報源を開く ↗
26GitHubスター安定 - 21活動中
SkillForge
スキルA skill creator that proves its skills work. Evidence-driven skill creation for Claude Code and Codex: baseline-tested generation, per-skill regression evals, ecosystem doctor, cross-runtime compile, and an opt-in proactive advisor.
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tripleyak/SkillForge ~/.claude/skills/SkillForge884GitHubスター+1 (+0.11 %) - 22休止中
otel-agent-provenance
エージェントOpenTelemetry semantic conventions and instrumentation for agent provenance, derivation lineage, and acceptance criteria evaluation. Fills the Microsoft AI stack observability gap.
github測定済み成長情報源を開く ↗
0GitHubスター安定 - 23活動中
Sentry instrumentation skill for system-behavior tracking
github測定済み成長情報源を開く ↗
インストール
git clone https://github.com/tortastudios/sentry-instrumentation ~/.claude/skills/sentry-instrumentation24GitHubスター安定 - 24活動中
Deterministic, local-only audit reports for Claude Code AI agent sessions. Rust + SQLite. Zero network calls. MCP server for agents.
github測定済み成長情報源を開く ↗
22GitHubスター+2 (+10.0 %) - 25活動中
Log Claude Code sessions to Opik, the open-source LLM observability and evaluation platform, built by Comet. Tracing, evaluation, and skills for observable AI applications.
github測定済み成長情報源を開く ↗
22GitHubスター+1 (+4.8 %)