ツールの説明は英語です。

LLM可観測性

LLMアプリの監視、トレース、評価、品質管理。

用途

アクティビティ

並び順

122

ツールランキング

情報源間で正規化した測定済み成長順です。 各値は情報源固有の期間を使用し、変化の算出には7日間で2回以上の測定が必要です。

  1. 1
    活動中

    Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.

    mcp測定済み成長情報源を開く ↗

    インストール claude mcp add mcp-server -- npx @spanlens/mcp-server

    12GitHubスター安定
  2. 2

    その他
    活動中

    Test your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…

    github測定済み成長情報源を開く ↗

    24 768GitHubスター+142 (+0.58 %)
  3. 3

    その他
    活動中

    AI SRE AgenticOps for Kubernetes and cloud infrastructure.

    github測定済み成長情報源を開く ↗

    782GitHubスター+1 (+0.13 %)
  4. 4

    エージェント
    活動中

    The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…

    github測定済み成長情報源を開く ↗

    27 783GitHubスター+82 (+0.30 %)
  5. 5

    ライブラリ
    活動中

    Prefect is a workflow orchestration framework for building resilient data pipelines in Python.

    github測定済み成長情報源を開く ↗

    23 766GitHubスター+66 (+0.28 %)
  6. 6

    その他
    活動中

    A high-performance observability data pipeline.

    github測定済み成長情報源を開く ↗

    22 509GitHubスター+40 (+0.18 %)
  7. 7

    その他
    活動中

    Open source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…

    github測定済み成長情報源を開く ↗

    21 615GitHubスター+96 (+0.45 %)
  8. 8

    その他
    活動中

    SigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…

    github測定済み成長情報源を開く ↗

    32 003GitHubスター+50 (+0.16 %)
  9. 9

    その他
    活動中

    Mastra is the modern TypeScript framework for AI-powered applications and agents.

    github測定済み成長情報源を開く ↗

    27 658GitHubスター+128 (+0.46 %)
  10. 10

    その他
    活動中

    The fastest path to AI-powered full stack observability, even for lean teams.

    github測定済み成長情報源を開く ↗

    80 412GitHubスター+85 (+0.11 %)
  11. 11
    活動中

    🫖 Status page with uptime monitoring & API monitoring as code 🫖

    github測定済み成長情報源を開く ↗

    9 056GitHubスター+31 (+0.34 %)
  12. 12

    その他
    活動中

    Self-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.

    github測定済み成長情報源を開く ↗

    131GitHubスター-4 (-3.0 %)
  13. 13

    その他
    活動中

    eBPF-based Networking, Security, and Observability

    github測定済み成長情報源を開く ↗

    25 047GitHubスター+29 (+0.12 %)
  14. 14

    エージェント
    活動中

    eBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.

    github測定済み成長情報源を開く ↗

    12 068GitHubスター+8 (+0.07 %)
  15. 15

    エージェント
    活動中

    DataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.

    github測定済み成長情報源を開く ↗

    642GitHubスター+32 (+5.2 %)
  16. 16

    その他
    活動中

    Evidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.

    github測定済み成長情報源を開く ↗

    140GitHubスター+24 (+20.7 %)
  17. 17

    スキル
    活動中

    The Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.

    github測定済み成長情報源を開く ↗

    インストール git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method

    2 272GitHubスター+16 (+0.71 %)
  18. 18

    エージェント
    活動中

    The Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…

    github測定済み成長情報源を開く ↗

    156GitHubスター+19 (+13.9 %)
  19. 19
    活動中

    Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…

    github測定済み成長情報源を開く ↗

    3GitHubスター安定
  20. 20

    スキル
    休止中

    An assembly line for AI software development. 35 skills, 11 agent personas, 29 commands. From raw idea to shipped code.

    github測定済み成長情報源を開く ↗

    インストール /plugin marketplace add aneja5/forge-skills

    3GitHubスター安定
  21. 21
    活動中

    Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.

    github測定済み成長情報源を開く ↗

    0GitHubスター安定
  22. 22

    エージェント
    活動中

    Multi-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.

    github測定済み成長情報源を開く ↗

    3GitHubスター安定
  23. 23

    MCP
    活動中

    AURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.

    github測定済み成長情報源を開く ↗

    254GitHubスター-3 (-1.2 %)
  24. 24

    エージェント
    活動中

    Generic DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.

    github測定済み成長情報源を開く ↗

    1GitHubスター安定
  25. 25

    MCP
    活動中

    A typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.

    github測定済み成長情報源を開く ↗

    2GitHubスター安定