Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
117
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add mcp-server -- npx @spanlens/mcp-server12GitHub-Sternestabil - 2Aktiv
promptfoo
SonstigeTest your prompts, agents, and RAGs. Red teaming/pentesting/vulnerability scanning for AI. Compare performance of GPT, Claude, Gemini, DeepSeek, and more. Simple declarative configs with command line and CI/CD integration. Used by OpenAI…
githubgemessenes WachstumQuelle öffnen ↗
24 768GitHub-Sterne+142 (+0.58 %) - 3Aktiv
Flawless
SonstigeAI SRE AgenticOps for Kubernetes and cloud infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
782GitHub-Sterne+1 (+0.13 %) - 4Aktiv
mlflow
AgentThe open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlling costs and managing access to models…
githubgemessenes WachstumQuelle öffnen ↗
27 783GitHub-Sterne+82 (+0.30 %) - 5Aktiv
prefect
BibliothekPrefect is a workflow orchestration framework for building resilient data pipelines in Python.
githubgemessenes WachstumQuelle öffnen ↗
23 766GitHub-Sterne+66 (+0.28 %) - 6Aktiv
vector
SonstigeA high-performance observability data pipeline.
githubgemessenes WachstumQuelle öffnen ↗
22 509GitHub-Sterne+40 (+0.18 %) - 7Aktiv
openobserve
SonstigeOpen source observability platform for logs, metrics, traces, frontend monitoring, pipelines and LLM observability. A sophisticated, simple and highly performant alternative to Datadog, Splunk, and Elasticsearch with 140x lower storage…
githubgemessenes WachstumQuelle öffnen ↗
21 615GitHub-Sterne+96 (+0.45 %) - 8Aktiv
signoz
SonstigeSigNoz is an open-source, OpenTelemetry-native observability platform for your team and their AI agents. Get logs, metrics, and traces in one tool with features like APM, distributed tracing, log management, infra monitoring, etc. Combined…
githubgemessenes WachstumQuelle öffnen ↗
32 003GitHub-Sterne+50 (+0.16 %) - 9Aktiv
mastra
SonstigeMastra is the modern TypeScript framework for AI-powered applications and agents.
githubgemessenes WachstumQuelle öffnen ↗
27 658GitHub-Sterne+128 (+0.46 %) - 10Aktiv
netdata
SonstigeThe fastest path to AI-powered full stack observability, even for lean teams.
githubgemessenes WachstumQuelle öffnen ↗
80 412GitHub-Sterne+85 (+0.11 %) - 11Aktiv
openstatus
MCP🫖 Status page with uptime monitoring & API monitoring as code 🫖
githubgemessenes WachstumQuelle öffnen ↗
9 056GitHub-Sterne+31 (+0.34 %) - 12Aktiv
rocketplaneIO
SonstigeSelf-hosted AI SRE for Kubernetes — zero-instrumentation eBPF observability plus a copilot that fixes issues through guardrailed, self-verifying actions. BYO-LLM, air-gapped capable.
githubgemessenes WachstumQuelle öffnen ↗
131GitHub-Sterne-4 (-3.0 %) - 13Aktiv
cilium
SonstigeeBPF-based Networking, Security, and Observability
githubgemessenes WachstumQuelle öffnen ↗
25 047GitHub-Sterne+29 (+0.12 %) - 14Aktiv
kubeshark
AgenteBPF-powered network observability for Kubernetes. Indexes L4/L7 traffic with full K8s context, decrypts TLS without keys. Queryable by AI agents via MCP and humans via dashboard.
githubgemessenes WachstumQuelle öffnen ↗
12 068GitHub-Sterne+8 (+0.07 %) - 15Aktiv
databuff
AgentDataBuff is an AI-native APM built on Opentelemetry,with multi-agent troubleshooting out of the box.
githubgemessenes WachstumQuelle öffnen ↗
642GitHub-Sterne+32 (+5.2 %) - 16Aktiv
VeriRun
SonstigeEvidence-first infrastructure for reproducible, isolated executable evaluation and online rewards.
githubgemessenes WachstumQuelle öffnen ↗
140GitHub-Sterne+24 (+20.7 %) - 17Aktiv
fable-method
SkillThe Fable Workflow: how Claude Fable 5 worked, distilled into skills any model can run, with the eval that keeps it honest. Think / act / prove.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/Sahir619/fable-method ~/.claude/skills/fable-method2 272GitHub-Sterne+16 (+0.71 %) - 18Aktiv
agent-kernel
AgentThe Operating System for Scalable Enterprise AI Agents - Run, orchestrate, and deploy Compliant Enterprise AI Agents at scale across frameworks, without lock-in, rewrites or fragile glue code. Native support for MCP, A2A. Interface with all…
githubgemessenes WachstumQuelle öffnen ↗
156GitHub-Sterne+19 (+13.9 %) - 19Aktiv
Production reliability, benchmarking, and evaluation layer for MCP agent runtimes — protocol negotiation, A2A interop, circuit breakers, connection pooling, adaptive streaming, and autoscaling, with measured throughput/latency and a…
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil - 20Ruhend
forge-skills
SkillAn assembly line for AI software development. 35 skills, 11 agent personas, 29 commands. From raw idea to shipped code.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add aneja5/forge-skills3GitHub-Sternestabil - 21Aktiv
Architecture-first Python scaffold for an auditable multi-agent corporate credit desk using A2A, MCP, deterministic credit policies, model routing, and OpenTelemetry.
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 22Aktiv
sre-on-call
AgentMulti-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil - 23Aktiv
aura
MCPAURA is a production-tested SRE agent platform you can deploy in minutes. AURA handles the guardrails, APIs, state management, streaming, and failure handling required to put AI to work safely on production infrastructure.
githubgemessenes WachstumQuelle öffnen ↗
254GitHub-Sterne-3 (-1.2 %) - 24Aktiv
dsh-plugins
AgentGeneric DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 25Aktiv
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubgemessenes WachstumQuelle öffnen ↗
2GitHub-Sternestabil