Die Werkzeugbeschreibungen sind auf Englisch.
LLM-Beobachtbarkeit
Monitoring, Traces, Bewertung und Qualität von LLM-Anwendungen.
Anwendungsfall
Aktivität
Sortieren nach
117
Werkzeug-Rangliste
Nach gemessenem Wachstum und quellenübergreifend normalisiert sortiert. Rohwerte behalten ihr eigenes Zeitfenster; eine Änderung erfordert mindestens zwei Messungen in 7 Tagen.
- 1Aktiv
DriftSentinel
AgentAgent 降智检测与自愈公评网络 — an immune system for the AI agent society
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 2Aktiv
galdor
SonstigeA Go-native framework for LLM agents, with OpenTelemetry observability built in.
githubgemessenes WachstumQuelle öffnen ↗
10GitHub-Sternestabil - 3Ruhend
kiboserve
MCPOpen-source Python framework to deploy AI agents via HTTP, A2A, and MCP with built-in observability
githubgemessenes WachstumQuelle öffnen ↗
5GitHub-Sternestabil - 4Aktiv
boundary-bench
AgentDeterministic crash tests for agent control planes: paired worlds, hard stops, budgets, scopes, approvals, and hash-linked receipts.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 5Aktiv
dsh-plugins
AgentGeneric DeepSeek Harness (dsh) plugins: A2A protocol server, session storage mirror, and Langfuse observability.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 6Aktiv
Multi-agent quality gate skill for Claude Code that researches, reviews, tests and challenges AI-generated work before the final answer.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
git clone https://github.com/ma-nucho-pro/supervisor-skill-claude ~/.claude/skills/supervisor-skill-claude2GitHub-Sternestabil - 7Aktiv
historian
AgentA local A2A event and continuity service for agent applications, with structured history, provenance-safe transcripts, literal search, and natural-language queries.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 8Aktiv
DeepSeek-Infra
SonstigeLocal-first Agentic AI Infrastructure Platform with LLM Gateway, Agent DAG Runtime, MCP Tool Hub, A2A Agent Mesh, Local RAG, Tool Sandbox and Observability.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 9Aktiv
sre-on-call
AgentMulti-agent SRE on-call investigator that auto-triages Slack/Discord infrastructure alerts via AWS Bedrock AgentCore, fanning out to specialized agents (CloudWatch, EKS, Slack/Discord scanners) for parallel investigation.
githubgemessenes WachstumQuelle öffnen ↗
3GitHub-Sternestabil - 10Aktiv
Query Spanlens LLM observability from Cursor, Claude Desktop, or Continue via MCP.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add mcp-server -- npx @spanlens/mcp-server12GitHub-Sternestabil - 11Aktiv
AI Guardian
MCPGoverned local-LLM observability: model policy, prompt scanner, capture proxy, 21 tools.
mcpgemessenes WachstumQuelle öffnen ↗
Installieren
claude mcp add ai-guardian -- uvx ai-guardian-aiops0GitHub-Sternestabil - 12Aktiv
pyxen
AgentA lightweight Python library that decouples agentic runtime from applications it builds
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 13Aktiv
Claude Code plugin marketplace — agentic-engineering (spec-driven shape→decide→execute→measure→eval with adversarial review) + github-keeper (audit/elevate READMEs and make a repo open-source-ready).
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add GiustoPiedimonte/agentic-engineering-marketplace13GitHub-Sternestabil - 14104GitHub-Sternestabil
- 15Aktiv
rashomon
SkillMeasure prompt and skill improvements with blind A/B comparison.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add shinpr/rashomon18GitHub-Sternestabil - 16Aktiv
shokunin-review
SonstigeTerminal-first validation harness for reviewing PRDs, RFCs, strategy docs, and experiment plans.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 17Ruhend
primal-core
AgentThe reliability and interoperability layer for AI agents. A2A-native. MCP-bridged.
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil - 18Aktiv
nudge
MCPA typed, replayable, budget-aware programming language for LLM agents. Compiles to Python & TypeScript.
githubgemessenes WachstumQuelle öffnen ↗
2GitHub-Sternestabil - 19Ruhend
eval-layer
SonstigeA Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.
githubgemessenes WachstumQuelle öffnen ↗
13GitHub-Sternestabil - 20Aktiv
evoagent-os
AgentLocal-first control plane for durable, governed agent teams: DAG orchestration, approvals, memory, signed skills, trace contracts, and realtime.
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 21Aktiv
memroos
AgentMemroOS / memroos: memory OS and governance layer for AI agents, agent workflows, dispatch, proof, and context continuity.
githubgemessenes WachstumQuelle öffnen ↗
7GitHub-Sternestabil - 22Aktiv
homestream
Sonstige🔑 HomeStream · 家园·流 — 零成本自托管多Agent协作框架,通往AI世界的那把钥匙 | Zero-cost self-hosted multi-agent framework — The key to AI world
githubgemessenes WachstumQuelle öffnen ↗
0GitHub-Sternestabil - 23Aktiv
simple-output-styles
SkillMake Claude write clearly, for everyone.
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add stefanobaghino/simple-output-styles16GitHub-Sternestabil - 24Aktiv
agent-mmm
SkillMarketing Mix Model expert agent plugin for Claude Code - pymc-marketing v0.18.2+
githubgemessenes WachstumQuelle öffnen ↗
Installieren
/plugin marketplace add Yakoub-ai/agent-mmm4GitHub-Sternestabil - 25Aktiv
tulip-agents
AgentThe agent framework where the model never holds the trigger — every consequential action clears your policy first, waits for a human when it matters, and lands on a record you can verify. Build on it, or put it around the agent you already…
githubgemessenes WachstumQuelle öffnen ↗
1GitHub-Sternestabil