Le descrizioni degli strumenti sono in inglese.

coder_eval — scheda e storico dello strumento

Test that your Claude Code skills, MCP servers, and CLIs actually work when an agent uses them — sandboxed YAML suites, activation checks, A/B experiments, CI gates.

Attivo

Fonte canonica ↗

Linguaggio
Python
Licenza
Apache-2.0
Creazione
Ultima attività
Argomenti
  • agent-evaluation
  • agent-skills
  • agent-testing
  • anthropic
  • claude
  • claude-agent-sdk
  • claude-code
  • claude-code-plugins-marketplace
  • claude-code-skills
  • claude-skills
  • codex
  • coding-agents
  • evaluation-framework
  • gemini
  • github-actions
  • long-horizon-agents
  • regression-testing
  • skill-testing
  • swe-bench
  • terminal-bench

Misurazione attuale

119stelle GitHub+3 (+2.6 %)
crescita misurata

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

Storico delle rilevazioni

90 giorni · 90 rilevazioni massime
DataValoreMetrica
119stelle GitHub
119stelle GitHub
118stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
116stelle GitHub
115stelle GitHub
115stelle GitHub
113stelle GitHub
113stelle GitHub
113stelle GitHub
111stelle GitHub
110stelle GitHub
109stelle GitHub
109stelle GitHub

Classificazioni

Ambiti
Tipo
Server MCP
Utilizzi

Queste voci dichiarano gli stessi argomenti. Nessuna somiglianza viene dedotta: sono indicati solo gli argomenti in comune.