工具说明为英文。

eval-layer — 工具资料与历史

A Claude Code skill that adds a rubric-based eval layer to any agent project. Framework-agnostic — generates rubric, test cases, judge prompt, and harness. Returns a weighted score plus a judge-leniency signal.

休眠

官方来源 ↗

语言
HTML
创建时间
最近活动
主题
  • agent-evaluation
  • anthropic
  • claude-code
  • claude-code-skill
  • evals
  • llm-agents
  • llm-as-a-judge
  • rubric

当前测量

13GitHub 星标+1 (+8.3 %)
实测增长

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

测量历史

90 天 · 90 次测量上限
日期数值指标
13GitHub 星标
13GitHub 星标
13GitHub 星标
13GitHub 星标
13GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标
12GitHub 星标

分类

领域
类型
其他
用途

这些条目声明了相同的主题。此处不推断相似性,仅列出共同的主题。