Agents-eval — tool profile and history

A Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.

Dormant

Canonical source ↗

Language
Python
License
Apache-2.0
Created
Last activity
Topics
  • a2a-protocol
  • agent-evaluation
  • agentbeats
  • ai-agents
  • benchmarks
  • llm-evaluation
  • multi-agent-evaluation
  • peerread

Current measurement

2GitHub starsstable
measured growth

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

Reading history

90 days · 90 maximum readings
DateValueMetric
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars
2GitHub stars

Classifications

Domains
Type
Agents
Use cases

These entries declare the same topics. No similarity is inferred: only the shared topics are stated.