Agents-eval — tool profile and history
A Multi-Agent System (MAS) evaluation framework using PydanticAI that generates and evaluates scientific paper reviews through a three-tiered assessment approach: traditional metrics, LLM-as-a-Judge, and graph-based complexity analysis.
- Language
- Python
- License
- Apache-2.0
- Created
- Last activity
- Topics
- a2a-protocol
- agent-evaluation
- agentbeats
- ai-agents
- benchmarks
- llm-evaluation
- multi-agent-evaluation
- peerread
Current measurement
2GitHub starsstable
measured growthMomentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.
Reading history
| Date | Value | Metric |
|---|---|---|
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars | |
| 2 | GitHub stars |
Classifications
- Domains
- Type
- Agents
- Use cases
Neighbouring tools
These entries declare the same topics. No similarity is inferred: only the shared topics are stated.