aaabench — tool profile and history

A long-horizon benchmark harness: give a coding agent a real game engine, professional conditions and time, and ask it to build an open-world game. Harness only, no results.

Active

Canonical source ↗

Language
Shell
License
MIT
Created
Last activity
Topics
  • ai-agents
  • benchmark
  • game-development
  • llm-evaluation
  • mcp
  • unreal-engine

Current measurement

297GitHub stars
estimated momentum

Momentum uses available readings from the last 7 days: measured with at least two comparable readings, estimated otherwise.

Reading history

90 days · 90 maximum readings
DateValueMetric
297GitHub stars

Classifications

Domains
Type
Agents
Use cases

These entries declare the same topics. No similarity is inferred: only the shared topics are stated.