A long-horizon benchmark harness: give a coding agent a real game engine, professional conditions and time, and ask it to build an open-world game. Harness only, no results.
Test a coding agent's ability to build an open-world game using a real game engine and professional conditions.
Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.
Get an email alert on its next release or when it starts trending — never miss the moment.
Free · no card · unsubscribe anytimeA long-horizon benchmark harness: give a coding agent a real game engine, professional conditions and time, and ask it to build an open-world game. Harness only, no results.
aaabench has 377 stars on GitHub. It has been forked 70 times. aaabench is written mainly in Shell. It has been in active development since 2026. aaabench is available under the MIT license. Its main topics are ai-agents, benchmark, game-development, llm-evaluation.
A long-horizon benchmark harness: give a coding agent a real game engine, professional conditions and time, and ask it to build an open-world game. Harness only, no results.
aaabench is an open-source project. It is released under the MIT license.
Yes. aaabench is free and open source — you can use, modify and self-host it.
aaabench is available under the MIT license.
aaabench is written mainly in Shell.
Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.
[](https://olud.ai/project/ukanwat-aaabench.html)
Measured from GitHub topics shared by both projects, weighted by how rare each topic is.