Who it is built for
- Researchers evaluating coding-model game-building capability
- Developers inspecting playable coding-agent outputs
- Teams comparing model progress on a shared game-development benchmark
A playable benchmark for coding models building GTA in San Francisco
MirageML Bench tracks how far coding models can progress toward building a GTA-style game set in San Francisco from clean-room repositories. It presents playable submissions, their model and build information, source repositories, deployment links, and stated provenance for direct inspection.
The decision
Start with the job, the team, and the constraints. Product fit becomes much clearer when those three line up.
Inside the product
The core product capabilities, grouped around the work they enable.
The benchmark presents accepted submissions with a direct Play link so visitors can inspect the resulting game.
Each listed submission identifies the coding model and the date on which the build was completed.
Submission details include links to the source repository and the deployed game when available.
Coding-model game builds are submitted to the Mirage project for inclusion in the benchmark. Accepted entries are shown in a playable deck with implementation, source, deployment, and lineage details. Visitors can play the submission and inspect the linked project materials.
Plans and official links
See the entry price, free access options, company details, and direct vendor destinations in one place.
Choosing for a real workflow?
We map the workflow, connect existing systems, choose what to buy, and build what is missing.