Pac-Bench tests AI models ability to build Pac-Man in one HTML prompt
A new benchmark called Pac-Bench evaluates how well AI language models can generate a functional Pac-Man game from a single prompt. Each model receives exactly one instruction — 'Create a Pac-Man game in a single HTML page' — with no follow-up prompts or corrections allowed. The project uses the Harness framework to run and compare model outputs consistently. It was shared on Hacker News, where it drew early community attention. The benchmark aims to measure one-shot coding capability across different AI models in a fun, visual way.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in