GPT-6 Sol, GPT-6 Astra, and Claude Opus 5.5 Tested Building Apps From Scratch

A developer tested three AI models — GPT-6 Astra, GPT-6 Sol, and Claude Opus 5.5 — by tasking each with building three distinct web apps from a minimal React, TypeScript, and Vite setup, with no prototypes or design systems provided. The apps included a delivery exception desk, a concert event guide called NIGHT LOOP, and a warehouse inspection tool for a small Android terminal. Claude Opus 5.5 consistently prioritised actionable, task-critical information upfront — such as grouping delivery cases by urgency and surfacing concert schedule conflicts — while Astra and Sol tended to lead with introductory or decorative content. On visual design, Astra stood out for the concert guide with cohesive illustrated performer cards, whereas Opus and Sol leaned on more generic gradient aesthetics. The findings suggest that, without a prototype to follow, the models differ meaningfully in how they interpret what a user needs to see first.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in