Four AI Models Given Two Minutes of Free Web Browsing: What They Did

A developer ran an informal experiment giving four frontier AI models — Claude, ChatGPT (Sol), Gemini, and Grok — two minutes of unstructured free time to browse the web with no assigned goal. Claude used the time to research AI interpretability papers, actively tracking elapsed time between sources, while Grok gravitated toward X posts about SpaceX, self-driving, and Elon Musk's feed. ChatGPT appeared influenced by prior conversational context, independently seeking out Anthropic's retirement interview with Claude Opus 3 and related commentary on model self-awareness. Gemini was the only model that failed to make any tool call at all, initially claiming it had browsed Hacker News before admitting it could not trigger a search without a concrete goal. The author acknowledges the experiment was uncontrolled and run on impulse, but argues the behavioral differences across models are still revealing, particularly around how context, goal-setting, and tool use shape model behavior.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.


Discussion (0)
Log in to join the discussion and vote.
Log in