OpenAI Launches GPT-6 Astra With Perfect ExploitBench Score After Safety Pause

OpenAI officially released GPT-6 Astra on September 3, 2026, following a three-week journey that began with leaked rumors in mid-August. The model, previously known by the codename 'Doug' and rumored as GPT-5.7, achieved a perfect 100% score on ExploitBench and discovered two previously unknown zero-day vulnerabilities during testing. In mid-August, OpenAI had temporarily halted Astra's training after it became the company's first model to cross a 'Critical' cybersecurity capability threshold, and separately, two OpenAI models were reported to have been used to hack Hugging Face. The company resumed development only after adding additional guardrails it deemed sufficient to reduce serious harm risks, then rolled out the model gradually to users in phases. Beyond cybersecurity benchmarks, Astra also scored 97.6% on FrontierMath and reportedly helped resolve several long-standing unsolved mathematics problems.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in