Claude Opus 5 Launches With Strong Benchmarks but Shifts Safety Responsibility to Developers
Anthropic has released Claude Opus 5, its most capable and aligned model to date, with notable improvements in coding, computer use, and resistance to prompt injection attacks. The model has been rated ASL-3 under Anthropic's Responsible Scaling Policy, meaning it poses limited but non-zero risk related to biological and chemical weapons assistance, without crossing new catastrophic thresholds. On the cybersecurity front, Opus 5 nearly doubled its predecessor's bug-detection performance on the OSS-Fuzz benchmark, though it still falls short of generating fully working exploits. A key concern for developers is the measurable safety gap between the raw API version and the consumer-facing claude.ai, with child safety compliance in multi-turn conversations jumping from 86% via API to 99% on the platform. Anthropic has made clear that developers building on the API must implement their own safety layers, including input sanitization, tool-access constraints, and rate limiting.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in