Claude Opus 5.5 cuts costs 40%, delivers faster output with no API changes
Anthropic released Claude Opus 5.5 this week with a 40% price reduction and 30% faster output generation compared to its predecessor, while matching Fable 5.1 benchmark performance. The most significant saving is in cache reads, which drop from $0.50 to $0.20 per million tokens — a 60% reduction that benefits agentic and retrieval-heavy workflows. No API or prompt changes are required; developers can migrate simply by updating the model string or staying on the latest alias. Separately, Vercel's AI Gateway expanded support to include four new models — GLM-5.3 Flash, DeepSeek V4.1 Flash, Qwen 3.8 Flash, and Grok 4.7 — consolidating provider access under a single integration layer. Google's Gemini 3.5 Transcribe also joined the Gateway, offering WebSocket-based live transcription across 85+ languages without requiring a separate Google Speech-to-Text credential setup.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.

Discussion (0)
Log in to join the discussion and vote.
Log in