Wire flash
Anthropic Opus 5.5 beats Fable 5.1 on all benchmarks; OpenAI GPT-6 Luna scores 66.6% on DeepSWE 1.1
Editorial responsibility
- No named human review is recorded for this page.
- Source reporting is collected, normalized, translated or condensed automatically when needed.
- Automatically published source-backed update
A news intelligence analyst recaps the day's AI releases from Anthropic and OpenAI, concluding there is no clear winner. Anthropic addressed user feedback on Opus communication and introduced a subscriber reset. Opus 5.5 outperforms Fable 5.1 on every benchmark in Anthropic's comparison table, delivering Fable 5.1-level performance at roughly 40% lower cost than Opus 5. Sonnet 5.5 and Haiku 5.5 are expected in coming weeks. OpenAI's GPT-6 Luna achieved 66.6% on DeepSWE 1.1 at max effort, comparable to Fable 5 at medium effort but at 96% lower cost per task. Luna's API pricing is $0.10 per million input tokens and $0.50 per million output tokens. GPT-6 Sol approaches Astra's factual reliability on internal evaluations and outperforms low-effort Astra on AutomationBench at xhigh effort. The analyst also notes unverified rumors about an OpenAI counterpart to Grok Bot.
Source report
Both OpenAI and Anthropic delivered impressive updates today, each with distinct strengths. Here's a recap of the key announcements.
Anthropic's Updates
Anthropic addressed user feedback on how Opus communicates and introduced a new "reset" feature that subscribers can save for later — a welcome sign that the company is listening.
Benchmark Surprises
The benchmark results were a standout. Opus 5.5 outperforms Fable 5.1 on every benchmark in Anthropic's headline comparison table. According to Anthropic, Opus 5.5 delivers Fable 5.1-level performance on most tasks while costing approximately 40% less than Opus 5 on typical workloads at default settings.
"A huge statement. I really did not see that coming. Kudos, Anthropic."
Upcoming Models
- Sonnet 5.5 and Haiku 5.5 are expected in the coming weeks.
- Excitement is building around how far efficiency gains will carry over, and what a future flagship Fable 5.5 could deliver.
OpenAI's Announcements
OpenAI also impressed, with efficiency taking center stage.
GPT-6 Luna Performance
On DeepSWE 1.1, GPT-6 Luna at max effort scored 66.6% — comparable to Fable 5 at medium effort — at 96% lower cost per task in OpenAI's comparison.
- API pricing: $0.10 per million input tokens, $0.50 per million output tokens.
"Intelligence too cheap to meter? We're certainly getting a striking demonstration of how much cheaper capable models can become."
GPT-6 Sol Capabilities
GPT-6 Sol approaches Astra's factual reliability on OpenAI's internal evaluation and even outperforms low-effort Astra on AutomationBench at xhigh effort. While this doesn't establish Astra-level performance across the board, it remains an impressive result.
"Kudos, OpenAI. All of this makes me excited for the models still to come."
Other Developments
Beyond today's releases, rumors are circulating about an OpenAI counterpart to Grok Bot, though no launch announcement has been verified.
Final Verdict
There is no clear winner today. Both OpenAI and Anthropic delivered outstanding releases, each with its own unique focus. This is what makes competition exciting.
Source
kimmonismusNeutral / independent
Part of this Story
Anthropic launches Claude Opus 5.5 with Fable 5.1-level performance at 40% lower cost