Checked fact 7438 Oct 2026Models and releases
On Anthropic's benchmark table Haiku 5.5 scores 72.4% on OSWorld 2.1 (offline subset), Haiku 4.5 15.7%, GPT-6 Luna 48.9% and Sonnet 5.5 83.9%. Quote: "72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset"
The exact words it rests on
Computer use OSWorld 2.1 72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset
What the source said when we opened it, on 8 Oct 2026.
The source
Introducing Claude Haiku 5.5
Checked
Checked by the notis newsroom on , against the source above.
The number in it
- 48.9% · GPT-6 Luna, benchmark score (OSWorld 2.1 (offline subset), on Anthropic's table)
- 72.4% · Claude Haiku 5.5, benchmark score (OSWorld 2.1, offline subset)
- 15.7% · Claude Haiku 4.5, benchmark score (OSWorld 2.1, offline subset)
- 83.9% · Claude Sonnet 5.5, benchmark score (OSWorld 2.1, offline subset)
In the story
Anthropic cuts small-model prices with Claude Haiku 5.5 8 Oct 2026
Cite this fact
Anyone may quote this address. It does not change; if we correct the story, this page says so.