ainotis Join
My notis

Checked fact 7438 Oct 2026Models and releases

On Anthropic's benchmark table Haiku 5.5 scores 72.4% on OSWorld 2.1 (offline subset), Haiku 4.5 15.7%, GPT-6 Luna 48.9% and Sonnet 5.5 83.9%. Quote: "72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset"

The exact words it rests on

Computer use OSWorld 2.1 72.4% Offline subset 15.7% Offline subset 48.9% Offline subset 83.9% Offline subset

What the source said when we opened it, on 8 Oct 2026.

The source

Introducing Claude Haiku 5.5
Anthropic · 2026-10-07

Checked

Checked by the notis newsroom on , against the source above.

The number in it

  • 48.9% · GPT-6 Luna, benchmark score (OSWorld 2.1 (offline subset), on Anthropic's table)
  • 72.4% · Claude Haiku 5.5, benchmark score (OSWorld 2.1, offline subset)
  • 15.7% · Claude Haiku 4.5, benchmark score (OSWorld 2.1, offline subset)
  • 83.9% · Claude Sonnet 5.5, benchmark score (OSWorld 2.1, offline subset)

In the story

Anthropic cuts small-model prices with Claude Haiku 5.5 8 Oct 2026

Cite this fact

Anyone may quote this address. It does not change; if we correct the story, this page says so.