Anthropic Releases Claude Haiku 5.5 With Lower Prices And Major Performance Jumps
AI

Anthropic Releases Claude Haiku 5.5 With Lower Prices And Major Performance Jumps

TechNews Editorial
TechNews EditorialOct 7, 2026 · 2 min read
Share

Why it matters

The release matters because significant price cuts and performance gains in small AI models directly affect the economics of high-volume and agentic tasks.

The facts

  • Anthropic released Claude Haiku 5.5 with prices dropping by up to 90 percent for prompts under 100,000 tokens according to the company.
  • The new small model scores higher than its predecessor across multiple benchmarks and features adjustable reasoning levels.
  • Anthropic also cut cache read costs for Sonnet 5.5 by 50 percent and rolled out monthly API credits for subscribers.

Anthropic has released Claude Haiku 5.5. This release delivers the company's fastest and most affordable small model to date. Benchmark results show a major performance jump over its predecessor. Anthropic is also cutting prices for Sonnet 5.5 alongside this launch.

Haiku 5.5 is designed for high-volume, cost-sensitive tasks. These tasks include summarization, database queries, classification, and live customer support. On average, the model costs about 75 percent less than Haiku 4.5. For requests with prompts up to 100,000 tokens, prices drop by up to 90 percent. Anthropic notes that these shorter prompts account for roughly 90 percent of all previous Haiku requests. Prompts longer than 100,000 tokens cost five times as much.

Haiku 5.5 uses an updated tokenizer that consumes slightly more tokens per task than its predecessor. A similar tokenizer change on Opus 4.x models caused token usage to jump about 30 percent. Real-world savings are likely smaller than the per-token prices suggest. Anthropic points out this technical detail regarding the tokenizer updates.

Read nextGoogle Releases Nano Banana 2.1 Image Model With Lower Prices

Haiku 5.5 shows large benchmark improvements

Haiku 5.5 scores 1,620 on the knowledge benchmark GDPval-AA v2.1, which is more than double the 735 scored by its predecessor. On Humanity's Last Exam, it hits 45.9 percent without tools and 57.4 percent with tools. The previous version managed 10.2 percent and 18.7 percent on those tests.

The biggest performance jump appears in computer use. Haiku 5.5 scores 72.4 percent on OSWorld-2.1, up from 15.7 percent. On the agentic coding benchmark Terminal-Bench 4.0, it reaches 39.2 percent while Haiku 4.5 scored zero. Anthropic compares the model against OpenAI's budget model GPT-6 Luna, stating that Haiku 5.5 leads across every tested category. Sonnet 5.5 reference scores show that Haiku 5.5 still falls well behind Anthropic's larger model.

Haiku 5.5 is the first Haiku-class model with adjustable reasoning levels. This feature lets users balance cost against quality. Anthropic states the model works best for narrowly scoped tasks like compaction, summarization, or sub-agent work. Sonnet 5.5 and Opus 5.5 remain the better picks for complex agentic coding.

Cybersecurity safeguards are tighter than on the predecessor. They allow a broader range of defensive tasks than Sonnet 5.5 partly because the model is less capable overall. Penetration testing stays blocked. Organizations with broader needs can apply for verification programs for life sciences and cybersecurity.

Haiku 5.5 is available now across all platforms. These platforms include Amazon Web Services, Google Cloud, and Microsoft Azure.

A small automated security agent inspects a server’s network connections and identifies a vulnerable configuration without probing the system.
Illustration: AI & Tech News

Anthropic cuts Sonnet costs and adds API credits

Anthropic is cutting cache read costs for Sonnet 5.5 by 50 percent, dropping from $0.20 to $0.10 per million tokens. The company states this should reduce costs for most agentic tasks by about 20 percent. The change follows the introduction of OpenAI's new GPT-6.1 series.

Anthropic is also rolling out monthly API credits. Max-5x subscribers get $100, Max-20x subscribers get $200, and Team subscribers receive up to $500 per month. Users can spend these credits to experiment with tools, apps, and agents through the API. The company is updating its Python and TypeScript SDKs to add beta support for computer use and browser use.

Newsletter

Get the best AI & tech news daily

A concise daily digest. Unsubscribe anytime.

We use your email only to send this newsletter.

Keep reading