Claude Haiku 5.5: Fast Coding, Benchmarks, and Pricing

XMLans Posted on 1 days ago 5 Views


Anthropic announced Claude Haiku 5.5 on October 7, aiming at fast, inexpensive work such as coding subagents, summaries, and repetitive tasks. The company describes it as its fastest model at standard speed, with stronger results than GPT-6 Luna on several reported evaluations.

Claude Haiku 5.5 launch benchmark comparison

Low prices for prompts up to 100,000 tokens

The following standard Claude API prices are in US dollars per 1 million tokens:

Token categoryPrompt up to 100,000 tokensPrompt over 100,000 tokens
Input$0.10$0.50
Output$0.50$2.50
Cache reads$0.01$0.05

The official pricing documentation applies the higher rates to a request whose total prompt exceeds 100,000 tokens. Prompt length includes cache reads and cache writes. Each request is priced separately, so long conversation histories can move a request into the more expensive tier.

Benchmark gains and the right workload

Anthropic reports higher scores than GPT-6 Luna on Terminal-Bench 4.0, knowledge-work evaluations, and its computer-use test. These are the company's benchmark results; the useful question for a developer is how reliably the model completes their own small, frequent tasks.

The launch reinforces how competitive inexpensive models have become alongside the flagship race. Read the Anthropic announcement for the evaluations, or our Claude Opus 5.5 overview for the larger model in the family.

Adapted from the original Chinese article, published on October 8, 2026 (China Standard Time).

Hi! I frequently update with various articles about technology, practical tips, and cutting-edge news. I hope it will be helpful to you!
Last updated on 2026-10-09