Claude Haiku 5.5 With 90% Discount is Here
Although its high token usage makes the performance lead considerably less impressive in terms of efficiency, anthropic's new Claude Haiku 5.5 has taken the top spot among small AI models in Artificial Analysis' Intelligence Index.
Although its high token usage makes the performance lead considerably less impressive in terms of efficiency, anthropic's new Claude Haiku 5.5 has taken the top spot among small AI models in Artificial Analysis' Intelligence Index.
Article outline
- What happened
- The key numbers
- What comes next
- Why it matters
- The details
- The bottom line
Key points
- Haiku 5.5 scores 43 at its highest reasoning setting, ahead of GLM-5.3 Flash at 42, Gemini 3.8 Flash at 41, and GPT-6 Luna at 38.
- Cost per 1M Tokens Haiku 5.5 up to 100K Haiku 5.5 over 100K Haiku 4.5.
- Anthropic has additionally climbed the model's context window significantly, from 200, 000 tokens on Haiku 4.5 to one million tokens on Haiku 5.5.
- The model scored 1, 620 on GDPval-AA v2.1, compared with 735 for Haiku 4.5.
- While output costs $0.50 per million, for prompts of up to 100, 000 tokens, input costs just $0.10 per million tokens.
Haiku 5.5 scores 43 at its highest reasoning setting, ahead of GLM-5.3 Flash at 42, Gemini 3.8 Flash at 41, and GPT-6 Luna at 38. It still trails the larger Claude Sonnet 5.5 by 13 points. Solid Performance Comes With Heavy Token Usage.
Artificial Analysis discovered that Haiku 5.5 consumes around 162, 000 output tokens per task at its maximum reasoning setting.
GPT-6 Luna uses roughly 50, 000 tokens per task, meaning Haiku can consume more than three times as plenty of tokens.
Notably, the difference remains even when performance is matched. At its "high" reasoning setting, Haiku 5.5 scores 38 while using around 55, 000 tokens, compared with roughly 50, 000 tokens for GPT-6 Luna at the same score.
Artificial Analysis therefore gives GPT-6 Luna an advantage in Pareto efficiency. It measures how much performance a model delivers relative to resources and cost.
Moving Haiku 5.5 from its "xhigh" to "max" setting adds only around two Intelligence Index points while increasing token consumption by roughly 1.8 times.
While GPT-6 Luna reaches broadly comparable performance for roughly one-third of that amount, at maximum effort, Haiku 5.5 reportedly costs around $0.21 per Intelligence Index task.
Claude Sonnet 5.5 is Faster, Smarter and Up to 30% Cheaper. Haiku Hallucinates Less but Knows Less. Haiku 5.5 performed better in Artificial Analysis' hallucination testing.
Meanwhile, the model hallucinated in around 40% of tested cases, compared with 77% for GPT-6 Luna.
Nevertheless, GPT-6 Luna demonstrated stronger factual knowledge on AA-Omniscience, scoring 44% accuracy compared with Haiku 5.5's 36%.
Haiku partly compensates by being more willing to admit when it does not know an answer rather than generating an incorrect response.
Anthropic has additionally climbed the model's context window significantly, from 200, 000 tokens on Haiku 4.5 to one million tokens on Haiku 5.5. Massive Upgrade Over Haiku 4.5.
Anthropic positions Haiku 5.5 as its fastest and cheapest small model for high-volume workloads such as summarization, classification, database queries and customer backing.
Meanwhile, the model scored 1, 620 on GDPval-AA v2.1, compared with 735 for Haiku 4.5.
On Humanity's Last Exam, Haiku 5.5 reached 45.9% without tools and 57.4% with tools, compared with just 10.2% and 18.7% respectively for its predecessor.
Notably, the largest improvement came in computer employ. Haiku 5.5 scored 72.4% on the offline subset of OSWorld 2.1, up from 15.7% for Haiku 4.5 and ahead of GPT-6 Luna's 48.9%.
Benchmark Haiku 5.5 Haiku 4.5 GPT-6 Luna Sonnet 5.5. GDPval-AA v2.1 1, 620 735 1, 437 1, 840. AA-Briefcase v1.1 1, 578 614 1, 336 1, 824. OSWorld 2.1 72.4% 15.7% 48.9% 83.9%. Terminal-Bench 4.0 39.2% 0% 16.4% 70.6%. FrontierCode 1.1 46.4% – 42.4% 52.1%. Chartography 46.4% 6.4% 29.1% 61.6%. Anthropic Slashes Haiku Pricing.
Haiku 5.5 is additionally substantially cheaper per token than Haiku 4.5.
Cost per 1M Tokens Haiku 5.5 up to 100K Haiku 5.5 over 100K Haiku 4.5. Cache reads $0.01 $0.05 $0.10. Cache writes $0.125 $0.625 $1.25. Input $0.10 $0.50 $1.00. Output $0.50 $2.50 $5.00. AI Giants Need $6 Trillion Revenue to Justify Data Centers.
Anthropic notes this makes Haiku 5.5 around 75% cheaper on average than Haiku 4.5, with savings of up to 90% for shorter prompts.
Nevertheless, the model uses a new tokenizer and can generate significantly more tokens, meaning real-world savings may be smaller than the headline per-token rate reductions suggest. Adjustable Reasoning Comes to Haiku.
Haiku 5.5 is the first model in the Haiku family to offer adjustable reasoning levels, allowing developers to trade higher performance for greater cost and token consumption.
Anthropic recommends it mainly for tightly defined tasks, including summarization, compaction, and sub-agent workloads.
For more demanding agentic coding tasks, the firm still recommends Claude Sonnet 5.5 or Opus 5.5.
Haiku 5.5 is available throughout Anthropic's platforms as well as Amazon Web Services, Google Cloud and Microsoft Azure.
Anthropic has additionally cut Sonnet 5.5 cache-read pricing by 50% to $0.10 per million tokens and introduced monthly API credits of up to $500 for some paid subscribers. Stay Connected with ProPakistani.
Obtain the latest tech news, telecom insights, and product launches wherever you prefer. Follow on Google Discover. Follow on Google News Join WhatsApp. See more ProPakistani stories in Google Search and Top Stories.
Technology and Automotive Specialist covering the latest cars, smartphones, AI breakthroughs, and.
For now, claude Haiku 5.5 With 90% Discount is Here remains the part of the story worth watching, and further updates are likely as more details are confirmed.




