Skip to content

model-releases

Haiku 5.5 cuts token prices 90% on prompts under 100k

Claude News

On short prompts, input now costs $0.10 per million tokens against $1.00 on Haiku 4.5, and output costs $0.50 against $5.00. Anthropic puts the average saving at around 75%. Prompts over 100k get a smaller cut, and The News Stack reports that an updated tokenizer uses slightly more tokens per task.

The capability gap shows on Terminal-Bench 4.0, where Haiku 5.5 scores 39.2% and Haiku 4.5 scored 0.0%. Anthropic still recommends Sonnet 5.5 and Opus 5.5 for hard agentic coding. It aims Haiku 5.5, the first Haiku with an effort setting, at compaction, summaries and subagent work.

Anthropic also halved cache reads on Sonnet 5.5, from $0.20 to $0.10 per million tokens, and says that makes most agentic work around 20% cheaper. Max and Team subscribers get monthly API credits this week. Only the token count per task went up.

Related stories

  1. Sonnet 5.5 edges past Opus 5.5 on Terminal-Bench 4.0
  2. Opus 5.5 matches Fable 5.1 on most work for less money
  3. Cache reads get 75% cheaper on Claude Fable 5.1
  4. Claude Opus 5 lands at Opus 4.8 pricing and becomes the Max default
  5. Anthropic releases Claude Fable 5 and Mythos 5 models
  6. Claude Startups widens its door, with up to $45K in perks

Comments

No comments yet. Be the first.

Join the conversation

Sign in with Google to leave a comment. Your name and avatar come from your Google profile, and the comment appears after moderation.

We only use your name and avatar from Google. We never store your email address.