BRIGHT EVIDENCE PACK / Deployed
Haiku 5.5 cuts API prices for small teams
Anthropic’s new small model lowers API prices sharply. Teams still need to watch prompt length, token counts and the cost of getting a usable answer.
Canonical Bright record · JSON evidence pack · Key-facts embed
Dates and assessment
- Source published
- 2026-10-07
- Bright published
- 2026-10-07
- Substantive update
- None recorded
- Evidence state
- Deployed
- Independent verification
- Not established by this source review
- Last source review
- 2026-10-07
The claim in context
The human problem
Small teams need routine AI tasks to produce usable results at sustainable costs.
The prior constraint
Repeated API calls can accumulate costs, and unit rates alone do not describe the cost of a finished job.
AI’s actual role
Haiku 5.5 is a small language model for tasks such as classification, extraction and first-pass summaries.
The documented result
Anthropic released Claude Haiku 5.5 on October 7, with a substantial price cut for the short, repeated jobs that can add up across a business. For a small team, the opportunity is to test more routine work: sorting support tickets, extracting fields from documents or making first-pass summaries before someone reads the originals.
Why it may matter
Lower unit prices could make bounded workflows affordable to test and use regularly while keeping human quality checks.
Limitations
- On Anthropic’s platform, prompts of up to and including 100,000 tokens cost $0.10 per million input tokens and $0.50 per million output tokens. Above that prompt length, the rates rise to $0.50 and $2.50. Haiku 4.5’s base rates were $1 and $5. These are US-dollar API charges, measured in tokens, rather than a price for a chat subscription.
- That is a 90% unit-price reduction for the shorter tier and 50% for the longer one. Anthropic’s separate estimate of about 75% lower average run costs reflects its previous request mix and changes in token consumption. It is a vendor estimate, not a savings guarantee for a particular app.
- A simple illustration: 10,000 calls billed at 2,000 input tokens and 500 output tokens each would incur $4.50 in base model charges at the lower rates, against $45 at Haiku 4.5 rates. This compares equal billable token counts. It excludes retries, tools, storage and other operating costs; it does not predict the bill for processing the same text.
- Anthropic’s migration guide says the same input text produces approximately 30% more tokens with Haiku 5.5’s newer tokenizer, depending on the content. A prompt that used to fit below a budget may cross it after migration. Teams should count their real prompts with the new model, check output limits and update thinking settings rather than only changing the model name.
- Caching and batching offer further options. At the shorter tier, cache reads are $0.01 per million tokens; five-minute cache writes are $0.125 and one-hour writes $0.20. The longer-tier equivalents are $0.05, $0.625 and $1. The asynchronous Batch API halves standard input and output prices. Cloud endpoint choices and paid tools can add charges, so these figures are not an all-in quotation.
- Haiku 5.5 supports a one-million-token context window and adjustable effort, with medium as the default. Availability includes the Claude API, Amazon Bedrock, Google Cloud and Microsoft Foundry; features vary by platform. Anthropic’s launch table reports 39.2% on Terminal-Bench 4.0 against Sonnet 5.5’s 70.6%. Those are vendor-reported benchmark results, not Bright testing; headline scores do not establish performance at the default setting. They support retaining stronger models for demanding multi-step work; effort settings also affect the cost-quality trade-off.
- Anthropic also lists Haiku 5.5 for users across Free, Pro, Max, Team and Enterprise plans. New monthly API credits roll out over several days for Max and Team subscribers: $100 for Max 5x, $200 for Max 20x, and a Team pool capped at $500, based on seat types. New subscribers must wait seven days to claim. Unused credits expire each billing cycle. They do not increase Claude or Claude Code plan limits, cover interactive Claude Code or apply to partner-cloud usage. Free, Pro and Enterprise plans are ineligible for these credits.
Original evidence
- Introducing Claude Haiku 5.5 · institution
- Claude pricing · institution
- Pricing · institution
- Claude Haiku 5.5 migration guide · institution
- Claude Haiku 5.5 overview · institution
- Claude Haiku · institution
- Monthly API credits for Max and Team plans · institution
- Introducing Claude Haiku 5.5 on AWS · institution
- Claude Haiku 5.5 on Google Cloud · institution
- Claude in Microsoft Foundry · institution
Attribution
Credit Bright AI Future and link the canonical Bright record.
- Link to the canonical Bright record.
- Keep material limitations with the claim they qualify.
- Link to the original evidence when repeating a substantive claim.
- Do not describe a source check or organization-reported result as independent verification.
Linked source material, quotations, trademarks and media remain subject to their owners’ terms. No reuse right is granted for third-party media.
