Guide · 10 steps

Which Claude model should you use? The current lineup and official prices

Opus 5.5 is Anthropic's default pick, Haiku 5.5 (October 7, 2026) is far cheaper than Haiku 4.5, and Fable 5.1 is for the hardest jobs. The verified lineup, every price, and how to choose.

By ShajanthanReviewed 9 min read
ByShajanthanFounder & Editor
Published
Reading9 MIN
Versions coveredClaude API: Haiku 5.5, Sonnet 5.5, Opus 5.5, Fable 5.1 (as of Oct 9, 2026)
Diagram of the four current Claude models arranged as steps by price, from Haiku 5.5 to Fable 5.1
In 20 seconds
  1. Anthropic's current lineup is Fable 5.1 ($10/$50 per million tokens), Opus 5.5 ($4/$20), Sonnet 5.5 ($2/$10) and Haiku 5.5 (from $0.10/$0.50). All four have a 1M-token context window.
  2. Anthropic recommends Opus 5.5 for most work. Use Fable 5.1 only when Opus 5.5 at higher effort still falls short, and Haiku 5.5 for high-volume, latency-sensitive jobs.
  3. Watch the details. Haiku 5.5 costs five times more once a prompt passes 100K tokens, newer models produce about 30% more tokens for the same text, and Sonnet 4.5 retires on November 30, 2026.
Contents

Anthropic shipped three new Claude models in 16 days: Claude Opus 5.5 on September 22, 2026, Claude Sonnet 5.5 on September 28, and Claude Haiku 5.5 on October 7. Claude Fable 5.1, released September 1, sits above all three. The result is a four-model lineup with a clear default and some pricing quirks you should know about before you switch a production workload. This guide lists what's current, what each model costs according to Anthropic's own pricing page, and how to choose. It also covers the API calls you need to check models and control spend. If you're weighing Claude against OpenAI's newest models, read our coverage of the GPT-6 launch alongside this.

Last verified: October 8, 2026 (prices and model status change often, so recheck the linked official pages before you commit).

The short answer

  • Default to Claude Opus 5.5. Anthropic's models overview recommends it "for most workloads". It is also cheaper per token than the Opus it replaces.
  • Move to Claude Fable 5.1 only for demanding reasoning and long-horizon agent work, or, in Anthropic's words, "when Opus 5.5 at higher effort still falls short on your evals". It costs 2.5 times as much as Opus 5.5.
  • Use Claude Sonnet 5.5 when you need speed and lower cost but still want strong coding and agent performance. It costs half as much as Opus 5.5.
  • Use Claude Haiku 5.5 for classification, extraction, routing, summaries and subagents. It is very cheap for prompts under 100,000 tokens.
  • Claude Mythos 5.1 is not something you can simply choose. Access is restricted to Project Glasswing participants and, since October 6, 2026, organizations admitted to Anthropic's Cyber Verification Program.

The current lineup at a glance

All four current models accept text and image input, support tool use, have a 1M-token context window and 128K maximum output, and have a "reliable knowledge cutoff" of June 2026, according to Anthropic's models overview.

ModelAPI IDLatency (Anthropic's label)Default effortThinkingReleased
Claude Fable 5.1claude-fable-5-1SlowerhighAdaptive, always onSep 1, 2026
Claude Opus 5.5claude-opus-5-5ModeratemediumAdaptive, always onSep 22, 2026
Claude Sonnet 5.5claude-sonnet-5-5FasthighAdaptiveSep 28, 2026
Claude Haiku 5.5claude-haiku-5-5FastestmediumAdaptiveOct 7, 2026

"Always on" matters. You can't turn thinking off on Opus 5.5 (the September 22 release notes list this as a breaking change), so you pay for some reasoning tokens even on simple requests. The effort setting is how you control that. More on this below.

Official API prices

Prices are in US dollars per million tokens (MTok), from Anthropic's pricing page. "Cache write" is the 5-minute cache; a 1-hour cache write costs 2x base input.

ModelInputOutputCache write (5 min)Cache readBatch input / output
Fable 5.1$10$50$12.50$0.25$5 / $25
Opus 5.5$4$20$5$0.20$2 / $10
Sonnet 5.5$2$10$2.50$0.10$1 / $5
Haiku 5.5, prompts ≤100K tokens$0.10$0.50$0.125$0.01$0.05 / $0.25
Haiku 5.5, prompts >100K tokens$0.50$2.50$0.625$0.05$0.25 / $1.25
Mythos 5.1 (restricted)$10$50$12.50$0.25$5 / $25

Last verified: October 8, 2026

What to notice:

  • Opus got cheaper. Opus 5 was $5/$25; Opus 5.5 is $4/$20. Anthropic also claims Opus 5.5 costs "about 40% lower" than Opus 5 on typical workloads at default settings, because it uses fewer tokens per task. That is a vendor claim, so measure it on your own traffic.
  • Sonnet 5.5's cache reads were halved on October 7, 2026, from $0.20 to $0.10. Anthropic's September 28 launch page still shows $0.20, but the pricing page and release notes show $0.10.
  • Haiku 5.5 is priced by prompt length. Under 100K tokens it costs a tenth of Haiku 4.5 ($1/$5). Over 100K tokens the whole request moves to the higher tier, five times the price. Anthropic says Haiku 5.5 costs about 75% less to run than Haiku 4.5 on average: about 90% less on short requests and about 50% less on long ones.
  • Cache reads are discounted more than usual on the new models. The standard cache-hit rate is 0.1x base input. Opus 5.5 and Sonnet 5.5 charge 0.05x, and Fable 5.1 and Mythos 5.1 charge 0.025x.
  • Batch and caching stack. The Batch API halves input and output prices, and the discount combines with prompt caching.
  • Extras. Fast mode for Opus 5.5 (research preview, Claude API only) costs $8 input and $40 output, which is 2x standard; Anthropic says it is up to 2.5x faster. US-only inference (inference_geo: "us") multiplies all token prices by 1.1x. Web search costs $10 per 1,000 searches plus tokens.

The other cost you won't see in the table is the tokenizer. Anthropic's pricing page says Opus 4.7 and later models "produce approximately 30% more tokens for the same text", while per-token prices are unchanged. If you're moving from Sonnet 4.6 or Haiku 4.5, a lower per-token price doesn't mean a proportionally lower bill. We work through the arithmetic in why token pricing isn't the story.

How to choose: a decision guide by use case

If your workload is…Start withWhyEscalate to
Long-running agentic coding, complex refactorsOpus 5.5 (medium effort)Anthropic's default recommendationOpus 5.5 at high/xhigh, then Fable 5.1
Hardest reasoning or multi-day agent runs where failure is costlyFable 5.1Anthropic positions it for "demanding reasoning and long-horizon agentic work"n/a
Everyday coding help, chat assistants, document Q&ASonnet 5.5Half Opus 5.5's price; Anthropic calls it "the best combination of speed and intelligence"Opus 5.5
Classification, extraction, tagging, routingHaiku 5.5 (low/medium effort)$0.10/$0.50 for prompts under 100K tokensSonnet 5.5
Subagents inside a coding agent (search, summarize, compact)Haiku 5.5Anthropic recommends pairing it with Opus 5.5 or Sonnet 5.5Sonnet 5.5
Summarizing very long documents (>100K tokens)Sonnet 5.5 or Haiku 5.5Haiku's higher tier ($0.50/$2.50) still undercuts Sonnet, but price bothOpus 5.5
Overnight bulk jobs (evals, backfills)Any, via Batch API50% offn/a
Defensive security research needing fewer cyber blocksApply to the Cyber Verification ProgramTiers include Mythos 5.1, Opus 5.5 and Sonnet 5.5n/a

Two cautions. First, Anthropic's benchmark numbers for these models are vendor-reported, so test on your own tasks. Second, mixing models in one product is a design decision of its own. Our explainer on model routing covers when sending easy requests to Haiku and hard ones to Opus actually saves money.

Step by step: check models and set effort via the API

The examples use the official Python SDK. We documented against anthropic 1.12.1 (released October 8, 2026, requires Python 3.10+).

1. Install and set your key.

BASH
pip install --upgrade anthropic
export ANTHROPIC_API_KEY="sk-ant-..."

2. List the models your key can use. The Models API returns each model's id, display_name, max_input_tokens, max_tokens and a capabilities object. By default it lists active and deprecated models.

BASH
curl https://api.anthropic.com/v1/models \
    -H 'anthropic-version: 2023-06-01' \
    -H "X-Api-Key: $ANTHROPIC_API_KEY"

Read max_input_tokens and capabilities from the API rather than hard-coding them. Anthropic adds fields regularly. In October 2026 it added capabilities.thinking.types.disabled and capabilities.server_tools.

3. Count tokens before you estimate costs. Token counting is free (it has its own rate limits) and uses the tokenizer of the model you name. That matters because of the ~30% tokenizer difference.

PYTHON
import anthropic

client = anthropic.Anthropic()
count = client.messages.count_tokens(
    model="claude-opus-5-5",
    system="You are a scientist",
    messages=[{"role": "user", "content": "Hello, Claude"}],
)
print(count.input_tokens)

4. Set effort to control cost and latency. Effort applies to all output tokens, including thinking and tool calls. Anthropic calls it "a behavioral signal, not a strict token budget". The levels are low, medium, high, xhigh and max. Opus 5.5 and Haiku 5.5 default to medium, and most other models default to high.

PYTHON
response = client.messages.create(
    model="claude-opus-5-5",
    max_tokens=4096,
    messages=[{"role": "user", "content": "Review this migration plan for risks: ..."}],
    output_config={"effort": "medium"},
)
for block in response.content:
    if block.type == "text":
        print(block.text)

What Anthropic's docs say to expect: a normal Messages API response. On Opus 5.5 it may include thinking blocks before the text, because thinking can't be disabled. Check response.usage to see what you were billed for.

If you want a full agent loop rather than single calls, our guide to building your first agent with the Claude Agent SDK picks up from here.

Troubleshooting common migration errors

The new models changed several defaults. These issues come from Anthropic's release notes:

  • 400 error with tool_choice: "any" or "tool". Forced tool use is rejected on Fable 5.1, Opus 5.5 and Sonnet 5.5. Use auto with strict tool use instead.
  • 400 error when disabling thinking on Opus 5.5. You can't. Lower the effort instead. On Sonnet 5.5, the release notes say to use thinking: {"type": "between_tools"} at high effort or below rather than "disabled".
  • 400 error with budget_tokens on Haiku 5.5. Manual thinking budgets were removed. Use effort.
  • inference_geo returns 400. It is supported only on Claude 4.6 and later models.
  • Bills higher than expected after migrating. Recount your prompts with the new model's tokenizer (step 3) and check usage for thinking tokens.

In the Claude apps

The app plans are priced separately from the API. According to claude.com/pricing:

  • Free: no Opus and no Claude Code.
  • Pro: $20/month, or $17/month billed annually. Adds Opus and Claude Code.
  • Max: from $100/month, with 5x or 20x Pro usage.
  • Team: standard seats are $25/month ($20 billed annually) and premium seats are $125/month ($100 billed annually).

Fable models are on all paid plans, but on Pro and standard seats they run on pay-as-you-go usage credits from the start. On Max and premium seats you can spend up to 50% of your weekly limit on Fable, per Anthropic's Help Center. (That is standard plan access. An earlier Fable 5 promotion ended on July 19, 2026, and Fable 5.1 was never part of it.) Since October 7, Max plans also include monthly API credits ($100 for Max 5x, $200 for Max 20x), and Team plans get up to $500/month pooled. Anthropic's pages don't say which app plans offer Haiku 5.5 in the model picker. For the app itself, see our Claude app explainer.

In Claude Code on the Anthropic API, the opus, sonnet and haiku aliases point to Opus 5.5, Sonnet 5.5 and Haiku 5.5. Other providers differ, according to Claude Code's model-configuration docs (checked October 9, 2026). On Amazon Bedrock and Google Cloud, opus is Opus 5.5 but sonnet and haiku still resolve to Sonnet 4.5 and Haiku 4.5. On Claude Platform on AWS they resolve to Opus 5.5, Sonnet 4.6 and Haiku 4.5. On Microsoft Foundry they resolve to Opus 4.6, Sonnet 4.5 and Haiku 4.5. Pin full model IDs if you depend on a specific model.

Cloud availability

Anthropic says Opus 5.5, Sonnet 5.5 and Haiku 5.5 are available on Amazon Web Services, Google Cloud and Microsoft Azure. The Haiku 5.5 release note lists Bedrock, Claude Platform on AWS, Google Cloud and Microsoft Foundry. On Bedrock and Google Cloud the cloud provider sets the price. Anthropic's docs say regional endpoints there cost 10% more than global ones. Fast mode is available only on Anthropic's own API.

Older models: what's still available and what's retiring

ModelStatusRetirement
Fable 5, Opus 5, Sonnet 5ActiveNot before Jun 9, Jul 24, Jun 30, 2027 respectively
Opus 4.8, 4.7, 4.6ActiveNot before May 28, Apr 16, Feb 5, 2027
Sonnet 4.6ActiveNot before Feb 17, 2027
Opus 4.5ActiveNot before Nov 24, 2026
Haiku 4.5ActiveNot before Oct 15, 2026
Sonnet 4.5Deprecated Sep 30, 2026Nov 30, 2026 (replacement: Sonnet 5.5)
Opus 4.1, Opus 4, Sonnet 4RetiredAug 5 and Jun 15, 2026

Last verified: October 9, 2026 (Anthropic's deprecations page still lists Haiku 4.5 as active, with no deprecation notice)

If you still run Sonnet 4.5, you have until November 30, 2026. Haiku 4.5 can be retired no sooner than October 15, 2026. Note that "not sooner than" is not an announced retirement, but Haiku 5.5 is both cheaper and, per Anthropic, more capable, so there's little reason to wait. For background on Anthropic's model strategy, see Anthropic explained.

What's confirmed and what isn't

  • Confirmed (Anthropic's docs): model IDs, prices, context windows, maximum output, default effort, retirement dates, and Mythos 5.1 access routes.
  • Vendor claims: "75% cheaper than Haiku 4.5", "about 40% lower cost than Opus 5", Sonnet 5.5 "costs up to 30% less for most work", and all benchmark scores.
  • Not stated: which consumer plans offer Haiku 5.5 in the app, and Bedrock/Vertex list prices (set by those providers).
Did this guide work for you?

About this storyBased on the sources linked below. Editorial standards

Was this useful?Report an error
Comments
0

More on Claude & models

The Week in AI

New guides and explainers, every Friday.

0