Skip to main content

Claude Models: All 28 Models Compared

All 28 Claude models compared: Context, pricing, strengths and use cases, plus recommendation for which model excels for text, code, or research.

FHFinn Hillebrandt
AI Tools
Claude Models: All 28 Models Compared
Links marked with * are affiliate links. If a purchase is made through such links, we receive a commission.

28 models. That's how many Claude versions Anthropic has released since March 2023. And keeping track of which model does what, which ones are still active, and what they cost? Not exactly straightforward.

I've used every generation of Claude since version 1.0, from the early 9K context window days to today's 1M-token models like Sonnet 5 and Opus 4.8. The progress has been remarkable.

In this article, I'll break down all 28 Claude models, compare their features, pricing, and availability, and give you a clear recommendation for which one to pick.

TL;DRKey Takeaways
  • Claude Fable 5 (Mythos class) is back online since early July 2026 and remains Anthropic's most capable model, but since July 24, 2026 Claude Opus 5 is the recommended choice for agentic coding and enterprise work: close to Fable 5 at half the price, with a 1M token context window and a May 2026 knowledge cutoff. Its predecessor Claude 4.8 Opus (88.6% on SWE-bench Verified) is now a legacy model
  • Claude Sonnet 5 (late June 2026) is the new default model for Free and Pro users, with a 1M token context window at an introductory price of $2/$10 per million tokens through the end of August 2026
  • All modern Claude models (version 3+) have at least a 200,000 token context window and multimodal capabilities for text and images

What Are Claude Models?

Claude models are the AI models behind Anthropic's chatbot. What sets them apart from GPT or Gemini?

A few things:

  • Constitutional AI: Claude is trained with a set of principles that ensure responses are helpful, harmless, and honest. More details in Anthropic's Constitutional AI research.
  • Large Context Windows: Modern Claude models can process up to 1 million tokens (Fable 5, Sonnet 5, Opus 5, Opus 4.8, 4.7 and the 4.6 models), and every model from Claude 3 onwards handles at least 200,000 tokens
  • Multimodal Capabilities: From Claude 3 onwards, models can understand and analyze both text and images.
  • Hybrid Reasoning: The latest models offer both fast responses and deeper "thinking" for complex problems.

Anthropic has deliberately chosen a different strategy than OpenAI: instead of focusing on maximum performance, safety and control of the AI are the priority.

Before we look at the individual models, here are the key milestones of Claude's evolution from 2023 to today.

March 2023
Claude 1.0
First model with Constitutional AI training and a 9,000 token context window
July 2023
Claude 2.0
First publicly available Claude with a 100,000 token context window
March 2024
Claude 3 family
Haiku, Sonnet, and Opus, the first multimodal generation with a 200,000 token context
June 2024
Claude 3.5 Sonnet
Surpasses Claude 3 Opus in many benchmarks at a much lower price
February 2025
Claude 3.7 Sonnet
First Claude with Hybrid Reasoning and controllable thinking time
May 2025
Claude 4 Opus and Sonnet
Hybrid Reasoning family, Opus reaches 72.5% on SWE-bench
September 2025
Claude 4.5 generation
4.5 Sonnet beats the Opus variants for the first time, 4.5 Haiku follows in October
February 2026
Claude 4.6 Opus and Sonnet
1 million token context window and Agent Teams
April 2026
Claude 4.7 Opus
87.6% on SWE-bench Verified, 1M context becomes standard without a surcharge
May 2026
Claude 4.8 Opus
88.6% on SWE-bench Verified, dynamic workflows and effort control, since succeeded by Opus 5
June 2026
Claude Sonnet 5
New default model for Free and Pro plans with a 1M token context window
June 2026
Claude Fable 5 and Mythos 5
New Mythos class above Opus, back online since July 1, 2026 after a brief suspension
July 2026
Claude Opus 5
New recommended coding model, close to Fable 5 at half the price, succeeds Opus 4.8

Comparison of All Claude Models

Here is a detailed overview of all Claude models with their key properties:

Column groups:
Model
Release
Status
Input
Output
Cache Write
Cache Read
Claude 3 Haiku03/2024Discontinued$0.25$1.25$0.3$0.03
Claude 3.5 Sonnet06/2024Discontinued$3$15$3.75$0.3
Claude 3.5 Haiku11/2024Discontinued$0.8$4$1$0.08
Claude 3.7 Sonnet02/2025Discontinued$3$15$3.75$0.3
Claude 4 Sonnet05/2025Discontinued$3$15$3.75$0.3
Claude 4 Opus05/2025Discontinued$15$75$18.75$1.5
Claude 4.1 Opus08/2025Discontinued$15$75$18.75$1.5
Claude 4.5 Sonnet09/2025Active$3$15$3.75$0.3
Claude 4.5 Haiku10/2025Active$1$5$1.25$0.1
Claude 4.5 Opus11/2025Active$5$25$6.25$0.5
Claude 4.6 Opus02/2026Active$5$25$6.25$0.5
Claude 4.6 Sonnet02/2026Active$3$15$3.75$0.3
Claude 4.7 Opus04/2026Active$5$25$6.25$0.5
Claude Opus 507/2026Active$5$25$6.25$0.5
Claude 4.8 Opus05/2026Legacy$5$25$6.25$0.5
Claude Sonnet 506/2026Active$2$10$2.5$0.2
Claude Fable 506/2026Active$10$50$12.5$1

Claude Fable 5 and Claude Mythos 5

Released: June 2026

Claude Fable 5 and Claude Mythos 5 launched on June 9, 2026 as a new tier above the Opus class that Anthropic calls "Mythos-class". The two models share the same underlying system and the same specs. The difference is the safety classifiers.

Key Features:

  • 1 million token context window and 128,000 output tokens
  • Adaptive thinking as the only reasoning mode, controlled via the effort parameter
  • Safety classifiers on Fable 5: classified refusals fall back to another Claude model (Opus 4.8 by default) and the refusal response is not billed
  • 30-day data retention (zero data retention not available, Covered Models)
  • Raw chain-of-thought is never returned, only a summary via `thinking.display`
  • Pricing: $10 per 1M input tokens, $50 per 1M output tokens (double the Opus 4.8 rate)

Fable 5 vs. Mythos 5: the difference. Claude Fable 5 is the openly available model with active safety classifiers. It is generally available via the Claude API, Amazon Bedrock, Google Cloud Vertex AI, Microsoft Foundry, and Claude Platform on AWS.

Claude Mythos 5 is the same model without those classifiers. Access runs through Project Glasswing only, for vetted customers in cybersecurity, biomedical research, and critical infrastructure. Customers without Glasswing access use Fable 5.

Brief suspension in June, back online since July 1. Days after the launch, a US export control directive paused access to both Fable 5 and Mythos 5 after security researchers found a jailbreak. Access was restored on July 1, 2026. For subscription plans (Pro, Max, Team, seat-based Enterprise), Fable 5 was included at no extra cost between June 9 and June 22; usage credits were planned to apply from June 23 onward.

Claude Sonnet 5

Released: June 30, 2026

Claude Sonnet 5 became the new default model for Free and Pro users in the Claude apps on June 30, 2026, replacing Claude 4.6 Sonnet as the recommended Sonnet model. Anthropic says it rivals Opus 4.8 at a lower cost.

Key Features:

  • 1 million token context window and 128,000 output tokens, now the default in the Claude apps
  • Adaptive thinking with an effort parameter that defaults to "high"
  • New tokenizer (the same one used by Opus 4.7): about 30% more tokens for the same text than older Sonnet models
  • No priority tier available
  • Knowledge Cutoff: January 2026
  • Pricing: introductory price of $2 per 1M input tokens and $10 per 1M output tokens through August 31, 2026, then $3 per 1M input tokens and $15 per 1M output tokens

Sonnet 5 is now the model most Claude users actually see day to day, since it powers the Free and Pro plans by default. For the price of a Sonnet model, you get a 1-million-token context window that used to be reserved for the Opus models.

Availability: Claude Sonnet 5 is available via the Claude API, Claude.ai Web Interface, and the Claude iOS and Android apps.

Claude Opus 5

Released: July 24, 2026

Claude Opus 5 is Anthropic's new model for complex agentic coding and enterprise work, succeeding Claude 4.8 Opus as the recommended Opus model. Anthropic describes it as "close to the frontier intelligence of Claude Fable 5 at half the price." Fable 5 remains the model with the highest capability tier; Opus 5 is the cheaper near-frontier choice for coding.

Key Features:

  • 1 million token context window and 128,000 output tokens (up to 300,000 via the batch beta)
  • Knowledge cutoff May 2026, the most current of any Claude model
  • New default model on Claude Max, strongest available model on Claude Pro
  • Leads Frontier-Bench v0.1 and surpasses Fable 5 on the OSWorld 2.0 benchmark at roughly a third of the cost
  • CursorBench 3.2: within 0.5% of Fable 5
  • Pricing: $5 per 1M input tokens, $25 per 1M output tokens (half the price of Fable 5, the same as the now-legacy Opus 4.8). Fast mode: double the base price. Batch API: $2.50/$12.50
  • Availability: Claude API, Claude.ai Web Interface, Claude iOS and Android apps, and as the new default model on Claude Max and strongest model on Claude Pro

The practical core is the price-to-performance ratio: Opus 5 comes close to Fable 5's capabilities, according to Anthropic, at half the price. Compared with Opus 4.8, Anthropic describes it as "greatly improved performance for the same cost." On ARC-AGI 3, Anthropic reports roughly triple the score of the next-best model, and about 1.5x on Zapier's AutomationBench. It gains 10.2 percentage points over Opus 4.8 on chemistry tasks and 7.7 points on protein-related tasks. On cybersecurity tasks, Opus 5 still trails Claude Mythos 5. Anthropic has not published an official SWE-bench Verified score for Opus 5.

Claude 4.8 Opus

Released: May 2026

Claude 4.8 Opus was Anthropic's recommended model for agentic coding and enterprise workloads from May to July 2026, before being succeeded by Claude Opus 5. It's now a legacy model but remains available. The jump over Opus 4.7 was smaller than the one from 4.6 to 4.7, but it was real. The biggest gain was once again in coding: SWE-bench Verified rose to 88.6%, up from 87.6% on 4.7. The harder SWE-bench Pro climbed to 69.2%, up from 64.3%.

Key Features:

  • 1 million token context window and 128,000 output tokens, same as Opus 4.7
  • Dynamic workflows: Claude can spin up hundreds of parallel subagents in a single session to break down very large tasks
  • Effort control: you decide how much effort Claude puts into a task (the "high" tier is the default)
  • Cheaper fast mode: runs at 2.5x the speed and is three times cheaper than on previous models
  • 88.6% on SWE-bench Verified, 69.2% on SWE-bench Pro, 93.6% on GPQA Diamond
  • Pricing: $5 per 1M input tokens, $25 per 1M output tokens, unchanged from Opus 4.7
  • Availability: Claude API, Amazon Bedrock, Google Cloud Vertex AI, Microsoft Foundry, Claude apps

The most practically useful improvement was reliability. Anthropic reported that Opus 4.8 was around four times less likely than its predecessor to let flaws in its own code pass unremarked. On pure knowledge it landed roughly level with 4.7, and on GPQA Diamond it even dipped slightly (93.6% vs. 94.2%). The gain was clearly in agentic work, not raw knowledge.

Claude 4.7 Opus

Released: April 2026

Claude 4.7 Opus was Anthropic's flagship from April to May 2026 and remains active and in use. It built directly on Opus 4.6 and took a clear leap forward in coding: 87.6% on SWE-bench Verified (up from 80.8% on 4.6) and 64.3% on the harder SWE-bench Pro.

Key Features:

  • 1 million token context window (now standard, no longer beta)
  • 128,000 output tokens
  • New tokenizer: The same text now splits into 1.0 to 1.35x more tokens, so real cost can rise slightly at the same per-token price
  • "xhigh" reasoning mode: New effort tier between "high" and "max" for finer control over thinking depth
  • Image resolution up to 3.75 MP, over 3x prior Claude models; long edge up to 2,576 pixels
  • 87.6% on SWE-bench Verified, 94.2% on GPQA Diamond, 78% on OSWorld-Verified, 92.4% on MMLU
  • Pricing: $5 per 1M input tokens, $25 per 1M output tokens, same as 4.6 but without the tiered surcharge above 200K
  • Availability: Claude API, Amazon Bedrock, Google Cloud Vertex AI, Microsoft Foundry, Claude apps

Claude 4.6 Opus

Released: February 2026

Claude 4.6 Opus was the flagship at release and the first Claude model to offer a 1M-token context window alongside Agent Teams for parallel task coordination. It has since been succeeded by Claude 4.7 Opus but remains active and in use.

Key Features:

  • 1 million token context window
  • Agent Teams for parallel task execution
  • 80.8% on SWE-bench Verified, 91.3% on GPQA Diamond
  • 200,000 token context window with up to 64,000 output tokens
  • Multimodal capabilities: Processes text, images, and documents

Claude 4.6 Sonnet

Released: February 2026

Claude 4.6 Sonnet continues the Sonnet tradition of offering excellent performance at a balanced price point. It replaced Claude 4.5 Sonnet as the recommended model for most professional applications and has since been succeeded by Claude Sonnet 5 (June 2026) as Anthropic's default Sonnet model.

Key Features:

  • Enhanced reasoning and coding capabilities over Claude 4.5 Sonnet
  • Improved instruction following compared to Claude 4.5 Sonnet
  • 200,000 token context window with up to 64,000 output tokens
  • Multimodal capabilities: Processes text, images, and documents
  • Best balance of performance and cost for most use cases

Claude 4.5 Haiku

Released: October 15, 2025

Claude 4.5 Haiku is the latest Haiku model and marks a significant advancement in the cost-effective, fast AI category. It's the first Haiku model in version 4.5 and offers an impressive combination of speed, low cost, and surprisingly high performance.

Key Features:

  • Significantly improved performance: Surpasses Claude 3.5 Haiku in all benchmarks at similar costs
  • Fastest response times: Optimized for high-throughput applications with minimal latency
  • Enhanced coding capabilities: Strong performance on programming tasks for a Haiku model
  • 200,000 token context window with up to 64,000 output tokens
  • Multimodal capabilities: Processes text, images, and documents
  • Knowledge Cutoff: January 2025, Training Data through July 2025
  • API Pricing: $1 per 1 million input tokens, $5 per 1 million output tokens
  • Model String: claude-haiku-4-5-20251001
  • Safety Level: ASL-2 (AI Safety Level 2 Standard)

What makes Claude 4.5 Haiku special?

Claude 4.5 Haiku fills an important gap in Anthropic's portfolio: It offers significantly higher performance than its predecessor Claude 3.5 Haiku while maintaining the speed and low costs that Haiku models are known for. This makes it the perfect choice for applications requiring thousands or millions of API calls per day.

Particularly impressive is the coding performance: While earlier Haiku models were primarily intended for simple text tasks, Claude 4.5 Haiku can deliver surprisingly good results even on medium to complex programming tasks, at a fraction of the cost of Sonnet or Opus models.

Ideal Use Cases:

  • Chatbots and customer service: Fast, precise responses at low cost
  • Content moderation: High-throughput analysis of user-generated content
  • Simple automation: Code snippets, formatting, data extraction
  • Real-time applications: Where minimal latency is critical
  • Batch processing: Processing large volumes of documents or text

Availability: Claude 4.5 Haiku is available via the Claude API, Claude.ai Web Interface, Claude iOS and Android Apps, Amazon Bedrock, and Google Cloud Vertex AI.

Claude 4.5 Sonnet

Released: September 29, 2025

Claude 4.5 Sonnet (September 2025) was a significant breakthrough for Anthropic. It combines world-class coding performance with the attractive price of the Sonnet series: For the first time, a Sonnet model surpassed the Opus variants in the most important benchmarks. It has since been succeeded by Claude 4.6 Sonnet (February 2026).

Key Features:

  • State-of-the-art coding performance: Leading in SWE-bench Verified Benchmark, surpasses even Claude 4.1 Opus
  • Impressive focus ability: Can stay focused for 30+ hours on complex, multi-step tasks
  • OSWorld leadership: 61.4% performance in OSWorld benchmark (vs. 42.2% for Claude 4 Sonnet four months earlier)
  • 200,000 token context window with up to 64,000 output tokens
  • Multimodal capabilities: Processes text, images, and documents
  • Knowledge Cutoff: January 2025, Training Data through July 2025
  • API Pricing: $3 per 1 million input tokens, $15 per 1 million output tokens
  • Model String: claude-sonnet-4-5-20250929
  • Safety Level: ASL-3 (AI Safety Level 3 Standard)

What makes Claude 4.5 Sonnet special?

Claude 4.5 Sonnet marks a paradigm shift in Anthropic's model strategy. While previously Opus models offered the absolute peak performance, Anthropic has managed to develop a model with Claude 4.5 Sonnet that surpasses the Opus variants in most areas at one-fifth of the cost.

The most remarkable capability is "Extended Focus": Claude 4.5 Sonnet can work focused for over 30 hours on complex development tasks without losing quality or coherence. This makes it ideal for large refactoring projects, complex system architectures, and multi-step debugging sessions.

Claude 4.1 Opus

Released: August 2025

Claude 4.1 Opus was the most advanced Opus model of its generation, building on the Claude 4 family, which represented a quantum leap in AI development. These Hybrid-Reasoning models can switch between two modes: lightning-fast responses for simple questions and extensive "thinking" for complex problems. It has since been succeeded by Claude 4.5 Opus and Claude 4.6 Opus.

Key Features:

  • Improved reasoning capabilities compared to Claude 4 Opus
  • Strong coding assistant with a 74.5% success rate on SWE-bench Verified
  • Extended Hybrid-Reasoning functions with even deeper "thinking"
  • Significantly reduced hallucination rate (under 2% in benchmarks)
  • API Pricing: $15 per 1 million input tokens, $75 per 1 million output tokens
  • 200,000 token context window
  • Improved multimodal capabilities, especially for technical diagrams

Claude 4 Opus

Released: May 2025

Claude 4 Opus was the original flagship model of the Claude 4 family, which introduced Hybrid-Reasoning technology.

Key Features:

  • Previous best coding assistant with 72.5% success rate in SWE-bench (now surpassed by 4.1 Opus)
  • Hybrid-Reasoning: Users can choose between fast and thoughtful responses
  • Excellent for complex scientific, mathematical, and programming tasks
  • API Pricing: $15 per 1 million input tokens, $75 per 1 million output tokens
  • 200,000 token context window
  • Multimodal capabilities (text, images, documents)

Claude 4 Sonnet

Released: May 2025

Claude 4 Sonnet is the balanced version of the Claude 4 family and offers the best balance between performance and cost for most applications.

Key Features:

  • Also equipped with Hybrid-Reasoning capabilities
  • Better price-performance ratio than Opus
  • API Pricing: $3 per 1 million input tokens, $15 per 1 million output tokens
  • Ideal for most professional applications
  • Same technical foundation as Opus
  • 200,000 token context window

Claude 3.7 Sonnet

Released: February 24, 2025

Claude 3.7 Sonnet was the first model to introduce Anthropic's new Hybrid-Reasoning technology. It allows users to switch between fast responses and an "extended thinking mode."

Key Features:

  • First Claude with controllable thinking time
  • Toggle between "Fast" and "Think" mode
  • Improved performance on complex reasoning tasks
  • 200,000 token context window
  • Multimodal input (text and images)
  • Test model for later Claude 4 development

Claude 3.5 Haiku

Released: November 4, 2024

Claude 3.5 Haiku is the fastest and most cost-effective model in the Claude 3.5 series, which significantly improved upon the Claude 3 family foundations.

Key Features:

  • Fastest and most cost-effective model in the series
  • Optimized for high throughput rates
  • Impressive quality despite speed
  • Perfect for chatbots and simple automation
  • Multimodal capabilities at low cost
  • 200,000 token context window
  • API Pricing: $0.80 per 1 million input tokens, $4 per 1 million output tokens

Claude 3.5 Sonnet

Released: June 21, 2024

Claude 3.5 Sonnet was the first model in the improved 3.5 series and offered better performance at similar costs to Claude 3.

Key Features:

  • Surpasses Claude 3 Opus in many benchmarks
  • Significantly cheaper than Opus with comparable performance
  • Improved coding capabilities and visual understanding
  • Better workflow integration and tool use
  • 200,000 token context window
  • Ideal for productive applications
  • API Pricing: $3 per 1 million input tokens, $15 per 1 million output tokens

Claude 3 Opus

Released: March 2024

Claude 3 Opus was the most powerful model in the Claude 3 family, which marked a turning point: for the first time, Claude models could process both text and images.

Key Features:

  • The most powerful model in the series
  • First Claude version with true multimodality
  • Excellent for complex analysis and reasoning tasks
  • 200,000 token context window
  • Discontinued on June 30, 2025 (Retirement on January 5, 2026)

Claude 3 Sonnet

Released: March 2024

Claude 3 Sonnet was the balanced model in the Claude 3 family, which introduced the Japanese poetry-inspired naming convention.

Key Features:

  • Balanced model between performance and cost
  • Good all-round capabilities for most applications
  • Multimodal capabilities (text and images)
  • 200,000 token context window
  • Was replaced by Claude 3.5 Sonnet on July 21, 2025

Claude 3 Haiku

Released: March 2024

Claude 3 Haiku was the fastest and cheapest model in the Claude 3 family.

Key Features:

  • Fastest and cheapest model in the Claude 3 series
  • First multimodal Claude version for high-throughput applications
  • Discontinued in April 2026
  • 200,000 token context window
  • Multimodal capabilities (text and images)
  • API Pricing: $0.25 per 1 million input tokens, $1.25 per 1 million output tokens

Claude 2.1

Released: November 2023

Claude 2.1 was the improved version of Claude 2, the first publicly available Claude.

Key Features:

  • 200,000 token context window (twice as large as Claude 2.0)
  • Reduced hallucinations compared to Claude 2.0
  • Improved accuracy for long documents
  • Discontinued

Claude 2.0

Released: July 2023

Claude 2.0 was the first publicly available Claude and brought significant improvements over the first generation.

Key Features:

  • 100,000 token context window
  • Improved coding capabilities
  • Better mathematical reasoning
  • Available via API and Claude.ai website
  • Discontinued

Claude 1.3

Released: June 2023

Claude 1.3 was the final version of the first Claude generation, laying the foundation for Anthropic's approach to AI safety.

Key Features:

  • Improved safety features through Constitutional AI
  • Better handling of complex instructions
  • Only available via API
  • Discontinued

Claude 1.2

Released: May 2023

Claude 1.2 was an intermediate version of the first Claude generation.

Key Features:

  • Incremental improvements over Claude 1.0
  • Better text understanding
  • Discontinued

Claude 1.0

Released: March 2023

Claude 1.0 was the very first model from Anthropic and laid the foundation for all later developments.

Key Features:

  • First model with Constitutional AI training
  • 9,000 token context window
  • Focus on safety and helpfulness
  • Foundation for all later developments
  • Discontinued

Which Claude Model Should You Choose?

The choice of the right Claude model depends on your specific requirements:

For Developers and Programmers

Claude Sonnet 5 (June 2026) is the current recommended default model for developers, building on the strong coding foundations of Claude 4.6 Sonnet with a 1M token context window at a Sonnet price point.

Alternative for special requirements: Claude Opus 5 (July 2026) is Anthropic's recommended coding model, close to Fable 5's frontier intelligence at half the price, with a 1M token context window and a May 2026 knowledge cutoff. For the single highest capability tier, Claude Fable 5 remains the top choice.

For Businesses and Professional Applications

Claude Sonnet 5 offers the best price-performance ratio for most business applications: it is the new default model for Free and Pro plans, brings a 1 million token context window, and starts at an introductory price before moving to the standard rate at the end of August 2026 (see the pricing table below for the exact figures). That makes it a strong fit for customer service, content creation, and data analysis. Its predecessor Claude 4 Sonnet was retired in June 2026.

For High-Volume Applications

Claude 4.5 Haiku (October 2025) is the newest and best Haiku model for chatbots, automated responses, and other applications that need fast, cost-effective responses with high quality. It offers significantly better performance than Claude 3.5 Haiku at a similar price point (see the pricing table below).

Price Comparison of All Claude Models

The following table shows a detailed overview of all Claude prices (all figures in $ per million tokens). All active models offer 50% discount when using the Batch API.

ModelClaude Fable 5
StatusActive
Input$10
Output$50
Cache Write$12.5
Cache Read$1
ModelClaude Sonnet 5
StatusActive
Input$2
Output$10
Cache Write$2.5
Cache Read$0.2
ModelClaude Opus 5
StatusActive
Input$5
Output$25
Cache Write$6.25
Cache Read$0.5
ModelClaude Opus 4.8
StatusLegacy
Input$5
Output$25
Cache Write$6.25
Cache Read$0.5
ModelClaude Opus 4.7
StatusActive
Input$5
Output$25
Cache Write$6.25
Cache Read$0.5
ModelClaude Opus 4.6
StatusActive
Input$5
Output$25
Cache Write$6.25
Cache Read$0.5
ModelClaude Sonnet 4.6
StatusActive
Input$3
Output$15
Cache Write$3.75
Cache Read$0.3
ModelClaude Sonnet 4.5
StatusActive
Input$3
Output$15
Cache Write$3.75
Cache Read$0.3
ModelClaude Haiku 4.5
StatusActive
Input$1
Output$5
Cache Write$1.25
Cache Read$0.1
ModelClaude Opus 4.1
StatusDiscontinued · 08/2026
Input$15
Output$75
Cache Write$18.75
Cache Read$1.5
ModelClaude Opus 4
StatusDiscontinued · 06/2026
Input$15
Output$75
Cache Write$18.75
Cache Read$1.5
ModelClaude Sonnet 4
StatusDiscontinued · 06/2026
Input$3
Output$15
Cache Write$3.75
Cache Read$0.3
ModelClaude Sonnet 3.7
StatusDiscontinued · 02/2026
Input$3
Output$15
Cache Write$3.75
Cache Read$0.3
ModelClaude 3.5 Sonnet
StatusDiscontinued · 10/2025
Input$3
Output$15
Cache Write$3.75
Cache Read$0.3
ModelClaude 3.5 Haiku
StatusDiscontinued · 02/2026
Input$0.8
Output$4
Cache Write$1
Cache Read$0.08
ModelClaude 3 Haiku
StatusDiscontinued · 04/2026
Input$0.25
Output$1.25
Cache Write$0.3
Cache Read$0.03

Important notes on the price table:

  • All prices are in USD per million tokens
  • Batch API usage provides 50% discount on all active models
  • Prompt caching can reduce costs by up to 90% for repeated content
  • Extended thinking tokens in Opus/Sonnet 4+ models are billed at the regular output rate

Fable 5 leads SWE-bench; Opus 4.7 wins GPQA Diamond

Models:
Claude Fable 5
Claude 4.8 Opus
Claude 4.7 Opus
Claude 4.6 Opus
Claude 4.6 Sonnet
Claude 4.5 Opus
Claude 4.5 Sonnet
Claude 4.1 Opus
Sources: Anthropic, Vals AI
|
CC BY 4.0
gradually.ai

It gets even more interesting when you plot coding performance against API price and compare Claude with GPT and Gemini.

DeepSeek and GPT-5.3 Codex define the efficiency frontier
Ideal: strong + cheap
Anthropic
OpenAI
Google
DeepSeek
Moonshot AI
Efficiency frontier (best price-performance)
Sources: Anthropic, OpenAI, Google DeepMind
|
CC BY 4.0
gradually.ai

Conclusion

For most users, Claude Sonnet 5 (June 2026) is the right choice. It offers world-class performance at a balanced price point, building on the strong coding foundations of Claude 4.6 Sonnet with a 1M token context window.

For high-volume, cost-sensitive applications, Claude 4.5 Haiku offers the best balance of speed, quality, and price. And for demanding coding and enterprise work, Claude Opus 5 (July 2026) is the recommended choice, close to Fable 5's frontier intelligence at half the price.

The good news: all modern Claude models are excellent. You can't really go wrong. Just pick the one that fits your budget and use case.

Frequently Asked Questions About Claude Models

FH

Finn Hillebrandt

AI Expert & Blogger

Finn Hillebrandt is the founder of Gradually AI, an SEO and AI expert. He helps online entrepreneurs simplify and automate their processes and marketing with AI. Finn shares his knowledge here on the blog in 50+ articles as well as through the AI Business Club.

Learn more about Finn and the team, follow Finn on LinkedIn, join his Facebook group for ChatGPT, OpenAI & AI Tools or do like 17,500+ others and subscribe to his AI Newsletter with tips, news and offers about AI tools and online business. Also visit his other blog, Blogmojo, which is about WordPress, blogging and SEO.