Skip to content
  • Models
  • Rankings
  • Ori
Sign Up
Sign Up
OpenRouterOpenRouter
© 2026 OpenRouter, Inc

Product

  • Chat
  • Rankings
  • Benchmarks
  • Apps
  • Discover
  • Models
  • Collections
  • Providers
  • Pricing
  • Business
  • Enterprise
  • Labs

Company

  • About
  • Blog
  • Careers
    Hiring
  • Privacy
  • Terms of Service
  • Trust Center
  • Support
  • Works With OR
  • Data
  • Brand

Developer

  • Documentation
  • API Reference
  • Developer Platform
  • Status

Connect

  • Discord
  • GitHub
  • LinkedIn
  • X
  • YouTube
Favicon for anthropic

Claude Opus 5 (batch)

anthropic/claude-opus-5:batch

Compare

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis of charts and documents, complex office deliverables, and coordinating parallel subagents.

The model maintains strong instruction following and tool use across extended tasks, while remaining effective at lower effort settings for workloads that prioritize latency and token efficiency.

Modalities

In / Out Price

$2.50 / $12.50per 1M

Context

1.0M

Released

Jul 24, 2026

Compare
ProvidersPricingPerformanceUptimeBenchmarksAppsActivityFAQExplore

Providers

Different companies host the same model. OpenRouter routes your request to one of them based on the routing mode you pick — Balanced (price + speed), Nitro (fastest), or Exacto (highest tool-calling accuracy).

Pricing

The average price customers actually pay for this model, next to the prices providers post. Caching and discounts mean the price actually paid is often well below the listed one.

Performance

Throughput is how fast the model writes (tokens per second — higher is better). Latency is total round-trip time (lower is better). TTFT is time-to-first-token — how long before you see anything appear (lower is better).

Uptime

Uptime is the percentage of the past 3 days that at least one provider was responding to requests. Availability is the percentage of time that inference was successfully served. OpenRouter continuously monitors and uses the next-best provider when one returns an error.

Benchmarks

Scores on standardized evaluations. Higher percentages are better — and rank percentile shows where this model lands among all models on OpenRouter.

Benchmark score summary for Claude Opus 5 (batch) (Artificial Analysis and Design Arena)
SourceBenchmarkScore
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) Intelligence Index49.7
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) Coding Index77.0
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) Agentic Index55.5
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) GPQA Diamond93.7%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) HLE54.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) AA-LCR80.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) GDPval-AA60.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) CritPt27.7%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) SciCode55.7%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) AA-Omniscience Accuracy59.5%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Xhigh Effort) AA-Omniscience Non-Hallucination Rate40.5%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) Intelligence Index39.8
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) Coding Index66.9
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) Agentic Index37.5
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) GPQA Diamond88.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) HLE43.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) AA-LCR81.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) GDPval-AA43.6%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) CritPt23.1%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) SciCode49.2%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) AA-Omniscience Accuracy55.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Low Effort) AA-Omniscience Non-Hallucination Rate37.8%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) Intelligence Index50.7
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) Coding Index78.0
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) Agentic Index56.2
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) GPQA Diamond93.2%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) HLE54.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) AA-LCR79.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) GDPval-AA61.8%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) CritPt29.1%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) SciCode56.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) AA-Omniscience Accuracy60.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Max Effort) AA-Omniscience Non-Hallucination Rate39.2%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) Intelligence Index45.1
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) Coding Index74.3
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) Agentic Index47.0
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) GPQA Diamond91.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) HLE51.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) AA-LCR82.0%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) GDPval-AA51.2%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) CritPt26.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) SciCode51.5%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) AA-Omniscience Accuracy57.1%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, Medium Effort) AA-Omniscience Non-Hallucination Rate39.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) Intelligence Index48.2
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) Coding Index76.5
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) Agentic Index52.7
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) GPQA Diamond93.7%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) HLE52.8%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) AA-LCR79.0%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) GDPval-AA56.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) CritPt28.3%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) SciCode55.4%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) AA-Omniscience Accuracy58.9%
Artificial AnalysisClaude Opus 5 (Adaptive Reasoning, High Effort) AA-Omniscience Non-Hallucination Rate38.8%
Design ArenaClaude Opus 5 Agents Arena Agenticgamedev Elo1261
Design ArenaClaude Opus 5 Agents Arena Androidnative Elo1266
Design ArenaClaude Opus 5 Agents Arena Full Stack Elo1324
Design ArenaClaude Opus 5 Agents Arena Mobile Apps Elo1348
Design ArenaClaude Opus 5 Agents Arena Python-Pptxslides Elo1265
Design ArenaClaude Opus 5 Agents Arena Webapps Elo1278
Design ArenaClaude Opus 5 Models Arena 3D Elo1363
Design ArenaClaude Opus 5 Models Arena Asciiart Elo1389
Design ArenaClaude Opus 5 Models Arena Code Categories Elo1338
Design ArenaClaude Opus 5 Models Arena Data Visualization Elo1356
Design ArenaClaude Opus 5 Models Arena Game Development Elo1363
Design ArenaClaude Opus 5 Models Arena SVG Elo1350
Design ArenaClaude Opus 5 Models Arena UI Component Elo1361
Design ArenaClaude Opus 5 Models Arena Website Elo1319

Apps

Public apps that send the most traffic to this model. Good signal for what real production workloads look like — and a hint at which use cases this model is best suited for.

Activity

Token volume and request traffic to this model over time.

Quick Start

Drop-in code to call this model. OpenRouter's API is OpenAI-compatible — most SDKs work by just swapping the base URL. The only thing that changes between models is the model slug below.

Explore more models

AI Models with Vision: Multimodal LLMs for Image UnderstandingCollectionAI Model RankingsRanking
$2.50$12.50$0.25----
--
AutoExacto Benchmarks
GPQA DiamondTAU-BenchGoogle Vertex (Europe)89.0%79.7%Claude Platform on AWS89.1%78.7%Azure89.5%78.1%Azure (US)89.8%77.0%Amazon Bedrock87.7%79.0%Anthropic89.6%76.4%Google Vertex (US)88.3%77.6%auto-routing86.8%79.0%Google Vertex89.3%76.4%Amazon Bedrock88.0%76.4%Amazon Bedrock88.8%75.3%
+1 more providers

Not enough uptime data to display yet.

When an error occurs in an upstream provider, we can recover by routing to another healthy provider, if your request filters allow it. You can access per-provider uptime data programmatically through the Endpoints API. Learn more about our load balancing and customization options.

Not enough data for apps using this model yet.

Frequently asked questions

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis of charts and documents, complex office deliverables, and coordinating parallel subagents.

Claude Opus 5 (batch) costs $2.50/M input tokens and $12.50/M output tokens, with separate rates for Cache Read at $0.25/M tokens, Cache Write at $3.125/M tokens, Cache Write (1h) at $5.00/M tokens and Web Search at $10.00/1K calls.

Claude Opus 5 (batch) has a 1,000,000 token context window. It supports up to 128,000 completion tokens.

Yes. Claude Opus 5 (batch) accepts tools and tool_choice for function calling. It also supports structured outputs via a JSON schema in response_format.

Claude Opus 5 (batch) accepts text, images and files such as PDFs as input and returns text.

Claude Fable 5.1, Claude Sonnet 5, Claude Fable 5 and 11 more are other text models from Anthropic.

Claude Opus 5 (batch) was released on July 24, 2026.

More models from Anthropic

Claude Fable 5.1

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual code generation, and finance and analysis tasks in particular. It also tends to be more concise than Fable 5 in its plans and summaries. We recommend testing it as a direct upgrade wherever you use Fable 5 today, and alongside Opus 5 on reasoning-heavy tasks.

Text1M context$10 / $50
Claude Fable 5.1

Claude Fable 5.1 improves on Claude Fable 5 across the board, with the biggest gains in agentic coding, long-running agentic workflows, and knowledge work: long code refactors, front-end and visual code generation, and finance and analysis tasks in particular. It also tends to be more concise than Fable 5 in its plans and summaries. We recommend testing it as a direct upgrade wherever you use Fable 5 today, and alongside Opus 5 on reasoning-heavy tasks.

Text1M context$5 / $25
Claude Opus 5

Fast-mode variant of Opus 5 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.

Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Note: As of September 1, 2026, this dedicated fast model is deprecated. Fast mode is now served by the fast service tier endpoint on Claude Opus 5 — request it with service_tier: "fast" or speed: "fast". Requests to this model keep working and are served by the same fast tier capacity, so no action is required, but new integrations should target the regular model. Learn more in our service tier docs: https://openrouter.ai/docs/guides/features/service-tiers

Text1M context
Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis of charts and documents, complex office deliverables, and coordinating parallel subagents.

The model maintains strong instruction following and tool use across extended tasks, while remaining effective at lower effort settings for workloads that prioritize latency and token efficiency.

Text1M context$5 / $25
Claude Opus 5

Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis of charts and documents, complex office deliverables, and coordinating parallel subagents.

The model maintains strong instruction following and tool use across extended tasks, while remaining effective at lower effort settings for workloads that prioritize latency and token efficiency.

Text1M context$2.50 / $12.50
Claude Sonnet 5

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max, and x-high), a 1M-token context window, and text, image, and file inputs. Sonnet 5 uses an updated tokenizer and includes real-time cyber safeguards that block certain high-risk dual-use activities.

Text1M context$2 / $10
Claude Sonnet 5

Sonnet 5 is Anthropic's most capable Sonnet-class model, with frontier performance across coding, agents, and professional work. It supports adaptive thinking with selectable reasoning effort levels (low, medium, high, max, and x-high), a 1M-token context window, and text, image, and file inputs. Sonnet 5 uses an updated tokenizer and includes real-time cyber safeguards that block certain high-risk dual-use activities.

Text1M context$1 / $5
Claude Fable Latest

This model always redirects to the latest model in the Claude Fable family.

Text1M context
Claude Fable 5

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token context window. It is suited for long-running, complex, and asynchronous tasks that previously required frequent human check-ins.

It is particularly strong at end-to-end work that would otherwise take a person hours, days, or weeks - taking on problems that are long-running, ambiguous, or highly multi-step. It executes well-scoped tasks with few mistakes, automatically self-correcting through verification loops, and ships with robust safeguards.

Text1M context$10 / $50
Claude Fable 5

Claude Fable 5 is a Mythos-class model from Anthropic, built for autonomous knowledge work and coding. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token context window. It is suited for long-running, complex, and asynchronous tasks that previously required frequent human check-ins.

It is particularly strong at end-to-end work that would otherwise take a person hours, days, or weeks - taking on problems that are long-running, ambiguous, or highly multi-step. It executes well-scoped tasks with few mistakes, automatically self-correcting through verification loops, and ships with robust safeguards.

Text1M context$5 / $25
Claude Opus 4.8

Fast-mode variant of Opus 4.8 - identical capabilities with higher output speed at 2x pricing relative to regular Opus 4.8.

Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Note: As of September 1, 2026, this dedicated fast model is deprecated. Fast mode is now served by the fast service tier endpoint on Claude Opus 4.8 — request it with service_tier: "fast" or speed: "fast". Requests to this model keep working and are served by the same fast tier capacity, so no action is required, but new integrations should target the regular model. Learn more in our service tier docs: https://openrouter.ai/docs/guides/features/service-tiers

Text1M context
Claude Opus 4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token context window. It is suited for highly autonomous agents, long-horizon agentic work, knowledge work, and memory-driven tasks where coherence over extended sessions matters.

It is particularly strong on multi-step reasoning, complex coding, and end-to-end project orchestration - large codebases, multi-stage debugging, and long-running asynchronous agent pipelines. Beyond coding, it handles knowledge work such as drafting documents, building presentations, and analyzing data, maintaining quality across very long outputs.

Text1M context$5 / $25
Claude Opus 4.8

Claude Opus 4.8 is Anthropic's most capable generally available model in the Opus family. It supports text, image, and file inputs with text output, with reasoning support and a 1M-token context window. It is suited for highly autonomous agents, long-horizon agentic work, knowledge work, and memory-driven tasks where coherence over extended sessions matters.

It is particularly strong on multi-step reasoning, complex coding, and end-to-end project orchestration - large codebases, multi-stage debugging, and long-running asynchronous agent pipelines. Beyond coding, it handles knowledge work such as drafting documents, building presentations, and analyzing data, maintaining quality across very long outputs.

Text1M context$2.50 / $12.50
Claude Opus 4.7

Fast-mode variant of Opus 4.7 - identical capabilities with higher output speed at premium 6x pricing.

Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Text1M context
Claude Haiku Latest

This model always redirects to the latest model in the Claude Haiku family.

Text200K context
Claude Sonnet Latest

This model always redirects to the latest model in the Claude Sonnet family.

Text1M context
Claude Opus Latest

This model always redirects to the latest model in the Claude Opus family.

Text1M context
Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on complex, multi-step tasks and more reliable agentic execution across extended workflows. It is especially effective for asynchronous agent pipelines where tasks unfold over time - large codebases, multi-stage debugging, and end-to-end project orchestration.

Beyond coding, Opus 4.7 brings improved knowledge work capabilities - from drafting documents and building presentations to analyzing data. It maintains coherence across very long outputs and extended sessions, making it a strong default for tasks that require persistence, judgment, and follow-through.

For users upgrading from earlier Opus versions, see our official migration guide here

Text1M context$5 / $25
Claude Opus 4.7

Opus 4.7 is the next generation of Anthropic's Opus family, built for long-running, asynchronous agents. Building on the coding and agentic strengths of Opus 4.6, it delivers stronger performance on complex, multi-step tasks and more reliable agentic execution across extended workflows. It is especially effective for asynchronous agent pipelines where tasks unfold over time - large codebases, multi-stage debugging, and end-to-end project orchestration.

Beyond coding, Opus 4.7 brings improved knowledge work capabilities - from drafting documents and building presentations to analyzing data. It maintains coherence across very long outputs and extended sessions, making it a strong default for tasks that require persistence, judgment, and follow-through.

For users upgrading from earlier Opus versions, see our official migration guide here

Text1M context$2.50 / $12.50
Claude Opus 4.6

Fast-mode variant of Opus 4.6 - identical capabilities with higher output speed at premium 6x pricing.

Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

Text1M context
Claude Sonnet 4.6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with memory, polished document creation, and confident computer use for web QA and workflow automation.

Text1M context$3 / $15
Claude Sonnet 4.6

Sonnet 4.6 is Anthropic's most capable Sonnet-class model yet, with frontier performance across coding, agents, and professional work. It excels at iterative development, complex codebase navigation, end-to-end project management with memory, polished document creation, and confident computer use for web QA and workflow automation.

Text1M context$1.50 / $7.50
Claude Opus 4.6

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective for large codebases, complex refactors, and multi-step debugging that unfolds over time. The model shows deeper contextual understanding, stronger problem decomposition, and greater reliability on hard engineering tasks than prior generations.

Beyond coding, Opus 4.6 excels at sustained knowledge work. It produces near-production-ready documents, plans, and analyses in a single pass, and maintains coherence across very long outputs and extended sessions. This makes it a strong default for tasks that require persistence, judgment, and follow-through, such as technical design, migration planning, and end-to-end project execution.

For users upgrading from earlier Opus versions, see our official migration guide here

Text1M context$5 / $25
Claude Opus 4.6

Opus 4.6 is Anthropic’s strongest model for coding and long-running professional tasks. It is built for agents that operate across entire workflows rather than single prompts, making it especially effective for large codebases, complex refactors, and multi-step debugging that unfolds over time. The model shows deeper contextual understanding, stronger problem decomposition, and greater reliability on hard engineering tasks than prior generations.

Beyond coding, Opus 4.6 excels at sustained knowledge work. It produces near-production-ready documents, plans, and analyses in a single pass, and maintains coherence across very long outputs and extended sessions. This makes it a strong default for tasks that require persistence, judgment, and follow-through, such as technical design, migration planning, and end-to-end project execution.

For users upgrading from earlier Opus versions, see our official migration guide here

Text1M context$2.50 / $12.50
Claude Opus 4.5

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and reasoning benchmarks, and improved robustness to prompt injection. The model is designed to operate efficiently across varied effort levels, enabling developers to trade off speed, depth, and token usage depending on task requirements. It comes with a new parameter to control token efficiency, which can be accessed using the OpenRouter Verbosity parameter with low, medium, or high.

Opus 4.5 supports advanced tool use, extended context management, and coordinated multi-agent setups, making it well-suited for autonomous research, debugging, multi-step planning, and spreadsheet/browser manipulation. It delivers substantial gains in structured reasoning, execution reliability, and alignment compared to prior Opus generations, while reducing token overhead and improving performance on long-running tasks.

Text200K context$5 / $25
Claude Opus 4.5

Claude Opus 4.5 is Anthropic’s frontier reasoning model optimized for complex software engineering, agentic workflows, and long-horizon computer use. It offers strong multimodal capabilities, competitive performance across real-world coding and reasoning benchmarks, and improved robustness to prompt injection. The model is designed to operate efficiently across varied effort levels, enabling developers to trade off speed, depth, and token usage depending on task requirements. It comes with a new parameter to control token efficiency, which can be accessed using the OpenRouter Verbosity parameter with low, medium, or high.

Opus 4.5 supports advanced tool use, extended context management, and coordinated multi-agent setups, making it well-suited for autonomous research, debugging, multi-step planning, and spreadsheet/browser manipulation. It delivers substantial gains in structured reasoning, execution reliability, and alignment compared to prior Opus generations, while reducing token overhead and improving performance on long-running tasks.

Text200K context$2.50 / $12.50
Claude Haiku 4.5

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications.

It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

Text200K context$1 / $5
Claude Haiku 4.5

Claude Haiku 4.5 is Anthropic’s fastest and most efficient model, delivering near-frontier intelligence at a fraction of the cost and latency of larger Claude models. Matching Claude Sonnet 4’s performance across reasoning, coding, and computer-use tasks, Haiku 4.5 brings frontier-level capability to real-time and high-volume applications.

It introduces extended thinking to the Haiku line; enabling controllable reasoning depth, summarized or interleaved thought output, and tool-assisted workflows with full support for coding, bash, web search, and computer-use tools. Scoring >73% on SWE-bench Verified, Haiku 4.5 ranks among the world’s best coding models while maintaining exceptional responsiveness for sub-agents, parallelized execution, and scaled deployment.

Text200K context$0.50 / $2.50
Claude Sonnet 4.5

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with improvements across system design, code security, and specification adherence. The model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking.

Sonnet 4.5 also introduces stronger agentic capabilities, including improved tool orchestration, speculative parallel execution, and more efficient context and memory management. With enhanced context tracking and awareness of token usage across tool calls, it is particularly well-suited for multi-context and long-running workflows. Use cases span software engineering, cybersecurity, financial analysis, research agents, and other domains requiring sustained reasoning and tool use.

Text1M context$3 / $15
Claude Sonnet 4.5

Claude Sonnet 4.5 is Anthropic’s most advanced Sonnet model to date, optimized for real-world agents and coding workflows. It delivers state-of-the-art performance on coding benchmarks such as SWE-bench Verified, with improvements across system design, code security, and specification adherence. The model is designed for extended autonomous operation, maintaining task continuity across sessions and providing fact-based progress tracking.

Sonnet 4.5 also introduces stronger agentic capabilities, including improved tool orchestration, speculative parallel execution, and more efficient context and memory management. With enhanced context tracking and awareness of token usage across tool calls, it is particularly well-suited for multi-context and long-running workflows. Use cases span software engineering, cybersecurity, financial analysis, research agents, and other domains requiring sustained reasoning and tool use.

Text1M context$1.50 / $7.50
Claude Opus 4.1

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains in multi-file code refactoring, debugging precision, and detail-oriented reasoning. The model supports extended thinking up to 64K tokens and is optimized for tasks involving research, data analysis, and tool-assisted reasoning.

Text200K context$15 / $75
Claude Opus 4.1

Claude Opus 4.1 is an updated version of Anthropic’s flagship model, offering improved performance in coding, reasoning, and agentic tasks. It achieves 74.5% on SWE-bench Verified and shows notable gains in multi-file code refactoring, debugging precision, and detail-oriented reasoning. The model supports extended thinking up to 64K tokens and is optimized for tasks involving research, data analysis, and tool-assisted reasoning.

Text200K context$7.50 / $37.50
Claude Opus 4

Claude Opus 4 is benchmarked as the world’s best coding model, at time of release, bringing sustained performance on complex, long-running tasks and agent workflows. It sets new benchmarks in software engineering, achieving leading results on SWE-bench (72.5%) and Terminal-bench (43.2%). Opus 4 supports extended, agentic workflows, handling thousands of task steps continuously for hours without degradation.

Read more at the blog post here

Text200K context$15 / $75
Claude Sonnet 4

Claude Sonnet 4 significantly enhances the capabilities of its predecessor, Sonnet 3.7, excelling in both coding and reasoning tasks with improved precision and controllability. Achieving state-of-the-art performance on SWE-bench (72.7%), Sonnet 4 balances capability and computational efficiency, making it suitable for a broad range of applications from routine coding tasks to complex software development projects. Key enhancements include improved autonomous codebase navigation, reduced error rates in agent-driven workflows, and increased reliability in following intricate instructions. Sonnet 4 is optimized for practical everyday use, providing advanced reasoning capabilities while maintaining efficiency and responsiveness in diverse internal and external scenarios.

Read more at the blog post here

Text1M context$3 / $15
Claude 3.7 Sonnet

Claude 3.7 Sonnet is an advanced large language model with improved reasoning, coding, and problem-solving capabilities. It introduces a hybrid reasoning approach, allowing users to choose between rapid responses and extended, step-by-step processing for complex tasks. The model demonstrates notable improvements in coding, particularly in front-end development and full-stack updates, and excels in agentic workflows, where it can autonomously navigate multi-step processes.

Claude 3.7 Sonnet maintains performance parity with its predecessor in standard mode while offering an extended reasoning mode for enhanced accuracy in math, coding, and instruction-following tasks.

Read more at the blog post here

Text200K context
Claude 3.5 Haiku

Claude 3.5 Haiku features offers enhanced capabilities in speed, coding accuracy, and tool use. Engineered to excel in real-time applications, it delivers quick response times that are essential for dynamic tasks such as chat interactions and immediate coding suggestions.

This makes it highly suitable for environments that demand both speed and precision, such as software development, customer service bots, and data management systems.

This model is currently pointing to Claude 3.5 Haiku (2024-10-22).

Text200K context
Claude 3.5 Haiku

Claude 3.5 Haiku features enhancements across all skill sets including coding, tool use, and reasoning. As the fastest model in the Anthropic lineup, it offers rapid response times suitable for applications that require high interactivity and low latency, such as user-facing chatbots and on-the-fly code completions. It also excels in specialized tasks like data extraction and real-time content moderation, making it a versatile tool for a broad range of industries.

It does not support image inputs.

See the launch announcement and benchmark results here

Text200K context
Claude 3.5 Sonnet

New Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at:

#multimodal

Text200K context
Claude 3.5 Sonnet

Claude 3.5 Sonnet delivers better-than-Opus capabilities, faster-than-Sonnet speeds, at the same Sonnet prices. Sonnet is particularly good at:

For the latest version (2024-10-23), check out Claude 3.5 Sonnet.

#multimodal

Text200K context
Claude 3 Haiku

Claude 3 Haiku is Anthropic's fastest and most compact model for near-instant responsiveness. Quick and accurate targeted performance.

See the launch announcement and benchmark results here

#multimodal

Text200K context$0.25 / $1.25
Claude 3 Opus

Claude 3 Opus is Anthropic's most powerful model for highly complex tasks. It boasts top-level performance, intelligence, fluency, and understanding.

See the launch announcement and benchmark results here

#multimodal

Text200K context
Claude 3 Sonnet

Claude 3 Sonnet is an ideal balance of intelligence and speed for enterprise workloads. Maximum utility at a lower price, dependable, balanced for scaled deployments.

See the launch announcement and benchmark results here

#multimodal

Text200K context
Claude v2.1

Claude 2 delivers advancements in key capabilities for enterprises—including an industry-leading 200K token context window, significant reductions in rates of model hallucination, system prompts and a new beta feature: tool use.

Text200K context
Claude v2

Claude 2 delivers advancements in key capabilities for enterprises—including an industry-leading 200K token context window, significant reductions in rates of model hallucination, system prompts and a new beta feature: tool use.

Text200K context
Claude Instant v1.1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text100K context
Claude Instant v1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text100K context
Claude v2.0

Anthropic's flagship model. Superior performance on tasks that require complex reasoning. Supports hundreds of pages of text.

Text100K context
Claude v1

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text100K context
Claude Instant v1.0

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text100K context
Claude v1.2

Anthropic's model for low-latency, high throughput text generation. Supports hundreds of pages of text.

Text100K context