See how ChatGPT, Perplexity and Google AI Overviews describe you today

Free AI Visibility Audit
Anthropic model°

Claude Opus 4.1

Claude Opus 4.1 is Anthropic's flagship Claude model, an updated Opus release built for coding, reasoning and agentic work, and it accepts text, images and files.

AnthropicVerified 24/09/2026Released 05/08/2025

01 / decision

Decision snapshot

Each figure sits next to the middle of the field, so it can be read as dear or cheap, wide or narrow, rather than floating on its own. Compared against the 126 models we track, not against an absolute standard.

Capability

Not measured

No published benchmark scores yet

Input / 1M tokens

$15.00

field median $0.30

Output / 1M tokens

$75.00

field median $1.25

Context

200K

field median 524K

02 / overview

What Claude Opus 4.1 is, and when to reach for it

Four questions, answered separately, because somebody arrives at one of them rather than at the top of the page.

What it is

The job it was built for, and the job it is not for.

Claude Opus 4.1 is a frontier general-purpose model from Anthropic aimed at the hard end of the workload: multi-step coding, extended reasoning and agent runs that call tools and act over several turns. It reads text, images and files, and Anthropic reports 74.5% on SWE-bench Verified for it. It is not the model you reach for when the job is high-volume classification or routine summarising, the cost per call does not earn its keep there.

When it arrived, and when to use it

Where it sits in its line, and when a sibling is the better pick.

Claude Opus 4.1 first appeared on 5 August 2025 as an incremental update to Anthropic's top Opus tier rather than a new generation. Pick it when a task has been failing on smaller Claude models, typically long refactors, debugging across a repository, or agent loops where one wrong step compounds. For chat, drafting and anything you run thousands of times a day, a cheaper sibling is the right call.

How you reach it

The API, the apps it powers, and what its limits let you do.

Claude Opus 4.1 is reached through the Anthropic API and powers the Claude apps, and it is available through the major cloud model platforms that carry Claude. Its context window holds roughly a large codebase slice or a stack of long documents in a single call, so you can hand it whole files rather than chunking them, and its image and file input means screenshots, PDFs and diagrams go in directly. Output length is generous enough for full files or long structured reports in one response.

Why it matters

What changes because this exists, or why it does not.

Claude Opus 4.1 matters mainly for agentic and coding pipelines, where the improvement over the previous Opus is the difference between an agent finishing a task and stalling halfway. Teams that had been splitting work into smaller supervised steps can hand over longer chains. As a release it is an increment, not a reset, the interesting part is reliability on long runs rather than any new capability.

Superseded by Claude Opus 4.5. See the whole line.

03 / evidence

How much of this is verified

Split by category, so a strong number never hides a thin evidence base. Verified means we read it on the benchmark's own published results; a provider's claim about its own model is shown and labelled rather than dropped.

No published benchmark scores for this model yet.

We publish a score only where we can link the result to where it was published. Until a benchmark result for this model exists in a source we read, this section stays empty rather than being filled with a provider's marketing figure.

How we decide what counts as evidence

04 / ledger

Benchmark ledger

Every published row, grouped by category, each compared with the best published score on the same benchmark. 'Is 64% good' is a question nobody can answer; '26 points behind the leader' is one anybody can.

Nothing in the ledger yet.

Each row here carries a score, the benchmark it came from, what the leading model scored on the same test, and a link to the published result. Rows appear as results are published and read.

How we decide what counts as evidence

05 / capability

Capability shape

Where this model is strong, and against how many peers. Ranks are against models with evidence in that category, not against everything we track: ranking against models nobody tested would rank who published, not who is better.

No category scores to shape yet.

A category score is the weighted mean of the benchmarks published for it. With no published rows there is nothing to average, and an empty chart drawn at zero would say something false.

How we decide what counts as evidence

06 / cost

What it costs

List API rates as last read from the provider, with the source on every row, plus every change we have recorded since we started tracking it.

Claude Opus 4.1 API pricingSurge45°
ChargePriceUnitRead onSource
Input$15.00per 1M tokens2026-09-24Check
Output$75.00per 1M tokens2026-09-24Check

Claude Opus 4.1 sits at the top of Anthropic's price list and among the most expensive models generally available, with output charged several times higher than input, as is standard. Treat it as a model you route to for the hard fraction of requests, with a cheaper model handling the rest, because running everything through it gets costly quickly.

What a month costsSurge45°
WorkloadInput / monthOutput / monthCost
A small product team20M tokens5M tokens$675.00
A busy support assistant200M tokens40M tokens$6,000.00
A document pipeline1000M tokens100M tokens$22,500.00

List API rates, no caching and no batch discount, which both providers offer and which change the answer a great deal. Treat these as the ceiling, not the bill.

07 / specs

Specifications

As listed by the provider's own catalogue and re-read every few hours. Anything absent is absent there too.

SpecificationSurge45°
Context window200,000 tokens
Maximum output32,000 tokens
Modalitiesimage, text, file
Released05/08/2025
StatusCurrent
Catalogue identifieranthropic/claude-opus-4.1

08 / lineage

Lineage

What this model replaced, what replaced it, and what else its provider has in the field.

Also from Anthropic

09 / line

The line

Every model in this family in release order, so a page from eight months ago says in one glance that two newer ones exist.

Came before

The first we track in this line.

  1. 01Claude Opus 4.105/08/2025
  2. 02Claude Opus 4.524/11/2025
  3. 03Claude Opus 524/07/2026
  4. 04Claude Opus 5.522/09/2026

Ordered by release date and worked out from the naming, so a new member slots in as soon as its page exists. A retired model keeps its page and its place in the line.

10 / notes

Our notes

What this model changes for a brand trying to be cited in AI answers, and every change we have logged since it launched.

What it changes for you

When a buyer researches software in Claude, the answer they read is assembled by a model of this class, working from long documents it can take in whole. That rewards source material that holds up at length, clear product pages, comparison content and documentation that state facts plainly rather than pages built around a single keyword. If your category page cannot be read end to end and summarised accurately, you will be described by whoever wrote something that can.

Where buyers meet this model

Buyers meet Claude Opus 4.1 in the Claude consumer and team apps, where it is the top-tier option for paid users asking substantial questions. Developers meet it through the Anthropic API and the cloud platforms that resell Claude. It also sits behind coding assistants and agent products that have selected Anthropic's flagship tier, so a buyer may be reading its output without ever seeing the model name.

Change log

Nothing published here yet. Changes appear within hours of a provider announcing them.

11 / questions

Questions

The things people ask about this model, answered from what is on this page rather than from anywhere else.

What is Claude Opus 4.1 best at?
Coding, reasoning and agentic tasks. Anthropic positions it as its flagship and reports 74.5% on SWE-bench Verified, so it is the Claude model to use for long refactors, debugging and agent runs that need to stay on track across many steps.
Can Claude Opus 4.1 read images and files?
Yes. It takes text, images and files as input, so screenshots, PDFs and documents can be passed in directly rather than converted to text first.
When should I use a cheaper Claude model instead?
For chat, drafting, summarising and anything you run at high volume. Opus 4.1 is priced at the top of the range, so most teams route only the hard requests to it and let a smaller model handle the rest.
Where can I access Claude Opus 4.1?
Through the Claude apps for end users and the Anthropic API for developers, as well as the major cloud platforms that carry Anthropic's models.
Surge45°

Is Claude Opus 4.1 recommending you?

Models change what gets cited. We measure whether AI answers name your brand or your competitors across every assistant, and show you what to change.