Anthropic

Claude Opus 4.1

200K ctxTextImage Tools Reasoning

Released Aug 5, 2025 · 41B tokens this week · 2 providers

Overview

The previous Opus generation, still excellent at deep analysis and large refactors. Priced at the premium tier and mostly kept around for teams that validated their evaluations against it. Superseded by Opus 4.5 on both quality and cost.

reasoningprogramminglegal

Providers

Reference data — not yet measured from live traffic

PROVIDERMAX OUTOUTPUT /MQUANT
AnthropicCheapest
200K32K$15.00$75.00239ms4499.90%Full
Amazon Bedrock
200K32K$15.75$78.751,340ms3899.68%Full

Latency p50 reflects a provider’s API endpoint responsiveness — the round-trip to its API, not per-token inference time. These figures, with throughput and uptime, are reference data until the gateway aggregates its own traffic.

Price comparison

Blended $ per 1M tokens

GPT-5.2$10.94
Claude Opus 4.1$60.00

Blended rate per 1M tokens, weighted one part prompt to three parts completion — roughly the shape of a chat workload. Your mix will move the number.

Call it

OpenAI-compatible — swap the base URL and go

curl https://model.cards/api/v1/chat/completions \
  -H "Authorization: Bearer $MODELCARDS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "anthropic/claude-opus-4.1",
    "messages": [
      { "role": "user", "content": "Summarize the tradeoffs of speculative decoding." }
    ]
  }'