OpenAI

GPT-5.2

400K ctxTextImage Tools Reasoning

Released Jan 15, 2026 · 641B tokens this week · 2 providers

Overview

OpenAI's flagship general model, tuned for long multi-step work and tool-heavy agents. It holds instructions across very long sessions and can switch between quick answers and extended deliberation. The default pick when you want maximum reliability rather than the lowest price.

reasoningprogrammingagentsacademia

Providers

Reference data — not yet measured from live traffic

PROVIDERMAX OUTOUTPUT /MQUANT
OpenAICheapest
400K128K$1.75$14.00208ms9299.94%Full
Azure AI Foundry
400K128K$1.89$15.12780ms7499.81%Full

Latency p50 reflects a provider’s API endpoint responsiveness — the round-trip to its API, not per-token inference time. These figures, with throughput and uptime, are reference data until the gateway aggregates its own traffic.

Price comparison

Blended $ per 1M tokens

Blended rate per 1M tokens, weighted one part prompt to three parts completion — roughly the shape of a chat workload. Your mix will move the number.

Call it

OpenAI-compatible — swap the base URL and go

curl https://model.cards/api/v1/chat/completions \
  -H "Authorization: Bearer $MODELCARDS_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai/gpt-5.2",
    "messages": [
      { "role": "user", "content": "Summarize the tradeoffs of speculative decoding." }
    ]
  }'