The Latent
EN▾
EnglishEspañol中文PortuguêsFrançaisالعربية日本語한국어
Sign Up
NEWSBENCHMARKSDATALEARNNEWSLETTERPARTNER WITH US →
Data/Pricing

Pricing

API Input List Price by Capability Tier and Vendor

Historical input-token list prices for separate OpenAI and Anthropic flagship and economy tiers. The chart compares vendor lines directly; it does not average the two vendors, because a median of two values is equivalent to their arithmetic mean.

API Input List Price by Capability Tier and Vendor

  • OpenAI flagship
  • Anthropic flagship
  • OpenAI economy
  • Anthropic economy
Anthropic flagship input pricing was $10 per million tokens in 2026-08-21.$30$20$10$0Jul '23Jan '24Jul '24Jan '25Jul '25Jan '26Jul '26{"f":[800,420,56,16],"s":[["OpenAI flagship","#10A37F"],["Anthropic flagship","#D97757"],["OpenAI economy","#34D399"],["Anthropic economy","#F59E0B"]],"p":[["2023-07-11T00:00:00.000Z","Jul '23",56,[[0,"$30",16,null],[1,"$11.02",246.29,null],[2,"$1.5",361.8,null],[3,"$1.63",360.22,null]],null],["2023-11-06T00:00:00.000Z","Nov '23",131.55,[[0,"$10",258.67,null],[1,"$11.02",246.29,null],[2,"$1",367.87,null],[3,"$1.63",360.22,null]],null],["2024-01-25T00:00:00.000Z","Jan '24",182.78,[[2,"$0.5",373.93,null],[3,"$1.63",360.22,null]],null],["2024-03-04T00:00:00.000Z","Mar '24",207.75,[[0,"$10",258.67,null],[1,"$15",198,null]],null],["2024-03-13T00:00:00.000Z","Mar '24",213.51,[[2,"$0.5",373.93,null],[3,"$0.25",376.97,null]],null],["2024-05-13T00:00:00.000Z","May '24",252.57,[[0,"$10",258.67,null],[1,"$15",198,null]],null],["2024-07-18T00:00:00.000Z","Jul '24",294.82,[[2,"$0.15",378.18,null],[3,"$0.25",376.97,null]],null],["2024-08-06T00:00:00.000Z","Aug '24",306.99,[[0,"$2.5",349.67,null],[1,"$15",198,null]],null],["2024-12-03T00:00:00.000Z","Dec '24",383.18,[[2,"$0.15",378.18,null],[3,"$0.8",370.29,null]],null],["2025-04-14T00:00:00.000Z","Apr '25",467.7,[[0,"$2",355.73,null],[1,"$15",198,null],[2,"$0.1",378.79,null],[3,"$0.8",370.29,null]],null],["2025-08-07T00:00:00.000Z","Aug '25",541.33,[[0,"$1.25",364.83,null],[1,"$15",198,null],[2,"$0.05",379.39,null],[3,"$0.8",370.29,null]],null],["2025-10-15T00:00:00.000Z","Oct '25",585.51,[[2,"$0.05",379.39,null],[3,"$1",367.87,null]],null],["2025-11-24T00:00:00.000Z","Nov '25",611.12,[[0,"$1.25",364.83,null],[1,"$5",319.33,null]],null],["2025-12-11T00:00:00.000Z","Dec '25",622.01,[[0,"$1.75",358.77,null],[1,"$5",319.33,null]],null],["2026-03-05T00:00:00.000Z","Mar '26",675.79,[[0,"$2.5",349.67,null],[1,"$5",319.33,null]],null],["2026-03-17T00:00:00.000Z","Mar '26",683.48,[[2,"$0.2",377.57,null],[3,"$1",367.87,null]],null],["2026-06-10T00:00:00.000Z","Jun '26",737.9,[[0,"$2.5",349.67,null],[1,"$10",258.67,null]],null],["2026-06-12T00:00:00.000Z","Jun '26",739.18,[[0,"$2.5",349.67,null],[1,"$5",319.33,null]],null],["2026-07-01T00:00:00.000Z","Jul '26",751.35,[[0,"$2.5",349.67,null],[1,"$10",258.67,null]],null],["2026-07-09T00:00:00.000Z","Jul '26",756.47,[[0,"$5",319.33,null],[1,"$10",258.67,null],[2,"$1",367.87,null],[3,"$1",367.87,null]],null],["2026-07-30T00:00:00.000Z","Jul '26",769.91,[[2,"$0.2",377.57,null],[3,"$1",367.87,null]],null],["2026-08-21T00:00:00.000Z","Aug '26",784,[[0,"$4",331.47,null],[1,"$10",258.67,null]],null]]}OpenAI flagship: $4 on Aug '26. Anthropic flagship: $10 on Aug '26. OpenAI economy: $0.2 on Jul '26. Anthropic economy: $1 on Jul '26
SOURCE: OpenAI and Anthropic official pricing and release pages; Internet Archive
RANGEALLYTD12M3M1M
Key takeaway

OpenAI and Anthropic input prices fell substantially across the period, but each vendor’s line remains separate so the chart does not hide divergent pricing moves inside a two-vendor average.

Anthropic flagship input pricing was $10 per million tokens in 2026-08-21.

Pro API coming soon

Methodology

The chart uses one canonical measure: USD list price per one million uncached text input tokens for standard synchronous API processing and the vendor's base short-context band. Input and output tokens are not blended because their prices differ substantially and a blend would require a workload-specific input/output ratio. Cached input, cache writes, batch or flex discounts, priority or fast premiums, long-context uplifts, fine-tuning, regional processing and negotiated commitments are excluded. Historical per-1,000-token quotes are multiplied by 1,000; no currency conversion is needed because every included quote is in USD.

The vendor panel is fixed to OpenAI and Anthropic from 2023-07-11, the first date on which official material provides token-denominated input prices for both a high-end and a low-cost general-purpose model from both vendors. The panel is intentionally not expanded when later vendors publish prices: adding entrants only after they become cheap would create entry and survivorship bias. At each change point, each vendor observation is published as its own line. The fixed panel is transparent, not market-share-weighted, and does not represent an industry average.

Tier assignment is prospective and based on the vendor's contemporaneous product hierarchy, not on today's model names or retrospective benchmark scores. The flagship tier is the generally available standard model the vendor describes as its flagship, most capable standard model, or top broadly available family tier. Premium 'Pro' variants, restricted safety-lifted models, and models positioned for a specialized modality are excluded. The economy tier is the vendor's current-generation model explicitly described as its smallest, fastest, most affordable, nano, mini, Instant or Haiku tier. A cheaper superseded model does not remain in the economy series after the vendor launches a current-generation replacement, which avoids selecting legacy survivors solely because their sticker price is low.

Models enter on a dated official release or price announcement and leave when a replacement tier becomes available, access is officially suspended, or an official price change takes effect. This rule is visible in the brief June 2026 Fable 5 entry, suspension and redeployment rather than smoothing the outage away. When OpenAI's 2024-01-25 post announced a GPT-3.5 Turbo price for release 'next week' without an exact effective day, the publication timestamp is used and the uncertainty is stated in the point note. Temporary prices are retained as real list-price observations: GPT-5.6 Sol's 2026-08-21 promotional price is included and marked as guaranteed only through at least 2026-11-21.

Every change point links to the official page that caused the constituent change. The initial GPT-4 value is additionally backed by the Internet Archive's 2023-03-14 capture of OpenAI's launch page, and Anthropic's July 2023 values come from its dated official model-pricing PDF. Live vendor pages can be edited after publication, so dated launch posts and exact Wayback captures are preferred where available. Archive coverage is uneven, especially for interactive pricing tables; the series therefore omits unsupported historical prices rather than interpolating them, and starts in July 2023 rather than fabricating a January 2023 baseline.

The former two-vendor median was removed. Each parsed OpenAI and Anthropic input price is emitted as its own series; no vendor is averaged, imputed, or filled when its price is unavailable.

Frequently asked questions

Why does the chart use input price rather than a blended token price?

Input and output tokens are separately metered products with very different list prices. A blended figure would depend on an assumed workload ratio and cache-hit rate. Using uncached input price keeps the unit consistent through time and lets readers apply their own output usage separately.

What does capability tier mean here?

It is a position in each vendor's contemporaneous product family, not a fixed benchmark-score threshold. Flagship means the broadly available standard model the vendor identifies as its leading general-purpose tier; economy means the current-generation model it identifies as smallest, fastest or most affordable. The mapping is made at release time and is not rewritten using later labels.

Why are capability-tier API prices relevant to the AI industry?

Token list prices help set the direct inference cost faced by developers choosing hosted models. Tracking fixed vendor tiers shows how entry prices for leading and lower-cost offerings changed, but it does not measure total task cost, model efficiency, demand, or realized enterprise pricing.

Is this the average price of the whole model market?

No. It is a fixed panel of separate OpenAI and Anthropic lines. Fixing the panel makes the history comparable and avoids adding later entrants only after their low prices are known, but it does not represent vendor market share or every API provider.

Why can a tier price rise even when inference is generally getting cheaper?

The series follows current-generation tier replacements. A new economy or flagship model can cost more per token while delivering materially greater capability, so the line mixes within-model price cuts with product-generation upgrades. Point notes identify which mechanism caused each move.

Are free tiers, batch discounts and cached-token discounts included?

No. The benchmark uses paid, standard synchronous, uncached, short-context input list prices. Free quotas, batch or flex processing, cached input, long-context premiums, regional uplifts and contract discounts are excluded because they are not directly comparable across vendors and time.

Related charts

  • Minimum API Price at ECI ≥125
  • PJM Base Residual Auction Clearing Price

The Latent

AI industry news. A sister publication to The Block.

Editorial

  • Standards
  • Corrections
  • Commercial policy
  • Contact

Company

  • LEARN
  • Data
  • Benchmarks
  • Score
  • Methodology
  • About
  • Team
  • Privacy Policy
  • Terms of Service
  • Security
  • The Block
  • Add The Latent as a preferred source