AI Model Pricing Watch 4. October 2026.

October starts with more choice. There are newer models with lower token prices than some of the models they sit beside. But that does not mean every model got cheaper.

That is the useful distinction this month: a cheaper route is not the same thing as a price cut on the route you already use.

I checked direct API prices across the same eight providers as September, added relevant new models, and kept the same three token baskets. This is a snapshot checked on 1 October, not an estimate of your whole October bill.

The short version

  • A cheaper small-model option: GPT-6 Luna's quick basket costs about £1.13 for 1,000 identical calls. That is token spend, not the cost of 1,000 finished agent tasks.
  • Newer everyday options: GPT-6.1 Sol and Claude Sonnet 5.5 both list $2 input and $10 output per million tokens. They are not interchangeable just because the rate is the same.
  • A lower-priced Opus option: Opus 5.5's input and output rates are 20% below Opus 5's. Opus 5 itself has not been repriced in this comparison.
  • A small currency headwind: the pound cost of an unchanged dollar basket is about 1.73% higher than in September's snapshot.

The full watch covers 32 entries across eight providers. Of the 25 directly matched September entries, zero uncached reference baskets changed in dollars; 25 were unchanged. Their recorded input, cache-read and output rates also match. The seven new entries have no invented September history.

What changed since September?

OpenAI: a new Luna, then two Sol releases

OpenAI's release log dates GPT-6 Luna and GPT-6 Sol to 22 September, followed by GPT-6.1 Sol on 29 September.

GPT-6 Luna lists $0.10 input and $0.50 output per million tokens. Its quick basket is $0.0015 instead of GPT-5.6 Luna's $0.0032: a 53.1% lower token cost. October's pound equivalent is £0.001129. This is a model substitution, not a same-model September price cut.

GPT-6.1 Sol keeps GPT-6 Sol's $2 input and $10 output rates, but halves its cache-read rate from $0.20 to $0.10 per million tokens. That difference does not help our uncached basket. It may help a real workflow with reusable context, once cache-write charges are included.

Anthropic: lower-priced Opus, same-priced Sonnet

Opus 5.5 arrived on 22 September. Its $4 input and $20 output rates make the fixed frontier basket 20% cheaper in dollars than Opus 5's $5/$25 tariff. At October's exchange rate, that is about £1.13 rather than £1.41.

Sonnet 5.5 arrived on 28 September at Sonnet 5's existing $2/$10 rates. Anthropic also reports token-efficiency improvements. That is a provider claim about task behaviour, not evidence from my own cross-provider test and not a lower token tariff.

DeepSeek: the Flash name now points to a different model

DeepSeek's 10 September update replaces V4 Flash with V4.1 Flash. The current endpoint is deepseek-flash; old V4 Flash names temporarily route to the replacement.

The new model's peak prices are $0.30 input and $1.20 output per million tokens, with off-peak rates at half. I have added it as a new entry. An ordinary price-trend line to the retired model would hide a change of model.

Pro remains available with version 0813 and its peak tariff retained. Peak prices apply only during the published UTC weekday windows, excluding Chinese public holidays.

Grok: a new model without a new standard token tariff

Grok 4.7 launched on 21 September. Its standard $2/$6 input/output rates match Grok 4.6's. Both use $4/$12 for our 250,000-input-token basket. The faster harness offering is outside this direct public API comparison.

Google, Mistral, Kimi and Qwen remain in the watch. Their retained rows keep the recorded dollar tariffs, with the qualifications below. This is a selected text/coding catalogue, not a claim that none of these companies released anything else.

What would the token bill look like?

A token is a unit the model processes, not a word or a finished job. Input is what you send; output is what the model generates. Billable reasoning can add to output. These examples contain no cache hits, tools, retries or taxes.

Illustrative token-only costs, at October's recorded rates
Model and basketOne call GBP1,000 identical calls GBP
GPT-6 Luna: quick0.0011291.13
GPT-6.1 Sol: general0.112909112.91
Claude Opus 5.5: frontier1.1290931129.09

Scale the basket, not the promise. A workflow that makes ten calls, searches the web and retries twice will not cost the same as one clean call. Different models may need different numbers of tokens to pass the same checks.

The same three baskets

Fixed workloads used in every monthly edition
BasketInput tokensOutput tokens
Quick10,0001,000
General50,0005,000
Frontier250,00025,000

The labels select a basket size. They do not certify equivalent capability. Calculate input tokens multiplied by the applicable input rate, plus output tokens multiplied by the applicable output rate, divided by one million.

The Bank of England's latest dated observation returned was 30 September: $1.3285 for £1. Its reciprocal is about £0.752729 per dollar, compared with September's stored £0.7399. This is an indicative spot conversion, not an executable bank quote.

October's full price table

Input, cache read and output are USD per million tokens. The final column is the uncached basket cost in GBP. A dash means a model-specific cache price was not recorded, not that caching is free. Model names link to official sources.

Quick basket

Quick: 10,000 input and 1,000 output tokens; standard direct API unless noted
ProviderModelInput USD/1MCache read USD/1MOutput USD/1MBasket tierBasket GBP
Alibaba CloudQwen3.8 Flash0.113—0.382standard0.001138
AnthropicClaude Haiku 4.510.105standard0.011291
DeepSeekDeepSeek V4.1 Flash (peak)0.300.0061.20peak0.003161
GoogleGemini 3.5 Flash-Lite0.300.032.50standard0.004140
MistralMistral Small 40.15—0.60standard0.001581
OpenAIGPT-5.6 Luna0.200.021.20standard0.002409
OpenAIGPT-6 Luna0.100.010.50standard0.001129

General basket

General: 50,000 input and 5,000 output tokens; standard direct API unless noted
ProviderModelInput USD/1MCache read USD/1MOutput USD/1MBasket tierBasket GBP
Alibaba CloudQwen3.7 Plus0.32—1.28standard0.016861
AnthropicClaude Sonnet 520.2010standard0.112909
AnthropicClaude Sonnet 5.520.2010standard0.112909
GoogleGemini 3.7 Flash0.750.0753.75standard0.042341
GoogleGemini 3.8 Flash0.750.0753.75standard0.042341
MistralMistral Large 30.50—1.50standard0.024464
MistralMistral Medium 3.51.50—7.50standard0.084682
Moonshot AIKimi K2.60.950.164standard0.050809
Moonshot AIKimi K2.7 Code0.950.194standard0.050809
OpenAIGPT-5.6 Terra20.2012standard0.120437
OpenAIGPT-6 Sol20.2010standard0.112909
OpenAIGPT-6.1 Sol20.1010standard0.112909
xAIGrok 4.31.250.202.50standard0.056455
xAIGrok Build 0.110.202standard0.045164

Frontier basket

Frontier: 250,000 input and 25,000 output tokens; standard direct API unless noted
ProviderModelInput USD/1MCache read USD/1MOutput USD/1MBasket tierBasket GBP
Alibaba CloudQwen3.8 Max 09021.65—4.951standard0.403670
AnthropicClaude Fable 5.1100.2550standard2.822732
AnthropicClaude Opus 550.5025standard1.411366
AnthropicClaude Opus 5.540.2020standard1.129093
DeepSeekDeepSeek V4 Pro (peak)1.320.0443.96peak0.322921
GoogleGemini 3.1 Pro Preview20.20124/18 long context1.091457
Moonshot AIKimi K330.3015standard0.846820
OpenAIGPT-5.6 Sol40.4020standard1.129093
OpenAIGPT-6 Astra10150standard2.822732
xAIGrok 4.620.5064/12 long context0.978547
xAIGrok 4.720.5064/12 long context0.978547
October 2026 fixed token basket costs in pounds, separated into quick, general and frontier groups.
Separate scales keep the smaller baskets readable. The tables are the accessible numerical record.

The qualifications that can change the bill

  • Long context: Gemini 3.1 Pro Preview and Grok 4.6/4.7 use higher tiers for the frontier basket. OpenAI's uplift starts above 272,000 input tokens, so that basket does not cross it.
  • Temporary prices: Gemini 3.7/3.8 Flash's introductory rates run through 31 December, with higher published rates from 1 January. GPT-5.6 Sol's promotion is available at least through 21 November, not necessarily ending that day.
  • Qwen regions: Flash and Max use Frankfurt Global cards. Plus uses the Singapore International alias's 20% reduction, with no expiry specified. These are not interchangeable regional tariffs or residency guarantees.
  • DeepSeek time bands: the table uses conditional peak prices. Off-peak rates are half; weekends and Chinese public holidays are entirely off-peak.
  • Caching: a cache-read price is not the price of creating or storing a cache. Opus 5.5's $0.20 reads sit beside $5 five-minute or $8 one-hour writes. Check whether your workflow can reuse context.
  • Extras: subscriptions, VAT, tools, grounding, cache storage, dedicated capacity, regional premiums, Batch, Flex and fast modes are outside these baskets.

September to October: what actually moved?

The 25 matched entries use the same recorded IDs and uncached baskets. Their original dollar rates are unchanged in this check. The pound lines rise because the conversion changed, not because those providers increased their recorded token rates.

Haiku can now be compared against September's standard tariff. Its earlier August batch-equivalent error has not been quietly rewritten. DeepSeek Pro's comparison uses the same resolved version and conditional peak tariff. Matching an alias does not prove its underlying model weights never changed.

September to October pound costs for 25 matched model entries; their unchanged dollar baskets rise about 1.73 percent with the exchange rate.
Two observations, not a forecast. New and replacement models have no historical lines.
See the 25 matched-model calculations
Uncached basket costs in GBP at each month's recorded exchange rate; provider-rate status is in USD
BasketProviderModelSeptember GBPOctober GBPDollar basket
QuickAlibaba CloudQwen3.8 Flash0.0011190.001138Unchanged
QuickAnthropicClaude Haiku 4.50.0110990.011291Unchanged
QuickGoogleGemini 3.5 Flash-Lite0.0040690.004140Unchanged
QuickMistralMistral Small 40.0015540.001581Unchanged
QuickOpenAIGPT-5.6 Luna0.0023680.002409Unchanged
GeneralAlibaba CloudQwen3.7 Plus0.0165740.016861Unchanged
GeneralAnthropicClaude Sonnet 50.1109850.112909Unchanged
GeneralGoogleGemini 3.7 Flash0.0416190.042341Unchanged
GeneralGoogleGemini 3.8 Flash0.0416190.042341Unchanged
GeneralMistralMistral Large 30.0240470.024464Unchanged
GeneralMistralMistral Medium 3.50.0832390.084682Unchanged
GeneralMoonshot AIKimi K2.60.0499430.050809Unchanged
GeneralMoonshot AIKimi K2.7 Code0.0499430.050809Unchanged
GeneralOpenAIGPT-5.6 Terra0.1183840.120437Unchanged
GeneralxAIGrok 4.30.0554930.056455Unchanged
GeneralxAIGrok Build 0.10.0443940.045164Unchanged
FrontierAlibaba CloudQwen3.8 Max 09020.3967900.403670Unchanged
FrontierAnthropicClaude Fable 5.12.7746252.822732Unchanged
FrontierAnthropicClaude Opus 51.3873121.411366Unchanged
FrontierDeepSeekDeepSeek V4 Pro (peak)0.3174170.322921Unchanged
FrontierGoogleGemini 3.1 Pro Preview1.0728551.091457Unchanged
FrontierMoonshot AIKimi K30.8323870.846820Unchanged
FrontierOpenAIGPT-5.6 Sol1.1098501.129093Unchanged
FrontierOpenAIGPT-6 Astra2.7746252.822732Unchanged
FrontierxAIGrok 4.60.9618700.978547Unchanged

What I would do this month

I would not change every model simply because there is a new name. Take one repeatable task, keep the acceptance checks the same, and compare the current route with one plausible alternative.

GPT-6 Luna is a new candidate for bounded high-volume work. GPT-6.1 Sol and Sonnet 5.5 deserve a look for normal delivery. Opus 5.5 offers a lower token entry price than Opus 5. None of that tells us which will finish your particular job properly.

Run the same task with two suitable models. Record uncached input, cache reads and writes, billable output, tools, retries, time and whether the result passed. Compare the cost per accepted result. Ask before changing the live default.

Check the next expiry date, set a spending limit, and keep human review in the cost calculation. A cheap model that creates more rework is not a bargain.

More choice is useful. Deliberate spending is better.

Previous edition: What Does AI Cost This Month? September 2026. The July opening edition explains why this watch measures prices rather than crowning a winner.

Sources and limits

This is provider-published direct self-service API pricing checked on 1 October 2026. It is not an audit of invoices, a recommendation to buy, or a measured performance benchmark. No paid model benchmark runs were made.

Each row links to its current official billing source. I also checked dated updates: OpenAI, Anthropic, Google, xAI, Mistral, DeepSeek, Kimi and Alibaba Cloud. Product or harness updates are not automatically changes to the direct text tariff.

The shortlist is deliberately limited. Image, audio, embeddings, specialist models, high-speed variants, self-hosting and negotiated services need different calculations. Prices, access and endpoint behaviour can change after this snapshot. Check current provider terms before committing spend.