Skip to content
llm-spend
GitHub
Cost anatomy

GPT-6 Astra

OpenAI / Azure OpenAIFoundry ·Global

Fresh input / cached input / output cost breakdown, active and scheduled rates, provenance, same-model deployment markup, and cost-comparable alternatives for this purchasable lane.

Start with a pattern

Illustrative workloads

Pick a useful starting shape, then fine-tune every input below.

Fine-tune

Workload and pricing scenario

Adjust the token shape, cache behavior, time window, and service tier. Every number below updates immediately.

Interactive

Workload cost calculator

60M
210K
90%

Share of input tokens served from cache. Applied only to models with a cache meter.

Time

Rows with a rate that changes by time of day, promo window or service tier are priced for this scenario instead of their flat listed rate; a brand-colored label under the model name names which one applies. Context size for band-priced rows uses the input-token count below.

Picking a specific hour also previews time-of-day rates that are scheduled but have not started billing yet; those rows are labelled · preview with their start instant. Now only ever shows rates that are billable at this moment.

Rate state

What this lane charges right now

Priced under the workload and scenario above. The full published schedule below always shows every variant regardless of which scenario is selected.

Input / 1M
$10.00
CHF 8.05
Cached input / 1M
$1.00
CHF 0.805
Output / 1M
$50.00
CHF 40.25
Cost anatomy

Where this workload's cost goes

60M input tokens (90% cache hit) and 210K output tokens, at the resolved rate above.

$60.00
Fresh input · CHF 48.30
$54.00
Cached input · CHF 43.47
$10.50
Output · CHF 8.45
$124.50
Workload total · CHF 100.22

$60.00 + $54.00 + $10.50 = $124.50

Provenance

How much to trust these numbers

InputofficialCached inputofficialOutputofficial

GPT-6 Astra Standard Global short-context rate. Microsoft also publishes a $12.50/M cache-write rate; this catalog models cached-input reads, not cache creation.

Microsoft Azure Foundry announcement (https://azure.microsoft.com/en-us/blog/gpt-6-astra-frontier-intelligence-for-work-now-generally-available-in-microsoft-foundry/), published 2026-09-03 and captured 2026-09-05: the GPT-6 Astra pricing table lists Standard Global short context at $10.00/M input, $1.00/M cached input, $12.50/M cached writes and $50.00/M output. A fresh full paged Azure Retail Prices API sweep on 2026-09-05 found no meter containing Astra or GPT-6 yet, so this row follows Microsoft's published Foundry table pending retail-meter publication. Cache writes are outside this schema.

Direct vs Foundry

Same-model deployment markup

How much this Foundry lane costs over the same model's Direct API rate.

Not available — no Direct API listing exists for this exact model in the catalog.

Cost-comparable

Other lanes within ±25% of this workload's cost

Same provider first, then ranked by cost proximity, for 60M in / 210K out. This is a cost ranking only — never a quality recommendation.

  • GPT-5.5 Long Context
    OpenAI / Azure OpenAISame provider
    Foundry ·Global
    $123.45
    CHF 99.38
    -$1.05 (-1%) vs this lane
  • GPT-5.6 Sol Long Context
    OpenAI / Azure OpenAISame provider
    Foundry ·Global
    $123.45
    CHF 99.38
    -$1.05 (-1%) vs this lane
  • GPT-6 Astra
    OpenAI / Azure OpenAISame provider
    Foundry ·Data Zone
    $136.95
    CHF 110.24
    +$12.45 (+10%) vs this lane
  • Claude Fable 5
    Claude
    Direct API
    $124.50
    CHF 100.22
    $0.00 (0%) vs this lane
  • Claude Mythos 5
    Claude
    Direct API
    $124.50
    CHF 100.22
    $0.00 (0%) vs this lane
  • DeepSeek-V4 Pro
    DeepSeek · Fireworks direct API
    Direct API
    $105.13
    CHF 84.63
    -$19.37 (-16%) vs this lane