GPT-6 Astra
Fresh input / cached input / output cost breakdown, active and scheduled rates, provenance, same-model deployment markup, and cost-comparable alternatives for this purchasable lane.
Illustrative workloads
Pick a useful starting shape, then fine-tune every input below.
Workload and pricing scenario
Adjust the token shape, cache behavior, time window, and service tier. Every number below updates immediately.
Workload cost calculator
Share of input tokens served from cache. Applied only to models with a cache meter.
Rows with a rate that changes by time of day, promo window or service tier are priced for this scenario instead of their flat listed rate; a brand-colored label under the model name names which one applies. Context size for band-priced rows uses the input-token count below.
Picking a specific hour also previews time-of-day rates that are scheduled but have not started billing yet; those rows are labelled · preview with their start instant. Now only ever shows rates that are billable at this moment.
What this lane charges right now
Priced under the workload and scenario above. The full published schedule below always shows every variant regardless of which scenario is selected.
Where this workload's cost goes
60M input tokens (90% cache hit) and 210K output tokens, at the resolved rate above.
$60.00 + $54.00 + $10.50 = $124.50
How much to trust these numbers
OpenAI API Standard short-context rate. Prompts above 272K input tokens use the separate long-context row below; Batch and Flex are half price, while Fast mode is 2x Standard.
OpenAI's official GPT-6 Astra model page (developers.openai.com/api/docs/models/gpt-6-astra), captured 2026-09-07: the live API model is `gpt-6-astra` with a 1,050,000-token context window, 128,000 max output, and $10/M input, $1/M cached input, $12.50/M cache writes and $50/M output. The OpenAI API pricing page confirms the same Standard row and publishes the Batch, Flex and Fast-mode schedules.
Same-model deployment markup
How much more Microsoft Foundry charges for this exact model, over this Direct rate.
| Lane | Deployment | Workload cost | Foundry markup |
|---|---|---|---|
| GPT-6 Astra | Foundry ·Data Zone | $136.95 | +$12.45 (+10%) |
| GPT-6 Astra | Foundry ·Global | $124.50 | $0.00 (0%) |
Other lanes within ±25% of this workload's cost
Same provider first, then ranked by cost proximity, for 60M in / 210K out. This is a cost ranking only — never a quality recommendation.
- GPT-6 AstraOpenAI / Azure OpenAISame providerFoundry ·Global$124.50CHF 100.22$0.00 (0%) vs this lane
- GPT-5.5 Long ContextOpenAI / Azure OpenAISame providerFoundry ·Global$123.45CHF 99.38-$1.05 (-1%) vs this lane
- GPT-5.6 Sol Long ContextOpenAI / Azure OpenAISame providerFoundry ·Global$123.45CHF 99.38-$1.05 (-1%) vs this lane
- GPT-6 AstraOpenAI / Azure OpenAISame providerFoundry ·Data Zone$136.95CHF 110.24+$12.45 (+10%) vs this lane
- Claude Fable 5ClaudeDirect API$124.50CHF 100.22$0.00 (0%) vs this lane
- Claude Mythos 5ClaudeDirect API$124.50CHF 100.22$0.00 (0%) vs this lane