GPT-6 Astra

OpenAIflagship New

OpenAI's 3 Sep 2026 flagship and the first model in the GPT-6 generation. Joint #1 on the Artificial Analysis Intelligence Index v4.3 at launch at 53, tied with Claude Fable 5.1 (both since passed by Claude Opus 5.5 at 58 and Claude Sonnet 5.5 at 56 on v4.3.2) — but at roughly a third of Fable's cost per task ($3.26 vs $7.63 at max effort), because it reaches the same score with markedly fewer steps and fewer output tokens. Trained on OpenAI's largest run to date, over 100,000 GPUs at the Stargate site in Texas, and the first OpenAI model to use other models in a significant supervisory role during training. Set a record on DeepSWE v1.1 at 74%. Context is 1,050,000 tokens but the shape matters: 922K maximum input, 128K maximum output, and requests above 272K input tokens bill at 2x input and 1.5x output for the entire request, not just the overflow — a cliff that a million-token window actively invites you to walk off. Five reasoning-effort levels (low through max) act as a cost dial: 53/51/50/46 at max/high/medium/low against $3.26/$1.72/$1.54/$0.82 per task, which makes medium the sensible default for most workloads. Batch and Flex run at half the standard rate; Fast mode doubles it. Chat Completions, Responses and Batch only — no Realtime, Assistants, audio, embeddings or moderation. First OpenAI model to reach the company's internal 'Critical' cybersecurity threshold, scoring 100% on ExploitBench and finding two zero-days in a modified run, so the advanced cyber capabilities are access-gated via the Daybreak programme. Note for compliance teams: the model uses opaque recurrence, which OpenAI's chief scientist has said makes chain-of-thought auditing harder.

Context Window

1M

tokens

Max Output

128K

tokens

Input Price

$10

per 1M tokens

Output Price

$50

per 1M tokens

Speed

53

tokens/sec

Released

Sep 2026

2026-09-03

Blended Cost

$30.00

per 1M tokens

Value Score

3.3

quality per $

Capabilities

ChatVisionFunction CallingCode GenerationReasoningSearch

Benchmarks

Quality Index
99

Price vs Quality

6080100$1.00$10.00$100.00Blended cost per 1M tokens (log scale) →GPT-4o — $6.25/1M, quality 85GPT-4o Mini — $0.375/1M, quality 72o3 Mini — $2.75/1M, quality 88o3 — $5.00/1M, quality 94GPT-4.1 — $5.00/1M, quality 89Claude Opus 4 — $45.00/1M, quality 91Claude Sonnet 4 — $9.00/1M, quality 88Claude 3.5 Haiku — $2.40/1M, quality 75Gemini 2.5 Pro — $5.63/1M, quality 92Gemini 2.0 Flash — $0.250/1M, quality 74Llama 4 Maverick — $0.400/1M, quality 80Llama 4 Scout — $0.275/1M, quality 71Mistral Large 2 — $4.00/1M, quality 79Codestral — $0.600/1M, quality 76DeepSeek V3 — $0.685/1M, quality 86DeepSeek R1 — $1.37/1M, quality 91Grok 3 — $9.00/1M, quality 87Grok 3 Mini — $0.400/1M, quality 78Command R+ — $6.25/1M, quality 68Amazon Nova Pro — $2.00/1M, quality 70Qwen 2.5 72B — $0.600/1M, quality 80Qwen 2.5 Coder 32B — $0.300/1M, quality 74Sonar Pro — $9.00/1M, quality 78GPT-5.5 — $17.50/1M, quality 97GPT-5.5 Pro — $105.00/1M, quality 96Claude Opus 4.7 — $15.00/1M, quality 96Gemini 3.1 Pro — $7.00/1M, quality 96DeepSeek V4 Pro — $1.32/1M, quality 89DeepSeek V4 Flash — $0.210/1M, quality 84Qwen 3.6 Plus — $3.50/1M, quality 86Grok 4.3 — $1.88/1M, quality 93Kimi K2.6 — $2.11/1M, quality 92Kimi K3 — $9.00/1M, quality 97Claude Sonnet 4.6 — $9.00/1M, quality 90GLM-5.1 — $2.03/1M, quality 88Claude Opus 4.8 — $15.00/1M, quality 98Qwen 3.7 Max — $5.00/1M, quality 94Claude Fable 5 — $30.00/1M, quality 99MiniMax M3 — $1.50/1M, quality 89GPT-5.6 Sol — $17.50/1M, quality 98GPT-5.6 Terra — $7.00/1M, quality 93GPT-5.6 Luna — $0.700/1M, quality 83Gemini 3.6 Flash — $2.25/1M, quality 90Gemini 3.5 Flash-Lite — $1.40/1M, quality 79Claude Sonnet 5 — $9.00/1M, quality 93GLM-5.2 — $2.90/1M, quality 91Claude Opus 5 — $15.00/1M, quality 99Claude Mythos 5 — $30.00/1M, quality 99Grok 4.5 — $4.00/1M, quality 94Qwen3.8 Max — $4.00/1M, quality 96Claude Fable 5.1 — $30.00/1M, quality 99DeepSeek V4.1 Flash — $0.750/1M, quality 92GLM-5.3 — $2.90/1M, quality 96GLM-5.3 Flash — $0.325/1M, quality 93Muse Spark 1.3 — $2.75/1M, quality 97Grok 4.6 — $4.00/1M, quality 95Gemini 3.8 Flash — $2.25/1M, quality 93Claude Opus 5.5 — $12.00/1M, quality 100Claude Sonnet 5.5 — $6.00/1M, quality 99GPT-6.1 Sol — $6.00/1M, quality 98GPT-6 Sol — $6.00/1M, quality 97GPT-6 Luna — $0.300/1M, quality 89GPT-6 Astra — $30.00/1M, quality 99GPT-6 Astra
Cheaper is left, more capable is up. 3 of 62 priced models are both cheaper and score at least as high. Self-host models are left off the price axis — their cost depends on the hardware you run them on, not on tokens.

Context Window in Context

  • ▸ GPT-6 Astra1.1M
  • GPT-5.6 Sol1.1M
  • GPT-5.6 Terra1.1M
  • GPT-5.6 Luna1.1M
  • GPT-6.1 Sol1.1M
  • GPT-6 Sol1.1M
  • GPT-6 Luna1.1M
  • Kimi K31.0M
Bar length is logarithmic — each full step is 10× the tokens. Shown against the models with the nearest context windows, not the directory extremes.

Ainda não está no nosso gateway

A Swfte Connect não tem uma chave para GPT-6 Astra, então não há nada para executar aqui. Adicione a sua própria chave do provedor e chame-o com a mesma API de qualquer outro modelo deste diretório.

A requisição
curl https://api.swfte.com/agents/v2/gateway/chat/completions \
  -H "Authorization: Bearer $SWFTE_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "model": "openai:gpt-6-astra",
    "messages": [{"role": "user", "content": "Hello"}]
  }'
Implante este modelo com a Swfte Connect

About GPT-6 Astra

GPT-6 Astra is a flagship AI model by OpenAI, released on September 3, 2026. It supports a context window of 1050K tokens and can generate up to 128K output tokens.

At $10 per million input tokens and $50 per million output tokens, its blended cost of $30.00/1M tokens places it in the premium pricing tier. Its value score of 3.3 reflects the balance of quality and cost.

Using GPT-6 Astra with Swfte

Access GPT-6 Astra through Swfte Connect, our unified LLM gateway. Connect gives you a single API for 50+ models, with automatic routing, cost optimization, and fallback handling. You can also try GPT-6 Astra in our AI Playground before integrating.

Deploy a model with Swfte Connect

One gateway, every provider, per-token cost visibility. Swap models without touching your code.