Claude Opus 5.5 is live on SeedRouter
OpenAIText generation

GPT-6 Astra — OpenAI's most capable model for hard problems

GPT-6 Astra is built for the hardest end-to-end work: reasoning, coding and research, with a 1.05M-token context. $8.00 / 1M input, $40.00 / 1M output per 1M tokens.

Available$8.00 / 1M input · $40.00 / 1M output$10.00 / 1M input · $50.00 / 1M output-20%
  • Reasoning effort control
  • 1.05M-token context
  • Image input
  • Prompt caching
  • Streaming
Context window
1.05M tokens
Max output
128K tokens
Effort levels
low to max
Knowledge cutoff
Apr 30, 2026

Try GPT-6 Astra

Chat with GPT-6 Astra, attach images, and see the tokens and cost of every answer.

Parameters

gpt-6-astra$8.00 in · $40.00 out per 1M tokens

Reasoning effort
Reasoning summary
Reasoning mode
Reasoning context
Verbosity

Temperature and Top P work only with reasoning effort none.

Conversation total$0.000000
Ready

Try one of these prompts

Messages are billed from your balance at the rates below, like an API call. Failed requests are free.

Real GPT-6 Astra output

One prompt, sent to GPT-6 Astra at low reasoning effort. Unedited.

Explain the difference between a mutex and a semaphore in two short paragraphs, and say when to use each.

A **mutex** is a lock that lets only one thread access a shared resource at a time. It has ownership: the thread that locks it must unlock it. Use a mutex to protect shared data or a critical section, such as updating a shared counter. A **semaphore** maintains a count of available permits: a thread acquires a permit to proceed and releases one when appropriate. Unlike a mutex, it has no ownership requirement, so one thread can release a permit for another. Use a semaphore to limit concurrent access to a resource pool or to signal that work is available.

Generated 2026-09-28. Answers vary from run to run.

GPT-6 Astra pricing

GPT-6 Astra is billed per token from your balance. Rates below are live, the same ones your account is charged.

One typical request2,000 input tokens, 1,000 output tokens
$0.056per request
Reasoning counts as output, so higher effort raises the output side.
What $20.00 buysOutput tokens at the current rate
500,000output tokens
Input and cached tokens cost less per token than output.

Example sizes are for illustration. Your bill follows the tokens each request reports.

ItemCounted asRateBilling
InputPrompt tokens not read from or written to the cache$8.00 / 1MPer 1M tokens
OutputAnswer and reasoning tokens$40.00 / 1MPer 1M tokens
Cached inputPrompt tokens served from the prompt cache$0.80 / 1MPer 1M tokens
Cache writePrompt tokens written to the prompt cache$10.00 / 1MPer 1M tokens
Input, over 272KInput when the prompt exceeds 272K tokens$16.00 / 1MPer 1M tokens
Output, over 272KOutput when the prompt exceeds 272K tokens$60.00 / 1MPer 1M tokens
Cached input, over 272KCache reads when the prompt exceeds 272K tokens$1.60 / 1MPer 1M tokens
Cache write, over 272KCache writes when the prompt exceeds 272K tokens$20.00 / 1MPer 1M tokens
Failed requestAny request that returns an errorFreeNot charged

What is GPT-6 Astra?

GPT-6 Astra is OpenAI's most capable model, built for the hardest end-to-end work: complex reasoning, coding, research and document creation. GPT-6 Astra reads up to 1.05M tokens of context and writes up to 128K tokens per request.

GPT-6 Astra always reasons: effort runs from low to max, and none is not supported. Raise the effort for deeper work and lower it when you need a faster, cheaper answer.

Model ID: gpt-6-astra · Input: text, images · Output: text · Formats: OpenAI Responses, OpenAI Chat Completions, Anthropic Messages.

Model ID
gpt-6-astra
Context window
1.05M tokens
Max output
128K tokens
Effort levels
low to max
Input
Text, images

Two ways to use GPT-6 Astra

Call it from your code, or let your coding agent do it for you.

API

Call it from your backend

Best for apps and services

Point the official OpenAI SDK at SeedRouter and keep your code. GPT-6 Astra answers both the Responses and Chat Completions formats.

  1. 1Create an API key
  2. 2Set the base URL to https://api.seedrouter.ai/v1
  3. 3Send gpt-6-astra as the model
Agent

Hand it to your coding agent

Best for Claude Code, Codex and other agents

Copy a ready prompt. Your agent shows the request and cost before anything is spent.

  1. 1Copy the prompt
  2. 2Paste it into your agent
  3. 3Approve the request

What GPT-6 Astra is good at

GPT-6 Astra is for tasks where getting it right matters more than getting it fast.

Complex reasoning

GPT-6 Astra works through multi-step problems that simpler models get wrong.

Hard coding tasks

Design systems, untangle tricky bugs and plan large refactors.

Research and analysis

Feed it reports and papers and get careful, grounded comparisons.

Document creation

GPT-6 Astra drafts design docs, specs and reports you can ship.

Images in the prompt

GPT-6 Astra reads diagrams, charts and screenshots alongside your text.

Cheaper repeat context

Cached input on GPT-6 Astra costs 10% of the input rate.

What teams build with GPT-6 Astra

Architecture reviews

Have GPT-6 Astra check a design before your team builds it.

Long agent runs

Let an agent plan and carry out multi-step work with the whole context in view.

Analyst reports

Turn raw research into structured findings with a text.format schema.

Start using GPT-6 Astra in four steps

  1. 01

    Create a key

    Sign in and create an API key. Add credit whenever you need it; there is no subscription.

  2. 02

    Change the base URL

    Point the OpenAI SDK at https://api.seedrouter.ai/v1. Your code stays the same.

  3. 03

    Pick GPT-6 Astra

    Send gpt-6-astra as the model. Set reasoning effort from low to max to trade depth for speed and cost.

  4. 04

    Check the bill

    Every request shows its tokens and charge in your usage records.

Get an API key

GPT-6 Astra on SeedRouter at a glance

What you can use today.

FeatureGPT-6 Astra
OpenAI ResponsesYes, official format
OpenAI Chat CompletionsYes
Anthropic MessagesYes
Image inputYes, by URL
StreamingYes
Prompt cachingYes, cached input and cache writes billed apart
Reasoning effort noneNo, the lowest effort is low
Failed requestsNot charged

Batch, Flex, background mode and service tiers are not offered.

GPT-6 Astra limits

1.05M-token context

On GPT-6 Astra, prompt, images and history share one 1.05M-token window; input is capped at 922K.

128K tokens out

One GPT-6 Astra request can return up to 128K tokens, reasoning included.

Long prompts cost more

Above 272K input tokens, the whole GPT-6 Astra request is billed at the higher long-context rates.

No temperature or top_p

GPT-6 Astra always reasons, so it does not accept temperature, top_p or logprobs.

Why call GPT-6 Astra through SeedRouter

One key, three formats

Call GPT-6 Astra with OpenAI Responses, Chat Completions or Anthropic Messages, using the same key.

Pay per token

Top up once and spend it on any model. No plan, no monthly fee.

Failures are free

A request that errors is not charged.

Live prices

The rates on this page are the rates your account pays.

Prompt caching

Repeated prompt prefixes are read from the cache at a lower GPT-6 Astra rate.

Image input

Send diagrams, charts and screenshots by URL with your prompt.

Streaming

Stream GPT-6 Astra tokens as they arrive with stream: true.

Related models

Other models you can call with the same key.

GPT-6 Luna
gpt-6-luna

Call GPT-6 Luna with the OpenAI SDK. The lowest-cost GPT-6 model, with a 1.05M-token context, live per-token pricing and a playground you can try right now.

View pricing
GPT-6 Sol
gpt-6-sol

Call GPT-6 Sol with the OpenAI SDK. 1.05M-token context, reasoning effort from none to max, live per-token pricing and a playground you can try right now.

View pricing
claude-fable-5
claude-fable-5

Anthropic

View pricing
claude-fable-5-1
claude-fable-5-1

Anthropic

View pricing
claude-opus-5-5
claude-opus-5-5

Anthropic

View pricing
dreamina-seedance-2-0
dreamina-seedance-2-0

ByteDance

View pricing

GPT-6 Astra FAQ

What is GPT-6 Astra?+

GPT-6 Astra is OpenAI's most capable model, built for the hardest end-to-end work: complex reasoning, coding, research and document creation. It has a 1.05M-token context window and writes up to 128K tokens per request.

Should I use GPT-6 Astra or GPT-6 Sol?+

Use GPT-6 Astra when a task is hard and mistakes are costly. GPT-6 Sol handles demanding coding and agent work at a lower price, and GPT-6 Luna suits focused, high-volume jobs.

How much does GPT-6 Astra cost?+

GPT-6 Astra is billed at $8.00 / 1M input and $40.00 / 1M output per 1M tokens. Cached input costs 10% of the input rate. Prompts above 272K tokens use the long-context rates shown in the pricing table.

Which reasoning effort levels does GPT-6 Astra support?+

GPT-6 Astra supports low, medium, high, xhigh and max. It does not support none, so it always reasons and does not accept temperature or top_p.

What is the knowledge cutoff of GPT-6 Astra?+

GPT-6 Astra has a knowledge cutoff of April 30, 2026. For newer facts, pass them in the prompt.

What is the GPT-6 Astra model ID and how do I call it?+

The model ID is gpt-6-astra. Create a SeedRouter API key, point the OpenAI SDK at https://api.seedrouter.ai/v1 and send gpt-6-astra as the model with the Responses or Chat Completions API.

Can GPT-6 Astra read images?+

Yes. GPT-6 Astra accepts images alongside text, sent by URL in the request. It returns text only.

Am I charged if a GPT-6 Astra request fails?+

No. A GPT-6 Astra request that returns an error is not charged. You pay only for the tokens a finished request reports.

Try GPT-6 Astra now

Send your first GPT-6 Astra request from the playground or your own code in minutes.