The official model page defines gpt-6-astra input and output capabilities, context and output limits, reasoning levels, pricing, endpoints, tools, and rate limits for API planning and cost estimation.
Suitable tasks: Complex reasoning, coding, computer use, research, and document creation in workflows with text and image input and text output.
Unsuitable tasks: Audio or video input/output, fine-tuning, and free-tier API usage.
Applicable model version: gpt-6-astra; the page lists gpt-6-astra for both snapshot and alias entries.
Applicable client, agent, or API: The listed API endpoints, including Chat Completions and Responses; treat tool support as Responses API support.
Recommended reasoning effort and parameters: low, medium, high, xhigh, or max; the page does not prescribe a default for a specific task.
model: gpt-6-astra
reasoning.effort: low | medium | high | xhigh | max
input: text, image
output: text
context_window: 1,050,000 tokens
max_output_tokens: 128,000
knowledge_cutoff: 2026-04-30This is a configuration summary normalized from the page's fields, not a verbatim code sample.
| Item | Official value |
|---|---|
| Standard input | $10.00 / 1M tokens |
| Cached input | $1.00 / 1M tokens |
| Cache writes | $12.50 / 1M tokens |
| Output | $50.00 / 1M tokens |
| More than 272K input tokens | 2× input/cache and 1.5× output rates for the full request |
| Batch / Flex | 50% of Standard rates |
| Fast mode | 2× applicable rates |
| Free tier | Not supported |
The page lists web search, file search, image generation, code interpreter, hosted shell, apply patch, skills, computer use, MCP, and tool search as Responses API tools. The model supports streaming, function calling, and structured outputs, and does not support fine-tuning.
The page lists RPM, TPM, and batch queue limits for Tiers 1–5. Actual limits vary by usage tier and should not be presented as a universal guarantee.
Chat Completions, Responses, Realtime, and Batch are listed as endpoints; the tools section explicitly scopes tool support to Responses API.
Pricing and capabilities are the official values visible on the collection date and may change. Reproduction should record collection date, input token count, cache status, and processing mode.
GPT-6 Astra