API models · Side by side

GPT-6 Astra vs Gemini 3.1 Pro Preview

Compare input formats, context limits and API pricing for two large-context models. Gemini is a preview listing, so version stability is part of the decision.

The short answer

  • Gemini lists audio and video input in addition to text, images and files. Astra lists text, images and files.
  • Base token prices differ substantially. Longer prompts trigger different pricing tiers for each model.
  • A context limit describes capacity, not a promise that every answer will use the entire document accurately.

Compare the details

Compare context, supported inputs and usage costs. Highlighted rows show a difference.

GPT-6 Astra vs Gemini 3.1 Pro Preview specifications
FeatureGPT-6 AstraGemini 3.1 Pro Preview
Input formatsfile, image, textaudio, file, image, text, video
Output formatstexttext
Context / position limit1,050,0001,048,576
Maximum output tokens128,00065,536
Developer featuresReasoning output, Completion limit, Output limit, Reasoning controls, Thinking effort, Response format, Seed, Structured output, Tool selection, Tool calling, Response lengthReasoning output, Output limit, Reasoning controls, Thinking effort, Response format, Seed, Stop sequences, Structured output, Temperature, Tool selection, Tool calling, Sampling controls
Cached input / 1M tokens$1.00$0.20
Cache write / 1M tokens$12.50$0.38
USD input / 1M tokens$10.00$2.00
USD output / 1M tokens$50.00$12.00
Model identifieropenai/gpt-6-astragoogle/gemini-3.1-pro-preview

API pricing

Prices are in USD per million tokens. Input is what you send; output is what the model generates. Claude and ChatGPT subscriptions are billed separately.

GPT-6 Astra: long-prompt rates
  • Long-prompt threshold: 272,000 prompt tokens. Input: $20.00 / 1M. Output: $75.00 / 1M.
Gemini 3.1 Pro Preview: long-prompt rates
  • Long-prompt threshold: 200,000 prompt tokens. Input: $4.00 / 1M. Output: $18.00 / 1M.

Tools, images and additional reasoning can add charges. The examples below cover input and output tokens only.

Compare local models

What those prices mean in practice

Each example uses the same token budget for both models. These are calculations, not measured workloads. Cache discounts, media, tools and extra reasoning are excluded.

Token cost examples
WorkloadGPT-6 AstraGemini 3.1 Pro Preview
1,000 short requestsPer request: 2,000 input + 500 output tokens$45.00$10.00
100 document summariesPer request: 20,000 input + 1,000 output tokens$25.00$5.20
10 long-document requestsPer request: 300,000 input + 2,000 output tokens$61.50$12.36

Estimate = requests × (input tokens × input price + output tokens × output price) ÷ 1,000,000. Long-prompt rates are applied where the pricing data provides them. Real requests can use different amounts of output.

Choosing for your project

Start with the formats you need to send, then compare the context allowance and the cost examples. For an existing application, check tool calls and response formats before changing its model. A higher context limit or a lower token price does not, on its own, tell you which answer will be more useful.