OpenAI

openai/gpt-5.6

250K context

GPT-5.6 has been released, and the Sol, Terra, and Luna models are fully available through OurToken with one OpenAI-compatible API.

Get API Key

Explore GPT-5.6 models

Model Introduction

OpenAI gpt-5.6

GPT-5.6 has been released, and the Sol, Terra, and Luna models are fully available through OurToken with one OpenAI-compatible API.

The GPT-5.6 family provides three production routes: Sol for maximum capability, Terra for balance, and Luna for low-cost high-volume traffic. Each model supports text and image input, text output, a 250K context window, up to 128K output tokens, function calling, and separate cache pricing.

Why It Looks Great

  • Choose Sol for demanding reasoning, coding, and agent workflows.
  • Choose Terra for a practical balance of capability, latency, and cost.
  • Choose Luna for frequent requests, high throughput, and the lowest token price.
  • Use the same OpenAI-compatible Responses API while changing only the model ID.
  • Plan repeated workloads with separate cached input and cache write rates.

Key Features

  • Models: gpt-5.6-sol, gpt-5.6-terra, gpt-5.6-luna
  • Context Window: 250000 tokens
  • Max Output: 128000 tokens
  • Input: Text and Image
  • Output: Text
  • Function Calling: Supported
  • OurToken Pricing: 20% of official price for each model

Specifications

ProviderOpenAI
FamilyGPT-5.6
AvailabilitySol, Terra, and Luna available on OurToken
ReleaseGPT-5.6 announced June 26, 2026
Model IDsgpt-5.6-sol, gpt-5.6-terra, gpt-5.6-luna
Context Window250000 tokens
Max Output128000 tokens
InputText, Image
OutputText
Function CallingSupported
OurToken Pricing20% of official price

Choose the Right GPT-5.6 API Model

Compare the three GPT-5.6 routes by workload fit, token price, caching, and production requirements.

GPT-5.6 Sol

The flagship route for difficult reasoning, advanced coding, repository work, tool use, and multi-step agents.

GPT-5.6 Terra

The balanced route for production coding, assistants, analysis, and tools when Sol-level capability is not required for every request.

GPT-5.6 Luna

The lowest-cost route for high-volume assistants, extraction, classification, content processing, and routine coding support.

Shared 250K Context

All three routes support a 250K context window and up to 128K output tokens for substantial requests.

Unified API Integration

Keep one OpenAI-compatible Responses API implementation and switch among the three complete model IDs.

Transparent Family Pricing

Sol, Terra, and Luna each show input, cached input, cache write, output, official, and discounted prices on their dedicated pages.

How to Select a GPT-5.6 Model

Match each workload to Sol, Terra, or Luna, then validate quality, latency, caching, and total request cost.

Classify the Workload

Identify task difficulty, request volume, latency target, context size, output length, and tool requirements.

01

Start with Terra

Use Terra as a balanced baseline for common production coding, assistant, analysis, and automation workloads.

02

Escalate to Sol

Move difficult reasoning, coding, or multi-step agent requests to Sol when the baseline misses quality targets.

03

Route Routine Work to Luna

Use Luna for frequent and cost-sensitive tasks that meet quality requirements without the higher-priced routes.

04

Plan Prompt Caching

Measure stable prompt prefixes and include both cached input and cache write rates in cost calculations.

05

Review Dedicated Pages

Open each model page for its exact prices, model ID, examples, features, how-to guidance, and FAQ.

06

GPT-5.6 Family FAQ

Answers about GPT-5.6 release, availability, Sol vs Terra vs Luna, context, pricing, caching, and API integration.

01

Is GPT-5.6 available through OurToken?

Yes. GPT-5.6 has been released, and gpt-5.6-sol, gpt-5.6-terra, and gpt-5.6-luna are available through OurToken.
02

What is the difference between Sol, Terra, and Luna?

Sol prioritizes maximum capability, Terra balances capability and cost, and Luna prioritizes low cost and high throughput.
03

What context window does GPT-5.6 support?

The three OurToken GPT-5.6 routes list a 250000-token context window and up to 128000 output tokens.
04

How much do GPT-5.6 models cost?

Each dedicated model page shows its own input, cached input, cache write, and output prices. OurToken lists every route at 20% of official pricing.
05

Does the family support prompt caching?

Yes. Cached input and cache writes are priced separately for Sol, Terra, and Luna, so include both in workload estimates.
06

Can I switch models without changing integrations?

Yes. Use the same OpenAI-compatible Responses API and change the model field among the three complete model IDs.
07

Which GPT-5.6 model should I choose first?

Terra is a practical baseline. Escalate hard tasks to Sol and route suitable high-volume work to Luna after testing your prompts.