/v1/responses). According to OpenAI’s official model guidance, gpt-5.6-terra balances intelligence and cost within the GPT-5.6 family.
Use gpt-5.6-terra when you want strong GPT-5.6 reasoning, coding, and tool-use capability while keeping cost below the Sol tier.
Key capabilities
- Responses API — Uses the newer
/v1/responsesendpoint withinputinstead ofmessages - Balanced tier — GPT-5.6 model that balances intelligence and cost
- Long context — Supports a 1.05M token context window and up to 128K output tokens
- Advanced reasoning — Supports reasoning effort levels: none, low, medium (default), high, xhigh, and max
- Pro mode — Set
reasoning.modetoprofor quality-first requests that can tolerate higher latency and token use - Persisted reasoning — Use
reasoning.contextto control whether prior reasoning is reused across turns - Verbosity control — Set response verbosity to low, medium, or high
- Streaming — Supports real-time token streaming via SSE
- Tool use — Supports function calling, web search, file search, code interpreter, computer use, and other Responses API tools
- Programmatic Tool Calling — Supports bounded tool-heavy workflows where the model can write JavaScript to coordinate eligible tools
Quick example
Parameters
API Reference
View the interactive API playground for GPT-5.6 Terra.