gpt-5.6-terra-pro

gpt-5.6-terra-pro

Enhanced mid-range large model from OpenAI’s GPT-5.6 family
2026-07-10
LLM
Model capability: imageModel capability: function_call
Cache Read:
$0.2/1M tokens
Input:
$2/1M tokens
Output:
$12/1M tokens
Bulk order? Contact your manager for exclusive deals

API Overview

GPT-5.6-Terra is OpenAI’s next-generation, closed-source, highly balanced large model specifically designed for large-scale production and intelligent agents. As the pivotal mid-range model bridging the GPT-5.6 family, Terra achieves a perfect balance between “intelligence and cost” in its design. The model not only seamlessly inherits the powerful capabilities of the original GPT-5.5 in everyday full-stack programming, long-chain tool calls, and logical routing, but also delivers a significant leap forward in inference throughput speed. In mainstream long-term agent benchmarks such as Agents' Last Exam, Terra outperforms its foreign counterparts at the flagship level in both efficiency and cost-effectiveness, making it the preferred choice for long-term hosting of high-frequency core enterprise business operations.

───────────────────────────────────────────────────────────────────

Core Capabilities

High-cost-performance agent foundation: Terra delivers stunning performance in the industry-recognized, heavy-duty long-chain agent benchmark Agents' Last Exam (covering 55 specialized fields). While maintaining extremely high planning success rates and tool-call success rates, Terra slashes enterprises’ implicit total token expenditure by more than 60% when tackling the same complex tasks.

Enterprise-grade high-concurrency, ultra-fast throughput: Terra has undergone rigorous operator optimization tailored specifically for large-scale, high-concurrency enterprise workflows. The model boasts an average throughput of up to 62 tokens per second, with a dramatically reduced first-token latency (TTFT). Its structured output and JSON-mode fault tolerance approach nearly 100%, making it perfectly suited for high-frequency backend automation pipelines and large-scale batch extraction.

Multi-modal perception and agile development collaboration: Terra features robust cross-modal alignment capabilities for both text and vision. In high-frequency scenarios such as developers’ daily interactions, agile debugging of medium-to-large projects involving multiple files, and OSWorld-level GUI automation, Terra maintains exceptionally stable instruction focus and remarkably low logical deviation rates.

Million-token-long-context alternative: By default, Terra supports an enormous context window of 1.05 million tokens. It can effortlessly handle enterprise-wide codebases spanning multiple directories, comprehensive industry compliance manuals, or lengthy conversation histories. Coupled with OpenAI’s newly evolved explicit boundary caching mechanism, Terra significantly reduces enterprises’ long-term memory maintenance costs while ensuring flagship-level performance in long-text recall.

Playground

Log in to explore more features! Click to Log In

API Analytics

API Reference (5)

API DescriptionAPI EndpointRequest MethodStabilityParameter Description
Chat(Talk)
POST
Stable
View Details
Chat (Image Analysis)
POST
Stable
View Details
Chat (Structured Output)
POST
Stable
View Details
Chat (function call)
POST
Stable
View Details
Responses
POST
Stable
View Details

API Pricing

$
ModelDescriptionContextOfficial Price302.AI PriceOfficial Price Gap

gpt-5.6-terra

-
1000000

Input$2 / 1M tokens
Output$12 / 1M tokens

Cache Read$0.2 / 1M tokens
Input$2 / 1M tokens
Output$12 / 1M tokens

Original Price