
gpt-5.6-terra
API Overview
GPT-5.6-Terra is OpenAI’s next-generation, closed-source, highly balanced large model specifically designed for large-scale production and intelligent agents. As the pivotal mid-range model bridging the GPT-5.6 family, Terra achieves a perfect balance between “intelligence and cost” in its design. The model not only seamlessly inherits the powerful capabilities of the original GPT-5.5 in everyday full-stack programming, long-chain tool calls, and logical routing, but also delivers a significant leap forward in inference throughput speed. In mainstream long-term agent benchmarks such as Agents' Last Exam, Terra outperforms its foreign counterparts at the flagship level in both efficiency and cost-effectiveness, making it the preferred choice for long-term hosting of high-frequency core enterprise business operations.
───────────────────────────────────────────────────────────────────
Core Capabilities
High-cost-performance agent foundation: Terra delivers stunning performance in the industry-recognized, heavy-duty long-chain agent benchmark Agents' Last Exam (covering 55 specialized fields). While maintaining extremely high planning success rates and tool-call success rates, Terra slashes enterprises’ implicit total token expenditure by more than 60% when tackling the same complex tasks.
Enterprise-grade high-concurrency, ultra-fast throughput: Terra has undergone rigorous operator optimization tailored specifically for large-scale, high-concurrency enterprise workflows. The model boasts an average throughput of up to 62 tokens per second, with a dramatically reduced first-token latency (TTFT). Its structured output and JSON-mode fault tolerance approach nearly 100%, making it perfectly suited for high-frequency backend automation pipelines and large-scale batch extraction.
Multi-modal perception and agile development collaboration: Terra features robust cross-modal alignment capabilities for both text and vision. In high-frequency scenarios such as developers’ daily interactions, agile debugging of medium-to-large projects involving multiple files, and OSWorld-level GUI automation, Terra maintains exceptionally stable instruction focus and remarkably low logical deviation rates.
Million-token-long-context alternative: By default, Terra supports an enormous context window of 1.05 million tokens. It can effortlessly handle enterprise-wide codebases spanning multiple directories, comprehensive industry compliance manuals, or lengthy conversation histories. Coupled with OpenAI’s newly evolved explicit boundary caching mechanism, Terra significantly reduces enterprises’ long-term memory maintenance costs while ensuring flagship-level performance in long-text recall.
Playground
Log in to explore more features! Click to Log In