doubao-seed-2-1-pro-260628

doubao-seed-2-1-pro-260628

ByteDance’s next-generation flagship model, designed for the era of Coding and Agents, featuring deep thinking capabilities.
2026-06-24
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Cache Creation:
$0.17/1M tokens
Cache Read:
$0.0024/1M tokens
Input:
$0.86/1M tokens
Output:
$4.28/1M tokens
Bulk order? Contact your manager for exclusive deals

API Overview

The doubao-seed-2-1-pro is ByteDance’s next-generation flagship model, designed for the era of Coding and Agents. This model has achieved a breakthrough in three key areas: complex software engineering delivery, fully autonomous execution of long-chain tasks, and multimodal screen-level understanding. Its core technical metrics comprehensively surpass those of leading-edge models from major international competitors. The model boasts exceptional cutting-edge logic and dynamic self-repair capabilities, enabling it to seamlessly support ultra-long-cycle closed-loop tasks lasting dozens of hours without interruption. It is well-suited for high-value enterprise R&D scenarios and complex agent orchestration. ───────────────────────────────────────────────────────────────────

Core Capabilities

Engineering-grade Fully Autonomous Coding Delivery: The model possesses the capability for long-term, fully closed-loop software engineering delivery. It has firmly established itself among the global leaders in cutting-edge coding benchmarks such as Terminal Bench 2.1 and SWE-Pro, closely mirroring real-world terminal environments. In rigorous RTL testing for chip design, the model can run autonomously for nearly 18 consecutive hours, undergoing multiple rounds of iteration and flawlessly completing all complex hardware engineering workflows—including simulation, testing, and comprehensive verification. Agent Swarm Collaboration: Specifically designed for agent clusters and swarm workflows, the model features exceptionally stable tool invocation and long-term planning constraints. In practical applications, it can natively support over 500 intelligent agents simultaneously collaborating at scale within the same 3D virtual space, efficiently executing thousands of complex tool and API coordination sequences. End-to-end Multimodal Perception Across All Scenarios: Its vision-language (VLM) capabilities rank among the top performers in real-world benchmarking platforms such as OSWorld and MobileWorld. The model can natively perform high-precision screen reading on complex GUIs, deeply understand cross-platform video content, high-density charts, and extremely intricate engineering drawings, fully integrating visual perception, logical reasoning, and automated execution into a seamless closed loop. Disruptive Production-Ready Cost Efficiency: While maintaining top-tier, cutting-edge “high-IQ” reasoning capabilities, thanks to Volcano Engine’s industry-leading distributed operator acceleration and exceptionally high MoE routing efficiency, the model’s overall invocation cost is only about 20% of that of comparable foreign models. It can perfectly handle enterprise-level daily traffic loads totaling trillions of requests per day.

Playground

Log in to explore more features! Click to Log In

API Analytics

API Reference (1)

API DescriptionAPI EndpointRequest MethodStabilityParameter Description
Chat (ByteDance Doubao-Vision)
POST
Stable
View Details

API Pricing

$
ModelDescriptionContextOfficial Price302.AI PriceCache Price

doubao-seed-2-1-pro-260628

-
256000

Input$0.86 / 1M tokens
Output$4.28 / 1M tokens

Input$0.86/ 1M tokens
Output$4.28/ 1M tokens
Original Price

Cache Creation$0.17/ 1M tokens
Cache Read$0.0024/ 1M tokens