doubao-seed-2-1-turbo-260628

doubao-seed-2-1-turbo-260628

ByteDance’s new-generation, low-cost, low-latency, deep-thinking large model designed for large-scale production scenarios
2026-06-24
LLM
Model capability: imageModel capability: videoModel capability: thinkingModel capability: function_call
Cache Creation:
$0.086/1M tokens
Cache Read:
$0.0024/1M tokens
Input:
$0.43/1M tokens
Output:
$2.14/1M tokens
Bulk order? Contact your manager for exclusive deals

API Overview

doubao-seed-2-1-turbo is ByteDance’s new-generation, low-cost, low-latency, deep-thinking large model designed for large-scale production scenarios. As an agile version of the 2.1-generation Pro flagship model, it boasts fully comprehensive functionality and delivers performance comparable to the Pro model in areas such as agile coding interactions, high-concurrency orchestration of lightweight agents, and multi-modal screen-level understanding. Thanks to its highly optimized distributed inference operators and compute resource allocation mechanism, the model achieves a groundbreaking leap in response speed and throughput. It is the top choice for enterprises seeking large-scale, batch deployment across multiple scenarios, agile workflow management, and extreme control over API budget expenditures.

───────────────────────────────────────────────────────────────────

Core Capabilities

Scalable and Agile Coding: Specifically optimized for developers’ high-frequency interactions and real-time code completion. Inheriting the core strengths of the 2.1 family from its top ranking on the SWE-Pro real-code delivery benchmark, it achieves ultra-low first-character latency for single-point function refactoring, full-stack agile fixes, and cross-file logical-flow completion, significantly reducing implicit time costs while enhancing the development experience.

High-Concurrency Multi-Agent Collaborative Workflow: Tailored specifically for high-frequency, high-throughput agent clusters and swarm orchestration scenarios. Featuring highly stable instruction adherence and tool-call success rates, it can serve as an execution sub-node in high-value workflows across various industries, seamlessly integrating into complex, multi-step automated processes and delivering highly reliable results.

Ultra-Low Latency Multi-Modal Panoramic Perception: Perfectly preserves the powerful visual-language core capabilities of the 2.1 generation. It supports low-latency, high-precision screen reading and end-to-end parsing of complex multi-platform GUIs, high-density reports, and video streams, providing millimeter-level agile feedback throughout the visual perception-to-agent-execution pipeline.

Unrivaled “Dimensionality-Reduction” Cost-Effectiveness: By default, it offers an ultra-high-fidelity context window of 256K tokens and an exceptionally large single-output limit of 128K tokens. While maintaining model performance on par with industry-leading models, its price has been cut in half again compared to the Pro level. With exceptionally strong performance and dramatic cost reductions, it easily handles enterprise-level daily traffic volumes in the trillions of concurrent requests.

Playground

Log in to explore more features! Click to Log In

API Analytics

API Reference (1)

API DescriptionAPI EndpointRequest MethodStabilityParameter Description
Chat (ByteDance Doubao-Vision)
POST
Stable
View Details

API Pricing

$
ModelDescriptionContextOfficial Price302.AI PriceCache Price

doubao-seed-2-1-turbo-260628

-
256000

Input$0.43 / 1M tokens
Output$2.14 / 1M tokens

Input$0.43/ 1M tokens
Output$2.14/ 1M tokens
Original Price

Cache Creation$0.086/ 1M tokens
Cache Read$0.0024/ 1M tokens