GPT Proto

GPTProto

  • Dashboard
  • LLM

    • claude
      Claude Opus 5New
    • google
      Gemini 3.6 Flash
    • google
      Gemini 3.5 Flash Lite
    • moonshotai
      Kimi K3
    • openai
      GPT 5.6 Luna

    Image

    • bytedance
      Dola Seedream 5.0 Pro 260628New
    • google
      Gemini 3.1 Flash Lite Image
    • google
      Gemini 3.1 Flash Image
    • openai
      GPT Image 2
    • google
      Gemini 3.1 Flash Image Preview

    Video

    • kling
      Kling v3.0 4kNew
    • bytedance
      Dreamina Seedance 2.0 Mini 260615
    • kling
      Kling v3 Omni 4k
    • bytedance
      Dreamina Seedance 2.0 Fast 260128
    • bytedance
      Dreamina Seedance 2.0 260128
    Explore 214+ Models >
  • Generator

    • Create Image
    • Create Video
    • Edit in Canvas

    Features

    • Anime to Real Life AINew
    • Anime AI Art Generator
    • AI Object Remover
    • AI Image Editor
    • Unrestricted AI Image Generator
    • AI Motion Transfer
    • AI Clothes Remover
    • AI Watermark Remover
    • AI Image Enhancer Online
    • Online Background Remover Tool
    Explore All >

    Prompts

    • Seedance 2.0 PromptsNew
    • GPT Image 2 Prompts
    • Nano Banana Pro Prompts
    • Seedream 5.0 Pro Prompts
  • AI Blog

    • GLM 5.2 vs MiniMax M3: Which Is Better for Coding and Frontend Work?
    • How to Create Your Own AI Character With an API—No Coding Required
    • Kimi K3 vs Claude Opus 5: Which Is Better for Coding and AI Agents?
    • 20 Free Seedream 5.0 Pro Packaging Design Prompts for Products and E-commerce
    • GLM-5.2 vs Kimi K3 for Coding: Which Is Better for Developers in 2026?
    Explore All >

    AI Insight

    • What Is Emochi AI—and Why Is It Growing So Fast? (2026)
    • What Is Kimi K3—and Is It Really Close to GPT-5.6 and Fable 5?
    • 12 Best AI Video Generation Tools in 2026 for YouTube, TikTok, Text and Images
    • What Is Qwen 3.8 Max? Release Date, 2.4T Preview, Pricing, and Early Benchmarks
    • Gemini 3.6 Flash and Gemini 3.5 Flash-Lite Explained: Which One Should You Use?
    Explore All >

    AI Docs

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    Explore All >

    AI Skills

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    Explore All >
Pricing
English繁體中文한국어日本語EspañolРусский
Get Started Now
  1. Home
  2. /Model
  3. /Bytedance
  4. /doubao-seedance-2-0-260128
Bytedance
doubao-seedance-2-0-260128
Documentation
Documentation
Doubao Seedance 2.0 is ByteDance's second-generation video model, built on a unified audio-video architecture that takes text, image, video, and audio in a single call and returns synchronized video plus audio in one pass. This page covers the text-to-video endpoint: write a prompt, get a 4–15 second clip up to 1080p with native sound. Call doubao-seedance-2-0-260128 on GPTProto from $0.2957/run — one balance, OpenAI-compatible access, no Jimeng/Volcano Ark account or regional payment setup required.

$ 0.2957

text

video

$ 0.2957

text

video

Playground
JSON
API

Input

Your browser does not support the video tag.
Your request will cost$0per run, for$100you can run this model approximately0times
Related Models
All Models
Kling
Kling
kling-v3.0-4k
$ 1.008
$ 1.26
Bytedance
Bytedance
dreamina-seedance-2-0-mini-260615
$ 0.2365
Vidu
Vidu
viduq3-turbo
$ 0.032
$ 0.04
Qwen
Qwen
wan-2.6
$ 0.45
$ 0.5
Google
Google
veo-3.1-fast-generate-preview
$ 1.2
MiniMax
MiniMax
hailuo-2.3-pro
$ 0.441
$ 0.49
Examples
Ultra-realistic F1 live TV broadcast screenshot, identity preserved exactly from reference image.

Young woman sitting in the VIP paddock / team garage during a Formula 1 race, shown on the official live race broadcast as the girlfriend of an F1 driver. It is the final lap, and she is listening to the team radio through a professional racing headset, watching the garage monitors nervously, leaning forward with one hand near her mouth, proud tense expression.

She wears a fitted white tank top, oversized racing team jacket draped over her shoulders, large black team-radio headset with boom mic, gold jewelry, soft glam makeup. A slim paddock pass hangs naturally from her neck. 

Add realistic F1 broadcast graphics: “FINAL LAP” banner, lap counter showing final lap, driver timing tower on the left, small F1-style logo bug, “LIVE” indicator, lower-third identifying her as driver partner / paddock guest. No fake oversized badge, no selfie angle.

Team staff, headsets, garage screens, mechanics, and race equipment blurred around her. Telephoto broadcast camera from across the garage, compression artifacts, digital noise, bright paddock lighting, natural skin texture, no smoothing, 8k quality
Ultra-realistic sports broadcast still of a glamorous woman sitting in a packed football stadium crowd during a night match, wearing a dark brown sleeveless high-neck satin top and black square earrings, shoulder-length light brown/blonde hair styled in soft waves. She is casually drinking from a tall blue aluminum can while holding a half-eaten cheeseburger in the other hand. Around her are fans in bright yellow and blue football jerseys and scarves, creating strong team-color contrast. The scene feels candid and cinematic, captured mid-game from a TV broadcast camera angle with shallow depth of field. Include realistic stadium seating, crowded audience atmosphere, broadcast overlay graphics in the top-left corner showing a live football score and match timer, and a sports network watermark in the top-right. Natural arena lighting, detailed skin texture, sharp focus on the woman, slightly blurred background crowd, authentic live sports broadcast aesthetic, 16:9 composition.
A shimmering soap film stretches across a circular frame, catching the light. The secrets to achieving a perfectly spherical bubble are unveiled as the surface tension and air pressure work in harmony. This exploration reveals the simple yet elegant physics at play, creating fleeting moments of iridescent beauty.
ltra high-end editorial photography, photorealistic, 8K, razor-sharp detail, perfect skin texture with natural pores and subsurface scattering, cinematic lighting, shot on medium format, Portrait: 31-year-old Black British architect in a cream blazer, standing in a brutalist concrete stairwell, soft directional skylight, composed posture, masterful color grading, natural catchlights in eyes, tactile fabric detail, magazine cover quality, no text, no logos, no watermarks

Doubao Seedance 2.0 API: Text-to-Video with Native Audio from $0.2957/Run

The Doubao Seedance 2.0 API is ByteDance's second-generation video model, available on GPTProto with OpenAI-compatible access and no separate Jimeng or Volcano Ark account. Built on a unified audio-video architecture, doubao-seedance-2.0 takes text, image, video, and audio in a single call and returns a synchronized clip with native sound — 4 to 15 seconds, up to 4K, across aspect ratios from 16:9 and 9:16 to 21:9. It's the same Doubao AI Seedance 2.0 engine developers use for cinematic shots, product spots, and short social video, callable through one standard endpoint.

For doubao seedance 2.0 text to video, generation starts at $0.2957 per run, and doubao seedance 2.0 api pricing scales predictably with resolution and duration — see the full per-run table below. Whether you're prototyping fast 480p clips or rendering 1080p and 4K shots with audio, seedance 2.0 doubao gives you a pay-as-you-go path to ByteDance's latest video generation, with one API key that works across every model on the platform.

What is Doubao Seedance 2.0

Released in February 2026 by ByteDance's Seed team, Doubao Seedance 2.0 (also written doubao-seedance-2.0 or seedance 2.0 doubao) is built on a unified audio-video architecture. The change that matters from version 1.5 Pro is not a bigger resolution number — it is that the model treats every input as a control signal. A single generation can take a text brief plus reference images, a clip for camera motion, and an audio bed, and reason over all of them together before the first frame is produced. For this text-to-video endpoint you work from text, but the same model id is what powers the fuller reference workflow elsewhere.

Spec Value
Provider ByteDance (Seed team)
Model string doubao-seedance-2-0-260128
This endpoint Text-to-video
Input modalities (model) Text, image, video, audio
Reference capacity (model) Up to 9 images, 3 video clips, 3 audio clips per call
Native audio Yes — generated jointly with video; multilingual lip-sync
Duration 4–15s (model); 4–14s priced on GPTProto
Resolution Up to 1080p (480p / 720p / 1080p tiers)
Aspect ratios 16:9, 9:16, 4:3, 3:4, 21:9, 1:1
Editing / extension Targeted clip/character/action edits; clip extension
Image-to-video Yes — separate endpoint at /image-to-video
Variants Full + Fast (doubao-seedance-2-0-fast-260128 on GPTProto)
Open source No — proprietary

How the model handles audio and motion

Two things separate Seedance 2.0 from a standard text-to-video model, and both affect how you prompt it. First, audio and video come out of one pass rather than a silent clip with a soundtrack bolted on afterward. Because sound and image are generated together, dialogue lands on the mouth and ambient effects follow the action — so a prompt that names the sound ("crowd noise builds as the car passes") tends to produce tighter sync than describing visuals alone. Second, the model holds a subject's identity across movement and angle changes more reliably than earlier versions, which is what makes multi-shot and recurring-character work usable instead of one-off. If you have written off AI video because a face drifts between cuts, this is the axis that improved most.

Doubao Seedance 2.0 vs Sora 2

Both models are live on GPTProto, so this is a like-for-like choice rather than a marketing claim.

  Doubao Seedance 2.0 Sora 2
Inputs Text, image, video, audio Text, image
Native audio Yes, joint generation Yes, synced
Max duration 15s 12s (standard) / 25s (Pro)
Max resolution Up to 1080p 720p (standard) / 1080p (Pro)
API roadmap Active OpenAI is sunsetting the Sora 2 API on 2026-09-24
GPTProto price from $0.2957/run (scales with res × duration) $0.40/run (flat)

 

Doubao Seedance 2.0 vs Seedance 1.5 Pro

If you already run Seedance 1.5 Pro, the upgrade is about input control, not a resolution bump. Version 1.5 Pro already generated audio and video together and followed multi-shot instructions. Seedance 2.0 adds the reference stack on top — up to 9 images, 3 video clips, and 3 audio clips in a single call — plus targeted editing and clip extension. In third-party motion and physics testing it scores higher than 1.5 Pro, with the largest gain in physical accuracy.

The cost gap is large: 1.5 Pro starts at $0.0408/run versus $0.2957 for 2.0. Practical rule: stay on 1.5 Pro for plain, high-volume text-to-video where budget dominates; move to 2.0 when you need reference-driven consistency, editing, or audio-led scenes. Both run from the same balance, so you can compare them on one prompt.

  Seedance 2.0 Seedance 1.5 Pro
Reference inputs Text, image, video, audio (9 img / 3 clip / 3 audio) Text, image
Native audio Yes Yes
Editing / extension Yes No
GPTProto price from $0.2957/run from $0.0408/run

Switching from Doubao, Jimeng, or Volcano Ark

If you already prototype Seedance 2.0 in ByteDance's own surfaces, GPTProto gives you the same model over one REST API: POST your prompt and parameters to the text-to-video endpoint, then poll /api/v3/predictions/{result_id}/result for the output. The parameter names (prompt, aspect_ratio, duration, resolution, generate_audio, camera_fixed, seed) map directly to what you already use, so porting a prompt is mostly copy-paste. The payoff beyond access is that the same key reaches the Fast variant, Seedance 1.5 Pro, and Sora 2, so you can A/B them on identical prompts from one balance instead of opening an account per provider.

What developers build with it

The features point at a few jobs it does well: short-form social and ad clips where native audio saves a separate dubbing step; character-consistent series and brand mascots that have to stay recognizable across shots; dialogue scenes that need lip-sync; and music- or beat-driven edits where sound and motion have to line up. For high-volume ideation, draft on the Fast variant and re-render keepers on the full model.

Doubao Seedance 2.0 prompt recipes

Seedance 2.0 rewards prompts that name sound and camera, not just visuals — because audio is generated jointly and camera language is a control signal. These are paste-and-edit starting points; tune duration and ratio to your shot.

1. Audio-led scene (use the joint audio pass)

A rainy Tokyo alley at night, neon reflections on wet asphalt. A figure in a dark coat walks toward camera. Diegetic sound: steady rain, distant traffic, footsteps in puddles. Slow dolly-in. 16:9.

Naming the sound lets the model sync ambience to the motion in one pass instead of you scoring it afterward. Keep generate_audio on.

2. Locked-off product shot (camera_fixed)

A ceramic coffee cup on a marble counter, morning light from the left, steam rising. Static camera, no movement. Shallow depth of field. 1:1.

Pair this with camera_fixed: true when you want the subject to move but the frame to hold — useful for e-commerce loops and packshots.

3. Character-consistent action

A young woman in a red jacket runs across a rooftop, jumps a gap, lands and keeps running. Keep her face, hair, and jacket identical across the whole clip. Handheld camera tracking from the side. Crowd and wind audio. 16:9, 8s.

Stating the consistency requirement in words leans on the model's strongest upgrade — holding a subject through movement.

4. Multi-shot in one generation

Shot 1: wide of a chef plating a dish. Shot 2: close-up of the garnish being placed. Shot 3: the chef looks up and smiles. Warm kitchen ambience, light sizzling. Smooth cuts. 16:9, 10s.

Numbering shots gives the model an explicit structure to follow, which holds better than one run-on description.

Prompting tips

  • Name diegetic sound explicitly when it matters; the audio is generated with the video.
  • Describe the camera move (dolly, pan, handheld, static) — it is treated as direction.
  • For recurring subjects, state the consistency requirement in words.
  • Draft on the Fast variant, then re-render the keeper on the full model.

How to Get a doubao-seedance-2-0-260128 API Key

Getting a doubao-seedance-2-0-260128 API key takes four steps and a few minutes. Create a free GPTProto account, add credits, generate your key, and make your first call — at $0.2957 it's a cheaper doubao-seedance-2-0-260128 API key than going direct, and one key works across every model on the platform. Full doubao-seedance-2-0-260128 Documentation is in the docs.

Sign up

Sign up

Create your free GPT Proto account to begin. You can set up an organization for your team at any time.

Top up

Top up

Your balance can be used across all models on the platform, including doubao-seedance-2-0-260128, giving you the flexibility to experiment and scale as needed.

Generate your API key

Generate your API key

In your dashboard, create an API key — you'll need it to authenticate when making requests to doubao-seedance-2-0-260128.

Make your first API call

Make your first API call

Use your API key with our sample code to send a request to doubao-seedance-2-0-260128 via GPT Proto and see instant AI-powered results.

Get API Key

FAQ

Everything you need to know about using Doubao Seedance 2.0 for professional video production.

What is Doubao Seedance 2.0?

ByteDance's second-generation video model — a unified audio-video architecture taking text, image, video, and audio, generating 4–15s clips up to 1080p with sound.

How does Doubao Seedance 2.0 API pricing work?

Per run, by resolution × duration. A 5-second clip is $0.39 at 480p, $0.83 at 720p, or $1.87 at 1080p; displayed from $0.2957/run.

Does Seedance 2.0 support image-to-video?

Yes, Seedance 2.0 supports image-to-video, but user feedback suggests it is currently most consistent in text-to-video mode. When using your own images, the model may occasionally prioritize motion over perfect image replication.

How much does the Doubao Seedance 2.0 API cost in practice?

A $100 balance buys roughly 258 five-second 480p drafts, 60 ten-second 720p scenes, or 26 ten-second 1080p shots.

Does Seedance 2.0 support image-to-video?

Yes — on a separate endpoint. Use the /image-to-video subpage, which adds an image field (a public URL) on top of the same parameters.

Does Seedance 2.0 generate audio?

Yes — audio is produced together with the video, not added in post.

How long can the videos be?

4 to 15 seconds per generation.

Is Seedance 2.0 open source?

No. It is a proprietary ByteDance model.

Doubao Seedance 2.0 vs Sora 2?

Seedance leads on multimodal reference inputs and 480p cost-per-second; Sora 2 Pro reaches 25s but its API is scheduled to sunset on 2026-09-24.

What is the best way to save credits on Seedance 2.0?

The best way to save on Seedance 2.0 is to avoid aggregators with complex pricing and stick to official rates. Always use low-resolution previews to test your prompts before doing a final high-resolution export.

Why is my Seedance 2.0 output inconsistent with my prompt?

Inconsistency in Seedance 2.0 often occurs when prompts are too vague or if you're trying to force the model to replicate a very specific static image. Focus on 'motion-first' prompting for the best results.

How do I troubleshoot API connection errors with Seedance 2.0?

If you experience inconsistencies with Seedance 2.0 API access, check your provider's status dashboard. Some platforms have removed access unexpectedly in the past, so using a stable aggregator like GPTProto is recommended.

Related Scenarios

Video Enhancer

Upscale resolution, sharpen details, and improve video quality instantly using our advanced video enhancer online tool.

Seedance 2.0

Integrate the Seedance 2.0 text-to-video model to render hyper-realistic action scenes using our developer-friendly AI API.

AI Video Editor

Integrate our AI video editor API to develop an automated AI tool. Empower creators with an advanced AI video editor powered by AI.

AI Video Generator

Scale video production with our AI video generator. This AI video maker and cinematic AI video creator provide realistic AI video outputs instantly.

Related Articles

More Blogs
Doubao 1.5 Pro: Benchmarks & API Access

Doubao 1.5 Pro: Benchmarks & API Access

ByteDance’s doubao 1.5 pro uses a leaner MoE architecture to rival heavier models with a fraction of compute. Read our guide to integrate it today.

Doubao AI: A Full Review of Features, Pros, Cons & Verdict

Doubao AI: A Full Review of Features, Pros, Cons & Verdict

Explore Doubao AI by ByteDance: Features multimodal capabilities, real-time answers, image generation & more. 50x cheaper than ChatGPT. Learn pricing, access options & how it compares to competitors.

Dreamina: Pro Guide to Seedance 2.0 Video AI

Dreamina: Pro Guide to Seedance 2.0 Video AI

Master dreamina Seedance 2.0 for realistic AI video. Learn about character consistency, physics, and VPN access tips. Optimize your workflow today.

Seedance AI: Bytedance's New Video Standard

Seedance AI: Bytedance's New Video Standard

Explore how Seedance by Bytedance is revolutionizing AI video with realistic facial expressions and low-cost API access. Learn more today.

GPT Proto

Empowering AI Innovation with Global Scale and Stability:

With our flagship product GPT Proto, we offer a unified interface to access and combine APIs from the world's leading AI providers—spanning text, vision, speech, and beyond. We empower developers and enterprises to simplify integration and accelerate innovation without limits.

Global Infrastructure, Local Compliance:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

Built to Scale:

We understand that stability is paramount. Our platform is built on a robust, decentralized architecture supporting dynamic Auto-scaling. Whether you are running a pilot or handling millions of concurrent requests, our system expands instantly to meet demand—guaranteeing that your business never outgrows our infrastructure.

Navigation

  • Dashboard
  • Models
  • Create Image
  • AI Image Upscale
  • AI Background Remover
  • Create Video
  • Edit in Canvas
  • Features
  • Pricing
  • AI Docs
  • AI Blog
  • AI Insight
  • AI Skills

Features

  • Anime to Real Life AI
  • Anime AI Art Generator
  • AI Object Remover
  • AI Image Editor
  • Unrestricted AI Image Generator
  • AI Motion Transfer
  • AI Clothes Remover
  • AI Watermark Remover
  • AI Image Enhancer Online
  • Online Background Remover Tool
  • AI Face Swap Image
  • AI Passport Photo Maker
  • MS Paint AI Generator
Explore all features >

LLM

  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
  • GPT 5.6 Sol
  • Grok 4.5
  • Claude Sonnet 5
  • Minimax M3
  • GLM 5.2
  • GPT 5.1 Chat Latest
  • Claude Fable 5
  • Qwen3.7 Max
  • Claude Opus 4.8 Thinking
  • Claude Opus 4.8
  • Gemini 3.5 Flash
  • DeepSeek v4 Flash
  • DeepSeek v4 Pro
  • Grok 4.3
Explore all models >

Image

  • Dola Seedream 5.0 Pro 260628
  • Gemini 3.1 Flash Lite Image
  • Gemini 3.1 Flash Image
  • GPT Image 2
  • Gemini 3.1 Flash Image Preview
  • Seedream 5.0 260128
  • Doubao Seedream 5.0 260128
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 251128
  • Doubao Seedream 4.5 251128
  • Grok Imagine 0.9
  • Gemini 3 Pro Image Preview
  • Qwen Image Lora
  • Qwen Image Plus Lora
  • Qwen Image Plus
  • Grok 4 Image
  • GPT Image 1 Mini
Explore all models >

Video

  • Kling v3.0 4k
  • Dreamina Seedance 2.0 Mini 260615
  • Kling v3 Omni 4k
  • Dreamina Seedance 2.0 Fast 260128
  • Dreamina Seedance 2.0 260128
  • Vidu 2.0
  • Doubao Seedance 2.0 260128
  • Doubao Seedance 2.0 Fast 260128
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Vidu Q3 Turbo
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
  • Vidu Q2 Pro Fast
  • Vidu Q2
  • Kling Video O1 Pro
Explore all models >

© 2026 Talent Tech Global Limited (Hong Kong) / Talent Tech Global LLC (US). All rights reserved.

  • About Us
  • Privacy Policy
  • Terms of Service
  • Sitemap