GPT Proto

GPTProto

  • 儀表板
  • LLM

    • claude
      Claude Opus 5新功能
    • google
      Gemini 3.6 Flash
    • google
      Gemini 3.5 Flash Lite
    • moonshotai
      Kimi K3
    • openai
      GPT 5.6 Luna

    影像

    • bytedance
      Dola Seedream 5.0 Pro 260628新功能
    • google
      Gemini 3.1 Flash Lite Image
    • google
      Gemini 3.1 Flash Image
    • openai
      GPT Image 2
    • google
      Gemini 3.1 Flash Image Preview

    影片

    • kling
      Kling v3.0 4k新功能
    • bytedance
      Dreamina Seedance 2.0 Mini 260615
    • kling
      Kling v3 Omni 4k
    • bytedance
      Dreamina Seedance 2.0 Fast 260128
    • bytedance
      Dreamina Seedance 2.0 260128
    探索 214+ 種模型 >
  • 生成器

    • 建立圖片
    • 建立影片
    • 在畫布中編輯

    功能

    • 動漫轉真人 AI新功能
    • 動漫 AI 藝術生成器
    • AI 物件移除器
    • AI 圖片編輯器
    • 無限制 AI 圖像生成器
    • AI 動作轉移
    • AI 衣物移除器
    • AI 浮水印移除工具
    • 線上 AI 圖片增強器
    • 線上背景移除工具
    探索全部 >

    提示詞

    • Seedance 2.0 提示詞新功能
    • GPT Image 2 提示詞
    • Nano Banana Pro 提示詞
    • Seedream 5.0 Pro 提示詞
  • AI 部落格

    • GLM 5.2 與 MiniMax M3:哪個更適合程式設計與前端工作?
    • 如何使用 API 建立自己的 AI 角色——無需撰寫程式碼
    • Kimi K3 與 Claude Opus 5:哪個更適合程式設計與 AI 代理?
    • 20 個免費 Seedream 5.0 Pro 產品與電子商務包裝設計提示詞
    • GLM-5.2 與 Kimi K3 程式設計比較:2026 年哪個更適合開發者?
    探索全部 >

    AI 洞察

    • 什麼是 Emochi AI?以及它為什麼成長得這麼快?(2026)
    • Kimi K3 是什麼?真的接近 GPT-5.6 與 Fable 5 嗎?
    • 2026 年 YouTube、TikTok、文字與圖片適用的 12 款最佳 AI 影片生成工具
    • 什麼是 Qwen 3.8 Max?發布日期、2.4T 預覽版、價格與早期基準測試
    • Gemini 3.6 Flash 與 Gemini 3.5 Flash-Lite 詳解:您應該使用哪一個?
    探索全部 >

    AI 文件

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    探索全部 >

    AI 技能

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    探索全部 >
定價
English繁體中文한국어日本語EspañolРусский
立即開始
  1. 首頁
  2. /模型
  3. /Bytedance
  4. /doubao-seedance-2-0-260128
Bytedance
doubao-seedance-2-0-260128
說明文件
說明文件
Doubao Seedance 2.0 是字節跳動的第二代影片模型,採用統一的音訊-影片架構,可在單次呼叫中接收文字、圖片、影片和音訊,並一次產生同步的影片與音訊。本頁介紹文字轉影片端點:撰寫提示詞,即可取得 4–15 秒、最高 1080p 且原生帶聲音的片段。在 GPTProto 上呼叫 doubao-seedance-2-0-260128,費用為 $0.2957/run — 共用一個餘額,提供與 OpenAI 相容的存取方式,無需 Jimeng/Volcano Ark 帳戶或地區性付款設定。

$ 0.2957

text

video

$ 0.2957

text

video

Playground
JSON
API

輸入

Your browser does not support the video tag.
您的請求將花費$0每次執行,對於$100您大約可以執行此模型0次
相關模型
所有模型
Kling
Kling
kling-v3.0-4k
$ 1.008
$ 1.26
Bytedance
Bytedance
dreamina-seedance-2-0-mini-260615
$ 0.2365
Vidu
Vidu
viduq3-turbo
$ 0.032
$ 0.04
Qwen
Qwen
wan-2.6
$ 0.45
$ 0.5
Google
Google
veo-3.1-fast-generate-preview
$ 1.2
MiniMax
MiniMax
hailuo-2.3-pro
$ 0.441
$ 0.49
範例
Ultra-realistic F1 live TV broadcast screenshot, identity preserved exactly from reference image.

Young woman sitting in the VIP paddock / team garage during a Formula 1 race, shown on the official live race broadcast as the girlfriend of an F1 driver. It is the final lap, and she is listening to the team radio through a professional racing headset, watching the garage monitors nervously, leaning forward with one hand near her mouth, proud tense expression.

She wears a fitted white tank top, oversized racing team jacket draped over her shoulders, large black team-radio headset with boom mic, gold jewelry, soft glam makeup. A slim paddock pass hangs naturally from her neck. 

Add realistic F1 broadcast graphics: “FINAL LAP” banner, lap counter showing final lap, driver timing tower on the left, small F1-style logo bug, “LIVE” indicator, lower-third identifying her as driver partner / paddock guest. No fake oversized badge, no selfie angle.

Team staff, headsets, garage screens, mechanics, and race equipment blurred around her. Telephoto broadcast camera from across the garage, compression artifacts, digital noise, bright paddock lighting, natural skin texture, no smoothing, 8k quality
Ultra-realistic sports broadcast still of a glamorous woman sitting in a packed football stadium crowd during a night match, wearing a dark brown sleeveless high-neck satin top and black square earrings, shoulder-length light brown/blonde hair styled in soft waves. She is casually drinking from a tall blue aluminum can while holding a half-eaten cheeseburger in the other hand. Around her are fans in bright yellow and blue football jerseys and scarves, creating strong team-color contrast. The scene feels candid and cinematic, captured mid-game from a TV broadcast camera angle with shallow depth of field. Include realistic stadium seating, crowded audience atmosphere, broadcast overlay graphics in the top-left corner showing a live football score and match timer, and a sports network watermark in the top-right. Natural arena lighting, detailed skin texture, sharp focus on the woman, slightly blurred background crowd, authentic live sports broadcast aesthetic, 16:9 composition.
A shimmering soap film stretches across a circular frame, catching the light. The secrets to achieving a perfectly spherical bubble are unveiled as the surface tension and air pressure work in harmony. This exploration reveals the simple yet elegant physics at play, creating fleeting moments of iridescent beauty.
ltra high-end editorial photography, photorealistic, 8K, razor-sharp detail, perfect skin texture with natural pores and subsurface scattering, cinematic lighting, shot on medium format, Portrait: 31-year-old Black British architect in a cream blazer, standing in a brutalist concrete stairwell, soft directional skylight, composed posture, masterful color grading, natural catchlights in eyes, tactile fabric detail, magazine cover quality, no text, no logos, no watermarks

Doubao Seedance 2.0 API: Text-to-Video with Native Audio from $0.2957/Run

The Doubao Seedance 2.0 API is ByteDance's second-generation video model, available on GPTProto with OpenAI-compatible access and no separate Jimeng or Volcano Ark account. Built on a unified audio-video architecture, doubao-seedance-2.0 takes text, image, video, and audio in a single call and returns a synchronized clip with native sound — 4 to 15 seconds, up to 4K, across aspect ratios from 16:9 and 9:16 to 21:9. It's the same Doubao AI Seedance 2.0 engine developers use for cinematic shots, product spots, and short social video, callable through one standard endpoint.

For doubao seedance 2.0 text to video, generation starts at $0.2957 per run, and doubao seedance 2.0 api pricing scales predictably with resolution and duration — see the full per-run table below. Whether you're prototyping fast 480p clips or rendering 1080p and 4K shots with audio, seedance 2.0 doubao gives you a pay-as-you-go path to ByteDance's latest video generation, with one API key that works across every model on the platform.

What is Doubao Seedance 2.0

Released in February 2026 by ByteDance's Seed team, Doubao Seedance 2.0 (also written doubao-seedance-2.0 or seedance 2.0 doubao) is built on a unified audio-video architecture. The change that matters from version 1.5 Pro is not a bigger resolution number — it is that the model treats every input as a control signal. A single generation can take a text brief plus reference images, a clip for camera motion, and an audio bed, and reason over all of them together before the first frame is produced. For this text-to-video endpoint you work from text, but the same model id is what powers the fuller reference workflow elsewhere.

Spec Value
Provider ByteDance (Seed team)
Model string doubao-seedance-2-0-260128
This endpoint Text-to-video
Input modalities (model) Text, image, video, audio
Reference capacity (model) Up to 9 images, 3 video clips, 3 audio clips per call
Native audio Yes — generated jointly with video; multilingual lip-sync
Duration 4–15s (model); 4–14s priced on GPTProto
Resolution Up to 1080p (480p / 720p / 1080p tiers)
Aspect ratios 16:9, 9:16, 4:3, 3:4, 21:9, 1:1
Editing / extension Targeted clip/character/action edits; clip extension
Image-to-video Yes — separate endpoint at /image-to-video
Variants Full + Fast (doubao-seedance-2-0-fast-260128 on GPTProto)
Open source No — proprietary

How the model handles audio and motion

Two things separate Seedance 2.0 from a standard text-to-video model, and both affect how you prompt it. First, audio and video come out of one pass rather than a silent clip with a soundtrack bolted on afterward. Because sound and image are generated together, dialogue lands on the mouth and ambient effects follow the action — so a prompt that names the sound ("crowd noise builds as the car passes") tends to produce tighter sync than describing visuals alone. Second, the model holds a subject's identity across movement and angle changes more reliably than earlier versions, which is what makes multi-shot and recurring-character work usable instead of one-off. If you have written off AI video because a face drifts between cuts, this is the axis that improved most.

Doubao Seedance 2.0 vs Sora 2

Both models are live on GPTProto, so this is a like-for-like choice rather than a marketing claim.

  Doubao Seedance 2.0 Sora 2
Inputs Text, image, video, audio Text, image
Native audio Yes, joint generation Yes, synced
Max duration 15s 12s (standard) / 25s (Pro)
Max resolution Up to 1080p 720p (standard) / 1080p (Pro)
API roadmap Active OpenAI is sunsetting the Sora 2 API on 2026-09-24
GPTProto price from $0.2957/run (scales with res × duration) $0.40/run (flat)

 

Doubao Seedance 2.0 vs Seedance 1.5 Pro

If you already run Seedance 1.5 Pro, the upgrade is about input control, not a resolution bump. Version 1.5 Pro already generated audio and video together and followed multi-shot instructions. Seedance 2.0 adds the reference stack on top — up to 9 images, 3 video clips, and 3 audio clips in a single call — plus targeted editing and clip extension. In third-party motion and physics testing it scores higher than 1.5 Pro, with the largest gain in physical accuracy.

The cost gap is large: 1.5 Pro starts at $0.0408/run versus $0.2957 for 2.0. Practical rule: stay on 1.5 Pro for plain, high-volume text-to-video where budget dominates; move to 2.0 when you need reference-driven consistency, editing, or audio-led scenes. Both run from the same balance, so you can compare them on one prompt.

  Seedance 2.0 Seedance 1.5 Pro
Reference inputs Text, image, video, audio (9 img / 3 clip / 3 audio) Text, image
Native audio Yes Yes
Editing / extension Yes No
GPTProto price from $0.2957/run from $0.0408/run

Switching from Doubao, Jimeng, or Volcano Ark

If you already prototype Seedance 2.0 in ByteDance's own surfaces, GPTProto gives you the same model over one REST API: POST your prompt and parameters to the text-to-video endpoint, then poll /api/v3/predictions/{result_id}/result for the output. The parameter names (prompt, aspect_ratio, duration, resolution, generate_audio, camera_fixed, seed) map directly to what you already use, so porting a prompt is mostly copy-paste. The payoff beyond access is that the same key reaches the Fast variant, Seedance 1.5 Pro, and Sora 2, so you can A/B them on identical prompts from one balance instead of opening an account per provider.

What developers build with it

The features point at a few jobs it does well: short-form social and ad clips where native audio saves a separate dubbing step; character-consistent series and brand mascots that have to stay recognizable across shots; dialogue scenes that need lip-sync; and music- or beat-driven edits where sound and motion have to line up. For high-volume ideation, draft on the Fast variant and re-render keepers on the full model.

Doubao Seedance 2.0 prompt recipes

Seedance 2.0 rewards prompts that name sound and camera, not just visuals — because audio is generated jointly and camera language is a control signal. These are paste-and-edit starting points; tune duration and ratio to your shot.

1. Audio-led scene (use the joint audio pass)

A rainy Tokyo alley at night, neon reflections on wet asphalt. A figure in a dark coat walks toward camera. Diegetic sound: steady rain, distant traffic, footsteps in puddles. Slow dolly-in. 16:9.

Naming the sound lets the model sync ambience to the motion in one pass instead of you scoring it afterward. Keep generate_audio on.

2. Locked-off product shot (camera_fixed)

A ceramic coffee cup on a marble counter, morning light from the left, steam rising. Static camera, no movement. Shallow depth of field. 1:1.

Pair this with camera_fixed: true when you want the subject to move but the frame to hold — useful for e-commerce loops and packshots.

3. Character-consistent action

A young woman in a red jacket runs across a rooftop, jumps a gap, lands and keeps running. Keep her face, hair, and jacket identical across the whole clip. Handheld camera tracking from the side. Crowd and wind audio. 16:9, 8s.

Stating the consistency requirement in words leans on the model's strongest upgrade — holding a subject through movement.

4. Multi-shot in one generation

Shot 1: wide of a chef plating a dish. Shot 2: close-up of the garnish being placed. Shot 3: the chef looks up and smiles. Warm kitchen ambience, light sizzling. Smooth cuts. 16:9, 10s.

Numbering shots gives the model an explicit structure to follow, which holds better than one run-on description.

Prompting tips

  • Name diegetic sound explicitly when it matters; the audio is generated with the video.
  • Describe the camera move (dolly, pan, handheld, static) — it is treated as direction.
  • For recurring subjects, state the consistency requirement in words.
  • Draft on the Fast variant, then re-render the keeper on the full model.

如何取得 doubao-seedance-2-0-260128 API 金鑰

取得 doubao-seedance-2-0-260128 API 金鑰只需四個步驟,僅需幾分鐘。建立免費的 GPTProto 帳號、儲值、產生金鑰,並進行第一次呼叫 — 在 $0.2957 這裡能以比直連更划算的價格取得 doubao-seedance-2-0-260128 API 金鑰,且單一金鑰即可在平台上所有模型通用。完整 doubao-seedance-2-0-260128 說明文件 請參閱說明文件。

註冊

註冊

建立您的免費 GPT Proto 帳號即可開始。您可以隨時為您的團隊設定組織。

儲值

儲值

您的餘額可用於平台上的所有模型,包括 doubao-seedance-2-0-260128,讓您能靈活地進行實驗並隨需求擴充。

產生您的 API 金鑰

產生您的 API 金鑰

在您的儀表板中建立 API 金鑰 — 進行 doubao-seedance-2-0-260128 請求時,您將需要它來進行驗證。

進行第一次 API 呼叫

進行第一次 API 呼叫

使用您的 API 金鑰搭配我們的範例程式碼,透過 GPT Proto 向 doubao-seedance-2-0-260128 發送請求,並立即查看 AI 生成的結果。

取得 API 金鑰

常見問題

了解使用 Doubao Seedance 2.0 進行專業影片製作所需的一切資訊。

什麼是 Doubao Seedance 2.0?

ByteDance 的第二代影片模型——採用統一的音訊-影片架構,可接收文字、圖片、影片和音訊,生成 4–15 秒、最高 1080p 且具備聲音的片段。

Doubao Seedance 2.0 API 的定價方式為何?

每次執行的費用取決於解析度與時長。5 秒片段在 480p、720p 或 1080p 下的費用分別為 0.39 美元、0.83 美元或 1.87 美元;顯示價格為每次執行 0.2957 美元起。

Seedance 2.0 支援圖片轉影片嗎?

是,Seedance 2.0 支援圖片轉影片,但使用者回饋顯示,目前在文字轉影片模式下的表現最穩定。使用自己的圖片時,模型偶爾可能會優先考慮動態效果,而無法完美重現圖片。

實際使用 Doubao Seedance 2.0 API 的費用是多少?

100 美元的餘額大約可購買 258 個 5 秒的 480p 草稿、60 個 10 秒的 720p 場景,或 26 個 10 秒的 1080p 鏡頭。

Seedance 2.0 支援圖片轉影片嗎?

是——但需要使用不同的端點。請使用 /image-to-video 子頁面,該頁面會在相同參數的基礎上新增圖片欄位(公開 URL)。

Seedance 2.0 會生成音訊嗎?

是——音訊會與影片一同生成,而不是在後製階段新增。

影片最長可以有多長?

每次生成的影片長度為 4 至 15 秒。

Seedance 2.0 是開源的嗎?

不可以。這是 ByteDance 的專有模型。

Doubao Seedance 2.0 與 Sora 2 相比如何?

Seedance 在多模態參考輸入和 480p 每秒成本方面具有優勢;Sora 2 Pro 可生成最長 25 秒的影片,但其 API 預計將於 2026-09-24 停止服務。

節省 Seedance 2.0 點數的最佳方式是什麼?

節省 Seedance 2.0 費用的最佳方式,是避免使用定價複雜的聚合服務,並遵循官方費率。進行最終的高解析度匯出前,請務必使用低解析度預覽來測試提示詞。

為什麼我的 Seedance 2.0 輸出結果與提示詞不一致?

當提示詞過於模糊,或試圖強迫模型重現非常特定的靜態圖片時,Seedance 2.0 往往會出現不一致的結果。請專注於使用「以動態為先」的提示詞,以獲得最佳效果。

如何排查 Seedance 2.0 的 API 連線錯誤?

如果你在使用 Seedance 2.0 API 時遇到不一致的情況,請查看服務供應商的狀態儀表板。過去有些平台曾在未預先通知的情況下移除存取權限,因此建議使用 GPTProto 這類穩定的聚合服務。

相關場景

影片增強工具

使用我們先進的線上影片增強工具,立即提升解析度、銳化細節並改善影片品質。

Seedance 2.0

整合 Seedance 2.0 文字轉影片模型,透過我們對開發者友善的 AI API,生成超逼真的動作場景。

AI 影片編輯器

整合我們的 AI 影片編輯器 API,打造自動化 AI 工具。透過 AI 驅動的先進 AI 影片編輯器,賦能創作者。

AI 影片生成器

使用我們的 AI 影片生成器擴大影片製作規模。這款 AI 影片製作工具與電影級 AI 影片創作工具能立即提供逼真的 AI 影片輸出。

相關文章

更多部落格
Doubao 1.5 Pro:基準測試與 API 存取

Doubao 1.5 Pro:基準測試與 API 存取

ByteDance 的 doubao 1.5 pro 採用更精簡的 MoE 架構,以少量運算資源與更大型的模型競爭。閱讀我們的指南,立即完成整合。

Doubao AI:功能、優缺點與評價完整評測

Doubao AI:功能、優缺點與評價完整評測

探索 ByteDance 推出的 Doubao AI:具備多模態功能、即時回答、圖像生成等更多能力。價格比 ChatGPT 低 50 倍。了解定價、存取選項,以及它與競爭對手的比較。

Dreamina:Seedance 2.0 影片 AI 專業指南

Dreamina:Seedance 2.0 影片 AI 專業指南

掌握 dreamina Seedance 2.0,製作逼真的 AI 影片。了解角色一致性、物理效果及 VPN 存取技巧,立即最佳化您的工作流程。

Seedance AI:Bytedance 的全新影片標準

Seedance AI:Bytedance 的全新影片標準

探索 Bytedance 推出的 Seedance 如何以逼真的面部表情和低成本 API 存取,徹底革新 AI 影片。立即深入了解。

GPT Proto

以全球規模與穩定性,賦能 AI 創新:

透過我們的旗艦產品 GPT Proto,我們提供統一的介面,讓您能存取並整合全球頂尖 AI 供應商的 API,涵蓋文字、視覺、語音等領域。我們協助開發者與企業簡化整合流程,並無限制地加速創新。

全球基礎設施,在地合規:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

為擴展而生:

我們深知穩定性至關重要。我們的平台建立在強大的去中心化架構之上,支援動態自動擴展。無論您是進行試點專案還是處理數百萬次併發請求,我們的系統都能即時擴展以滿足需求,確保您的業務永遠不會受限於基礎設施。

導覽

  • 儀表板
  • 模型
  • 建立圖片
  • AI 圖片放大
  • AI 背景移除
  • 建立影片
  • 在畫布中編輯
  • 功能
  • 定價
  • AI 文件
  • AI 部落格
  • AI 洞察
  • AI 技能

功能

  • 動漫轉真人 AI
  • 動漫 AI 藝術生成器
  • AI 物件移除器
  • AI 圖片編輯器
  • 無限制 AI 圖像生成器
  • AI 動作轉移
  • AI 衣物移除器
  • AI 浮水印移除工具
  • 線上 AI 圖片增強器
  • 線上背景移除工具
  • AI 臉部交換圖片
  • AI 護照照片製作器
  • MS Paint AI 生成器
Explore all features >

LLM

  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
  • GPT 5.6 Sol
  • Grok 4.5
  • Claude Sonnet 5
  • Minimax M3
  • GLM 5.2
  • GPT 5.1 Chat Latest
  • Claude Fable 5
  • Qwen3.7 Max
  • Claude Opus 4.8 Thinking
  • Claude Opus 4.8
  • Gemini 3.5 Flash
  • DeepSeek v4 Flash
  • DeepSeek v4 Pro
  • Grok 4.3
探索所有模型 >

影像

  • Dola Seedream 5.0 Pro 260628
  • Gemini 3.1 Flash Lite Image
  • Gemini 3.1 Flash Image
  • GPT Image 2
  • Gemini 3.1 Flash Image Preview
  • Seedream 5.0 260128
  • Doubao Seedream 5.0 260128
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 251128
  • Doubao Seedream 4.5 251128
  • Grok Imagine 0.9
  • Gemini 3 Pro Image Preview
  • Qwen Image Lora
  • Qwen Image Plus Lora
  • Qwen Image Plus
  • Grok 4 Image
  • GPT Image 1 Mini
探索所有模型 >

影片

  • Kling v3.0 4k
  • Dreamina Seedance 2.0 Mini 260615
  • Kling v3 Omni 4k
  • Dreamina Seedance 2.0 Fast 260128
  • Dreamina Seedance 2.0 260128
  • Vidu 2.0
  • Doubao Seedance 2.0 260128
  • Doubao Seedance 2.0 Fast 260128
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Vidu Q3 Turbo
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
  • Vidu Q2 Pro Fast
  • Vidu Q2
  • Kling Video O1 Pro
探索所有模型 >

© 2026 Talent Tech Global Limited (Hong Kong) / Talent Tech Global LLC (US). 保留所有權利。

  • 關於我們
  • 隱私權政策
  • 服務條款
  • 網站地圖