GPT Proto

GPTProto

  • 儀表板
  • LLM

    • deepseek
      DeepSeek Flash新功能
    • z-ai
      GLM 5.3
    • claude
      Claude Fable 5
    • deepseek
      DeepSeek v4 Pro
    • google
      Gemini 3.7 Flash
    • grok
      Grok 4.6
    探索模型 >

    影像

    • openai
      GPT Image 2.5 Sunburst新功能
    • openai
      GPT Image 2
    • google
      Nano Banana Pro (Gemini 3 Pro Image)
    • google
      Nano Banana 2 (Gemini 3.1 Flash Image)
    • midjourney
      Midjourney
    • openai
      GPT Image 2.5 Flare
    探索模型 >

    影片

    • minimax
      Minimax H3新功能
    • bytedance
      Seedance 2.5 (Build 260628)
    • bytedance
      Seedance 2.0 (Build 260128)
    • bytedance
      Seedance 2.0 Mini (Build 260615)
    • kling
      Kling v3.0 4k
    • vidu
      Vidu Q3 Turbo
    探索模型 >
    探索 232+ 種模型 >
  • 生成器

    • 建立圖片
    • 建立影片
    • 在畫布中編輯
    • 聊天

    功能

    • AI 貓跳舞影片生成器新功能
    • 無限制 AI 影片生成器
    • AI 包裝設計生成器
    • 動漫 AI 藝術生成器
    • AI 物件移除器
    • AI 圖片編輯器
    • AI 動作轉移
    • AI 浮水印移除工具
    • 線上 AI 圖片增強器
    • 線上背景移除工具
    探索全部 >

    提示詞

    • Seedance 2.0 提示詞新功能
    • GPT Image 2 提示詞
    • Nano Banana Pro 提示詞
    • Seedream 5.0 Pro 提示詞
    • Midjourney 提示詞
  • AI 部落格

    • OpenRouter 與 GPTProto:定價、模型、路由,以及 2026 年哪個 API 更好?
    • AI 產品廣告工作流程:從洗衣精圖片到 25 秒廣告片
    • DeepSeek V4 Pro 與 GLM 5.2:2026 年哪個更好?
    • 2026 年 7 款最佳影像編輯 AI 模型:API、批次編輯與產品照片
    • DeepSeek V4 Pro 與 Kimi K3:0813 更新後有何變化?
    探索全部 >

    AI 洞察

    • DeepSeek 尖峰定價已上線:API 何時費用更高?
    • 什麼是 GLM-5.3?Z.ai 悄然推出的程式設計方案、價格與已確認的升級
    • 什麼是 OpenAI 最新的 Astra 模型?發布日期、基準測試與比較(2026)
    • MiniMax H3 正式登場:其影片編輯升級實際帶來哪些改變
    • 什麼是 Emochi AI?以及它為什麼成長得這麼快?(2026)
    探索全部 >

    AI 文件

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    探索全部 >

    AI 技能

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    探索全部 >
定價+7% 贈送
English繁體中文한국어日本語EspañolРусский
立即開始
  1. 首頁
  2. /模型
  3. /MiniMax
  4. /minimax-h3
MiniMax
Minimax H3
$ 
MiniMax H3 is an open-weight multimodal video model for text-to-video, first- and last-frame image-to-video, reference-led generation, and video editing. It accepts text, images, video, and audio, then produces 4–15-second clips at 768P or 2K with native stereo sound.

模態

輸入: 文字輸入: 圖像
輸出: 影片

API 呼叫範例
$ 
curl --request POST "https://gptproto.com/api/v3/minimax/minimax-h3/text-to-video" \
  --header "Authorization: Bearer $GPTPROTO_API_KEY" \
  --header "Content-Type: application/json" \
  --data '{
    "prompt": "A tiny origami fox sailing a teacup across a moonlit puddle",
    "aspect_ratio": "16:9",
    "duration": 5,
    "resolution": "768P"
  }'
Minimax H3 定價

先從單次樣本成本開始,再選擇測試預算。GPTProto 費率比標價低 30%。

你的用量

$0.455 / 張

相關模型
所有模型
Minimax H3
目前
$ 
byMiniMax$0.224 / 每次2K · more
Wan 3.0
$ 
byQwen$0.09 / 每次more
Seedance 2.5 (Build 260628)
$ 
byBytedance$0.113 / 每秒480p · 720p
Kling v3.0 4k
$ 
byKling$0.336 / 每秒4k
Seedance 2.0 Mini (Build 260615)
$ 
byBytedance$0.0237 / 每秒480p · 720p
Kling v3 Omni 4k
$ 
byKling$0.336 / 每秒4k
Seedance 2.0 Fast (Build 260128)
$ 
byBytedance$0.0443 / 每秒480p · 720p
Vidu 2.0
$ 
byVidu$0.02 / 每秒720p · 1080p · more
Kling v3 Omni Pro
$ 
byKling$0.0896 / 每秒
Vidu Q3 Turbo
$ 
byVidu$0.032 / 每秒540p · 720p · 1080p
Vidu Q3 Pro
$ 
byVidu$0.04 / 每秒540p · 720p · 1080p
Wan 2.6
$ 
byQwen$0.09 / 每秒720p · 1080p · 2k
Video Watermark Remover
$ 
byGPTProto$0.05 / 每次
Veo 3.1 Fast Generate Preview
$ 
byGoogle$1.2 / 每次
Veo 3.1 Generate Preview
$ 
byGoogle$3.2 / 每次
Hailuo 2.3 Fast
$ 
byMiniMax$0.0285 / 每秒1080p · more
Hailuo 2.3 Pro
$ 
byMiniMax$0.441 / 每次
Hailuo 2.3 Standard
$ 
byMiniMax$0.042 / 每秒
Hailuo 02 Standard
$ 
byMiniMax$0.042 / 每秒
Hailuo 02 Pro
$ 
byMiniMax$0.441 / 每次
Hailuo 02 Fast
$ 
byMiniMax$0.0135 / 每秒
Wan 2.2 Plus
$ 
byQwen$0.09 / 每次1K · more
Veo 3.1
$ 
byGoogle$0.5 / 每次
Sora 2 Pro
$ 
byOpenAI$0.3 / 每秒720p · 1080p
Sora 2
$ 
byOpenAI$0.1 / 每秒720p
Higgsfield Turbo
$ 
byHiggsfield$0.2842 / 每次
Higgsfield Lite
$ 
byHiggsfield$0.0875 / 每次
Higgsfield Standard
$ 
byHiggsfield$0.3941 / 每次
模型解析度輸入 → 輸出
Minimax H3目前
$ 
$0.22 / 每次2K · more
輸入: 文字輸入: 圖像
輸出: 影片
Wan 3.0
$ 
$0.09 / 每次more
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Seedance 2.5 (Build 260628)
$ 
—$0.11 / 每秒480p · 720p
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Kling v3.0 4k
$ 
$0.34 / 每秒4k
輸入: 文字輸入: 圖像
輸出: 影片
Seedance 2.0 Mini (Build 260615)
$ 
—$0.02 / 每秒480p · 720p
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Kling v3 Omni 4k
$ 
$0.34 / 每秒4k
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Seedance 2.0 Fast (Build 260128)
$ 
—$0.04 / 每秒480p · 720p
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Vidu 2.0
$ 
$0.02 / 每秒720p · 1080p · more
輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Kling v3 Omni Pro
$ 
$0.09 / 每秒—
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Vidu Q3 Turbo
$ 
$0.03 / 每秒540p · 720p · 1080p
輸入: 文字輸入: 圖像
輸出: 影片
Vidu Q3 Pro
$ 
$0.04 / 每秒540p · 720p · 1080p
輸入: 文字輸入: 圖像
輸出: 影片
Wan 2.6
$ 
$0.09 / 每秒720p · 1080p · 2k
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Video Watermark Remover
$ 
—$0.05 / 每次—
輸入: 影片
輸出: 影片
Veo 3.1 Fast Generate Preview
$ 
—$1.20 / 每次—
輸入: 文字輸入: 圖像輸入: 影片
輸出: 影片
Veo 3.1 Generate Preview
$ 
—$3.20 / 每次—
輸入: 文字輸入: 圖像輸入: 影片
輸出: 影片
Hailuo 2.3 Fast
$ 
$0.03 / 每秒1080p · more
輸入: 圖像
輸出: 影片
Hailuo 2.3 Pro
$ 
$0.44 / 每次—
輸入: 文字輸入: 圖像
輸出: 影片
Hailuo 2.3 Standard
$ 
$0.04 / 每秒—
輸入: 文字輸入: 圖像
輸出: 影片
Hailuo 02 Standard
$ 
$0.04 / 每秒—
輸入: 文字輸入: 圖像
輸出: 影片
Hailuo 02 Pro
$ 
$0.44 / 每次—
輸入: 文字輸入: 圖像
輸出: 影片
Hailuo 02 Fast
$ 
$0.01 / 每秒—
輸入: 圖像
輸出: 影片
Wan 2.2 Plus
$ 
$0.09 / 每次1K · more
輸入: 文字輸入: 圖像
輸出: 影片
Veo 3.1
$ 
—$0.50 / 每次—
輸入: 文字輸入: 圖像輸入: 影片輸入: 音訊
輸出: 影片
Sora 2 Pro
$ 
—$0.30 / 每秒720p · 1080p
輸入: 文字輸入: 圖像
輸出: 影片
Sora 2
$ 
—$0.10 / 每秒720p
輸入: 文字輸入: 圖像
輸出: 影片
Higgsfield Turbo
$ 
$0.28 / 每次—
輸入: 圖像
輸出: 影片
Higgsfield Lite
$ 
$0.09 / 每次—
輸入: 圖像
輸出: 影片
Higgsfield Standard
$ 
$0.39 / 每次—
輸入: 圖像
輸出: 影片

MiniMax-H3 API for 2K Multimodal Video Generation

Access the MiniMax-H3 API through GPTProto for cinematic video generation from prompts, keyframes, and mixed reference assets. Build 4–15-second clips with native stereo audio, use up to nine reference images, three video clips, and three audio tracks, and manage MiniMax H3 alongside 200+ AI models with one API key and one balance.

Text, Image, Video, and Audio in One Context

Combine a prompt with image, video, and audio references. MiniMax H3 can use those assets to guide subject identity, motion, camera behavior, visual style, voice, and editing rhythm.

4–15 Seconds at 768P or 2K

Choose an integer duration from 4 to 15 seconds. Generate at 768P for drafts or request 2K output when sharper product detail, typography, or final-delivery quality matters.

What Is MiniMax H3?

MiniMax H3 is MiniMax's general-purpose, open-weight video generation system. Instead of treating text-to-video, keyframe animation, and multimodal references as unrelated tools, it interprets text, images, video, and audio within one context. The hosted API uses an asynchronous workflow: submit a task, store its task ID, check its status, and retrieve the finished video URL.

Output runs from 4 to 15 seconds at 24 fps, with 768P and 2K modes, landscape, square, and portrait ratios, and native stereo audio. The released H3-Base weights use the MiniMax H3 Community License; hosted API access and self-hosting remain separate deployment choices.

Specification MiniMax H3 API Details
Provider MiniMax
Official API model name MiniMax-H3
GPTProto model string MiniMax-H3
Generation modes Text-to-video, first-frame I2V, last-frame I2V, first-and-last-frame I2V, multimodal reference-to-video
Input types Text, images, video, and audio
Output duration 4–15 seconds, integer values
Output resolution 768P or 2K
Frame rate 24 fps
Output audio Native 32 kHz stereo
Aspect ratios 21:9, 16:9, 4:3, 1:1, 3:4, 9:16; adaptive where supported
Reference limits Up to 9 images, 3 video clips, and 3 audio clips; up to 12 files combined
Prompt limit Up to 7,000 characters in the official API
Processing Asynchronous task workflow

Which MiniMax H3 Input Mode Should You Use?

Choose the input mode before building the request. First/last-frame inputs and multimodal reference inputs are separate modes in the official API and cannot be mixed in one generation task.

Input Mode Use It When Practical Example
Text-to-video You need a new scene without an approved visual starting point Generate several cinematic concepts from one campaign brief
First-frame image-to-video The opening composition or product image is already approved Animate a product hero image while retaining its initial framing
Last-frame image-to-video The shot must resolve to a specific final composition End on a logo lockup, product pack, or call-to-action frame
First-and-last-frame video Both endpoints of a transition matter Move from a closed package to the product fully revealed
Reference-to-video Identity, motion, voice, music, setting, or style must come from source assets Keep the same product and spokesperson while borrowing motion from a reference clip

For reference-to-video, MiniMax H3 accepts up to nine images, three video clips, and three audio clips, with a maximum of 12 files in total. Each reference video or audio clip can be 2–15 seconds, and the total duration for each media type cannot exceed 15 seconds. Use public URLs for large assets because the complete request body is limited to 64 MB.

MiniMax H3 API for Ecommerce and Batch Video Workflows

For ecommerce, supply approved product images, state which details must remain unchanged, and add a reference video only when its motion or camera path matters. First-and-last-frame mode is often simpler for controlled product reveals.

For batch generation, treat every variation as a separate asynchronous job. Store its prompt, asset URLs, settings, task ID, and status; use a bounded queue; make retries idempotent; and copy successful outputs to your own storage. Keep logo treatment, framing, lighting, camera language, and CTA fixed while injecting variable SKU data per job.

MiniMax H3 vs Seedance 2.5 vs Veo 3.1

These models overlap but solve different constraints. MiniMax H3 emphasizes 2K output, open weights, stereo audio, and mixed references. Seedance 2.5 supports longer generations and larger reference sets. Veo 3.1 fits short Google Cloud workflows with a documented 4K path.

Decision Factor MiniMax H3 Seedance 2.5 Veo 3.1 Generate
Single-generation duration 4–15 seconds 4–30 seconds 4, 6, or 8 seconds
Documented output resolution 768P or 2K 480P, 720P, or 1080P 720P, 1080P, or 4K
Input modalities Text, image, video, audio Text, image, video, audio Text and image; no audio or video input in the documented Generate endpoint
Reference capacity Up to 9 images, 3 videos, 3 audio clips; 12 files total Up to 30 images, 10 videos, and 10 audio clips Up to 3 asset images
First/last-frame control Yes Yes Yes
Native generated audio Yes Yes Yes
Open weights H3-Base released under a community license No published open weights No
Best fit 2K multimodal reference work, motion transfer, editing, and self-hosting research Longer scenes, large reference packs, and timeline-led editing Short high-fidelity clips in Google Cloud and 4K delivery workflows

Compare duration, reference capacity, audio, retries, and finished-clip cost—not maximum resolution alone. Seedance 2.5 may reduce scene stitching; H3 may simplify mixed-reference 2K work; Veo 3.1 may suit short Google-centered 4K workflows.

How to Write Better MiniMax H3 Prompts

Write the prompt as a compact production brief that separates source roles, timeline, camera direction, audio, preserved details, and exclusions.

Prompt formula: subject and reference roles + scene goal + timed actions + camera movement + lighting and visual style + dialogue, sound effects, and music + details to preserve + final frame

Ecommerce Product Video Prompt

Use Image 1 for product identity and Image 2 for lighting. Create a 10-second vertical launch video for the matte-black speaker. Begin with a macro grille shot, pull back as it rotates, then show water droplets vibrating with the bass. Preserve the logo, controls, proportions, and finish. Use cool edge light, restrained motion, low electronic ambience, and no extra text.

First-and-Last-Frame Transition Prompt

Start from the supplied closed package and end exactly on the supplied final frame with the product assembled. Use one continuous camera move: the box opens, components rise, and the product locks into place. Preserve package graphics and final geometry. Add mechanical clicks and a short resolved tone.

MiniMax H3 API FAQ

What is the MiniMax H3 API?

It provides hosted access to MiniMax's open-weight multimodal video model for 4–15-second video at 768P or 2K with native stereo sound.

Does MiniMax H3 support both text-to-video and image-to-video?

Yes. It supports prompt-only generation, first- or last-frame I2V, two-keyframe transitions, and a separate multimodal reference mode.

What duration and resolution does MiniMax H3 support?

The official API accepts integer durations from 4 to 15 seconds at 24 fps, with 768P and 2K output modes.

Does MiniMax H3 generate audio with the video?

Yes. It generates native 32 kHz stereo sound. Prompts can direct dialogue, ambience, effects, and music; reference audio can guide voice or rhythm.

How many reference assets can I use?

Up to nine images, three video clips, and three audio clips, with 12 files total. Video and audio references are each limited to 15 seconds in aggregate.

Can I use MiniMax H3 for batch ecommerce video generation?

Yes, through application-level orchestration: submit one task per variation, store task IDs, limit concurrency, make retries idempotent, and save completed files to your storage.

MiniMax H3 or Seedance 2.5: which is better?

Choose H3 for 2K, open H3-Base weights, or compact mixed-reference jobs. Choose Seedance 2.5 for clips up to 30 seconds, larger reference packs, or timeline-led editing.

MiniMax H3 vs Veo 3.1: which is more cost-effective?

Compare the live rate for the same duration and resolution, then include retries and usable-output rate. H3 offers straightforward 768P/2K cost planning; Veo 3.1 may be preferable when Google Cloud integration or 4K delivery matters.

How do I access the MiniMax H3 API on GPTProto?

Create a GPTProto key and select H3 in the playground or documentation. Use the model string shown in the live API example: MiniMax-H3.

Is MiniMax H3 the same as Hailuo 3?

MiniMax H3 is the official name. “Hailuo 3” and “MiniMax Hailuo H3” are search aliases tied to MiniMax's Hailuo product; API docs identify the model as MiniMax-H3.

相關文章

與本模型相關的指南、對比與更新。

所有文章
5 Best Affordable AI Video APIs in 2026: Pricing, Ecommerce, and Short Drama

5 Best Affordable AI Video APIs in 2026: Pricing, Ecommerce, and Short Drama

Compare 5 affordable AI video APIs for ecommerce and AI short drama. See current pricing, clip costs, audio fees, and the best model for each job.

6 Cheapest AI Image Generators in 2026: Real Cost per Image

6 Cheapest AI Image Generators in 2026: Real Cost per Image

Compare 6 cheapest AI image generators in 2026, from about $0.0035 per image. See batch costs, hidden fees, and the best API for startups.

How to Make an AI Story Video for Kids with GPT Image 2 and Seedance 2.5

How to Make an AI Story Video for Kids with GPT Image 2 and Seedance 2.5

Learn how to make an AI story video for kids with GPT Image 2 and Seedance 2.5, including the full prompt, captions, sound, editing, and a real test.

Seedance 2.0 vs Seedance 2.5: Same Prompt, Storyboard, and Real Results

Seedance 2.0 vs Seedance 2.5: Same Prompt, Storyboard, and Real Results

See Seedance 2.0 vs 2.5 in the same 15-second storyboard test. Compare cinematic quality, emotion, pricing, ecommerce use cases, and API features.

GPT Proto

以全球規模與穩定性,賦能 AI 創新:

透過我們的旗艦產品 GPT Proto,我們提供統一的介面,讓您能存取並整合全球頂尖 AI 供應商的 API,涵蓋文字、視覺、語音等領域。我們協助開發者與企業簡化整合流程,並無限制地加速創新。

全球基礎設施,在地合規:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

為擴展而生:

我們深知穩定性至關重要。我們的平台建立在強大的去中心化架構之上,支援動態自動擴展。無論您是進行試點專案還是處理數百萬次併發請求,我們的系統都能即時擴展以滿足需求,確保您的業務永遠不會受限於基礎設施。

導覽

  • 儀表板
  • 模型
  • 建立圖片
  • AI 圖片放大
  • AI 背景移除
  • 建立影片
  • 在畫布中編輯
  • 聊天
  • 功能
  • 定價
  • AI 文件
  • AI 部落格
  • AI 洞察
  • AI 技能

功能

  • AI 貓跳舞影片生成器
  • 無限制 AI 影片生成器
  • AI 包裝設計生成器
  • 動漫 AI 藝術生成器
  • AI 物件移除器
  • AI 圖片編輯器
  • AI 動作轉移
  • AI 浮水印移除工具
  • 線上 AI 圖片增強器
  • 線上背景移除工具
  • AI 臉部交換圖片
  • AI 護照照片製作器
  • MS Paint AI 生成器
  • AI 衣物移除器
  • 無限制AI圖片生成器
  • AI 法式接吻生成器
  • AI 電影海報產生器
  • Artlist IO 工作室
  • 線上魔術橡皮擦
  • Luma Dream Machine
Explore all features >

LLM

  • DeepSeek Flash
  • GLM 5.3
  • Claude Fable 5
  • DeepSeek v4 Pro
  • Gemini 3.7 Flash
  • Grok 4.6
  • Hy4 Preview
  • GPT 6 Astra
  • Gemini 3.8 Flash
  • Claude Fable 5.1
  • Qwen3.8 Max 0902
  • GLM 5.3 Flash
  • DeepSeek v4 Flash Vision Exp
  • Qwen3.8 Max
  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
探索所有模型 >

影像

  • GPT Image 2.5 Sunburst
  • GPT Image 2
  • Nano Banana Pro (Gemini 3 Pro Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Midjourney
  • GPT Image 2.5 Flare
  • Grok Imagine Image 2.0
  • Seedream 5.0 Pro (Build 260628)
  • Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Seedream 5.0 (Build 260128)
  • Doubao Seedream 5.0 (Build 260128)
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 (Build 251128)
  • Doubao Seedream 4.5 (Build 251128)
  • Grok Imagine 0.9
  • Qwen Image Lora
探索所有模型 >

影片

  • Minimax H3
  • Seedance 2.5 (Build 260628)
  • Seedance 2.0 (Build 260128)
  • Seedance 2.0 Mini (Build 260615)
  • Kling v3.0 4k
  • Vidu Q3 Turbo
  • Wan 3.0
  • Kling v3 Omni 4k
  • Seedance 2.0 Fast (Build 260128)
  • Vidu 2.0
  • Doubao Seedance 2.0 (Build 260128)
  • Doubao Seedance 2.0 Fast (Build 260128)
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
探索所有模型 >

聯絡我們

有任何問題或回饋嗎?歡迎透過以下管道與我們聯繫。

TelegramWhatsApp

© 2026 Talent Tech Global Limited (Hong Kong). 保留所有權利。

註冊地址: Unit 1022a, Beverley Commercial Centre, 87-105 Chatham Road South, Tsim Sha Tsui, Hong Kong商業登記證號碼: 79462435-000-12-25-0
  • 關於我們
  • 隱私權政策
  • 服務條款
  • 網站地圖
友情連結logoto.videotopostudio.cc

輸入

輸出

Your browser does not support the video tag.