GPT Proto

GPTProto

  • Dashboard
  • LLM

    • deepseek
      DeepSeek FlashNew
    • z-ai
      GLM 5.3
    • claude
      Claude Fable 5
    • deepseek
      DeepSeek v4 Pro
    • google
      Gemini 3.7 Flash
    • grok
      Grok 4.6
    Explore models >

    Image

    • openai
      GPT Image 2.5 SunburstNew
    • openai
      GPT Image 2
    • google
      Nano Banana Pro (Gemini 3 Pro Image)
    • google
      Nano Banana 2 (Gemini 3.1 Flash Image)
    • midjourney
      Midjourney
    • openai
      GPT Image 2.5 Flare
    Explore models >

    Video

    • minimax
      Minimax H3New
    • bytedance
      Seedance 2.5 (Build 260628)
    • bytedance
      Seedance 2.0 (Build 260128)
    • bytedance
      Seedance 2.0 Mini (Build 260615)
    • kling
      Kling v3.0 4k
    • vidu
      Vidu Q3 Turbo
    Explore models >
    Explore 232+ Models >
  • Generator

    • Create Image
    • Create Video
    • Edit in Canvas
    • Chat

    Features

    • AI Cat Dance Video GeneratorNew
    • AI Story Board Generator
    • Unrestricted AI Video Generator
    • Cute Wallpaper Generator
    • AI French Kissing Generator
    • AI Age Filter
    • AI Packaging Design Generator
    • Anime to Real Life AI
    • Anime AI Art Generator
    • AI Object Remover
    Explore All >

    Prompts

    • Seedance 2.0 PromptsNew
    • GPT Image 2 Prompts
    • Nano Banana Pro Prompts
    • Seedream 5.0 Pro Prompts
    • Midjourney Prompts
  • AI Blog

    • How to Generate Product Images in Bulk with the Seedream 5.0 Pro API
    • How to Make an AI Cat Dance Video With Your Own Cat
    • DeepSeek Flash vs Kimi K3: Which Is Better for Coding and Agents?
    • 5 Best Affordable AI Video APIs in 2026: Pricing, Ecommerce, and Short Drama
    • 7 Best Venice API Alternatives in 2026 for Developers
    Explore All >

    AI Insight

    • What Is GPT-6 Sol? Release Status, Pricing, Tests, and What We Know
    • What Is Claude Opus 5.2? Release Status, Rumors, and What We Know
    • AI video generates wrong text: The real fix
    • OpenAI-compatible API 401 & 403 error Fix Guide
    • AI video has no sound? Why it happens and fixes
    Explore All >

    AI Docs

    • gpt-image-2
    • gpt-5.4
    • kimi-k2.5
    • claude-opus-4-6
    • kling-v3.0-pro
    Explore All >

    AI Skills

    • browser-use
    • claude-to-im
    • competitive-ads-extractor
    • content-creator
    • data-storytelling
    Explore All >
Pricing+7% bonus
English繁體中文한국어日本語EspañolРусский
Get Started Now
  1. Home
  2. /Model
  3. /MiniMax
  4. /minimax-h3
MiniMax
Minimax H3
$ 
MiniMax H3 is an open-weight multimodal video model for text-to-video, first- and last-frame image-to-video, reference-led generation, and video editing. It accepts text, images, video, and audio, then produces 4–15-second clips at 768P or 2K with native stereo sound.

Modalities

Input: TextInput: Image
Output: Video

API Usage Examples
$ 
curl --request POST "https://gptproto.com/api/v3/minimax/minimax-h3/text-to-video" \
  --header "Authorization: Bearer $GPTPROTO_API_KEY" \
  --header "Content-Type: application/json" \
  --data '{
    "prompt": "A tiny origami fox sailing a teacup across a moonlit puddle",
    "aspect_ratio": "16:9",
    "duration": 5,
    "resolution": "768P"
  }'
Minimax H3 pricing

Start from the cost of a single sample and pick a testing budget. GPTProto rates are 30% below list price.

Your usage

$0.455 / image

Related Models
All Models
Minimax H3
Current
$ 
byMiniMax$0.224 per time2K · more
Wan 3.0
$ 
byQwen$0.09 per timemore
Seedance 2.5 (Build 260628)
$ 
byBytedance$0.113 per second480p · 720p
Kling v3.0 4k
$ 
byKling$0.336 per second4k
Seedance 2.0 Mini (Build 260615)
$ 
byBytedance$0.0237 per second480p · 720p
Kling v3 Omni 4k
$ 
byKling$0.336 per second4k
Seedance 2.0 Fast (Build 260128)
$ 
byBytedance$0.0443 per second480p · 720p
Vidu 2.0
$ 
byVidu$0.02 per second720p · 1080p · more
Kling v3 Omni Pro
$ 
byKling$0.0896 per second
Vidu Q3 Turbo
$ 
byVidu$0.032 per second540p · 720p · 1080p
Vidu Q3 Pro
$ 
byVidu$0.04 per second540p · 720p · 1080p
Wan 2.6
$ 
byQwen$0.09 per second720p · 1080p · 2k
Video Watermark Remover
$ 
byGPTProto$0.05 per time
Veo 3.1 Fast Generate Preview
$ 
byGoogle$1.2 per time
Veo 3.1 Generate Preview
$ 
byGoogle$3.2 per time
Hailuo 2.3 Fast
$ 
byMiniMax$0.0285 per second1080p · more
Hailuo 2.3 Pro
$ 
byMiniMax$0.441 per time
Hailuo 2.3 Standard
$ 
byMiniMax$0.042 per second
Hailuo 02 Standard
$ 
byMiniMax$0.042 per second
Hailuo 02 Pro
$ 
byMiniMax$0.441 per time
Hailuo 02 Fast
$ 
byMiniMax$0.0135 per second
Wan 2.2 Plus
$ 
byQwen$0.09 per time1K · more
Veo 3.1
$ 
byGoogle$0.5 per time
Sora 2 Pro
$ 
byOpenAI$0.3 per second720p · 1080p
Sora 2
$ 
byOpenAI$0.1 per second720p
Higgsfield Turbo
$ 
byHiggsfield$0.2842 per time
Higgsfield Lite
$ 
byHiggsfield$0.0875 per time
Higgsfield Standard
$ 
byHiggsfield$0.3941 per time
ModelResolutionInput → Output
Minimax H3Current
$ 
$0.22 per time2K · more
Input: TextInput: Image
Output: Video
Wan 3.0
$ 
$0.09 per timemore
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Seedance 2.5 (Build 260628)
$ 
—$0.11 per second480p · 720p
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Kling v3.0 4k
$ 
$0.34 per second4k
Input: TextInput: Image
Output: Video
Seedance 2.0 Mini (Build 260615)
$ 
—$0.02 per second480p · 720p
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Kling v3 Omni 4k
$ 
$0.34 per second4k
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Seedance 2.0 Fast (Build 260128)
$ 
—$0.04 per second480p · 720p
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Vidu 2.0
$ 
$0.02 per second720p · 1080p · more
Input: ImageInput: VideoInput: Audio
Output: Video
Kling v3 Omni Pro
$ 
$0.09 per second—
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Vidu Q3 Turbo
$ 
$0.03 per second540p · 720p · 1080p
Input: TextInput: Image
Output: Video
Vidu Q3 Pro
$ 
$0.04 per second540p · 720p · 1080p
Input: TextInput: Image
Output: Video
Wan 2.6
$ 
$0.09 per second720p · 1080p · 2k
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Video Watermark Remover
$ 
—$0.05 per time—
Input: Video
Output: Video
Veo 3.1 Fast Generate Preview
$ 
—$1.20 per time—
Input: TextInput: ImageInput: Video
Output: Video
Veo 3.1 Generate Preview
$ 
—$3.20 per time—
Input: TextInput: ImageInput: Video
Output: Video
Hailuo 2.3 Fast
$ 
$0.03 per second1080p · more
Input: Image
Output: Video
Hailuo 2.3 Pro
$ 
$0.44 per time—
Input: TextInput: Image
Output: Video
Hailuo 2.3 Standard
$ 
$0.04 per second—
Input: TextInput: Image
Output: Video
Hailuo 02 Standard
$ 
$0.04 per second—
Input: TextInput: Image
Output: Video
Hailuo 02 Pro
$ 
$0.44 per time—
Input: TextInput: Image
Output: Video
Hailuo 02 Fast
$ 
$0.01 per second—
Input: Image
Output: Video
Wan 2.2 Plus
$ 
$0.09 per time1K · more
Input: TextInput: Image
Output: Video
Veo 3.1
$ 
—$0.50 per time—
Input: TextInput: ImageInput: VideoInput: Audio
Output: Video
Sora 2 Pro
$ 
—$0.30 per second720p · 1080p
Input: TextInput: Image
Output: Video
Sora 2
$ 
—$0.10 per second720p
Input: TextInput: Image
Output: Video
Higgsfield Turbo
$ 
$0.28 per time—
Input: Image
Output: Video
Higgsfield Lite
$ 
$0.09 per time—
Input: Image
Output: Video
Higgsfield Standard
$ 
$0.39 per time—
Input: Image
Output: Video

MiniMax-H3 API for 2K Multimodal Video Generation

Access the MiniMax-H3 API through GPTProto for cinematic video generation from prompts, keyframes, and mixed reference assets. Build 4–15-second clips with native stereo audio, use up to nine reference images, three video clips, and three audio tracks, and manage MiniMax H3 alongside 200+ AI models with one API key and one balance.

Text, Image, Video, and Audio in One Context

Combine a prompt with image, video, and audio references. MiniMax H3 can use those assets to guide subject identity, motion, camera behavior, visual style, voice, and editing rhythm.

4–15 Seconds at 768P or 2K

Choose an integer duration from 4 to 15 seconds. Generate at 768P for drafts or request 2K output when sharper product detail, typography, or final-delivery quality matters.

What Is MiniMax H3?

MiniMax H3 is MiniMax's general-purpose, open-weight video generation system. Instead of treating text-to-video, keyframe animation, and multimodal references as unrelated tools, it interprets text, images, video, and audio within one context. The hosted API uses an asynchronous workflow: submit a task, store its task ID, check its status, and retrieve the finished video URL.

Output runs from 4 to 15 seconds at 24 fps, with 768P and 2K modes, landscape, square, and portrait ratios, and native stereo audio. The released H3-Base weights use the MiniMax H3 Community License; hosted API access and self-hosting remain separate deployment choices.

Specification MiniMax H3 API Details
Provider MiniMax
Official API model name MiniMax-H3
GPTProto model string MiniMax-H3
Generation modes Text-to-video, first-frame I2V, last-frame I2V, first-and-last-frame I2V, multimodal reference-to-video
Input types Text, images, video, and audio
Output duration 4–15 seconds, integer values
Output resolution 768P or 2K
Frame rate 24 fps
Output audio Native 32 kHz stereo
Aspect ratios 21:9, 16:9, 4:3, 1:1, 3:4, 9:16; adaptive where supported
Reference limits Up to 9 images, 3 video clips, and 3 audio clips; up to 12 files combined
Prompt limit Up to 7,000 characters in the official API
Processing Asynchronous task workflow

Which MiniMax H3 Input Mode Should You Use?

Choose the input mode before building the request. First/last-frame inputs and multimodal reference inputs are separate modes in the official API and cannot be mixed in one generation task.

Input Mode Use It When Practical Example
Text-to-video You need a new scene without an approved visual starting point Generate several cinematic concepts from one campaign brief
First-frame image-to-video The opening composition or product image is already approved Animate a product hero image while retaining its initial framing
Last-frame image-to-video The shot must resolve to a specific final composition End on a logo lockup, product pack, or call-to-action frame
First-and-last-frame video Both endpoints of a transition matter Move from a closed package to the product fully revealed
Reference-to-video Identity, motion, voice, music, setting, or style must come from source assets Keep the same product and spokesperson while borrowing motion from a reference clip

For reference-to-video, MiniMax H3 accepts up to nine images, three video clips, and three audio clips, with a maximum of 12 files in total. Each reference video or audio clip can be 2–15 seconds, and the total duration for each media type cannot exceed 15 seconds. Use public URLs for large assets because the complete request body is limited to 64 MB.

MiniMax H3 API for Ecommerce and Batch Video Workflows

For ecommerce, supply approved product images, state which details must remain unchanged, and add a reference video only when its motion or camera path matters. First-and-last-frame mode is often simpler for controlled product reveals.

For batch generation, treat every variation as a separate asynchronous job. Store its prompt, asset URLs, settings, task ID, and status; use a bounded queue; make retries idempotent; and copy successful outputs to your own storage. Keep logo treatment, framing, lighting, camera language, and CTA fixed while injecting variable SKU data per job.

MiniMax H3 vs Seedance 2.5 vs Veo 3.1

These models overlap but solve different constraints. MiniMax H3 emphasizes 2K output, open weights, stereo audio, and mixed references. Seedance 2.5 supports longer generations and larger reference sets. Veo 3.1 fits short Google Cloud workflows with a documented 4K path.

Decision Factor MiniMax H3 Seedance 2.5 Veo 3.1 Generate
Single-generation duration 4–15 seconds 4–30 seconds 4, 6, or 8 seconds
Documented output resolution 768P or 2K 480P, 720P, or 1080P 720P, 1080P, or 4K
Input modalities Text, image, video, audio Text, image, video, audio Text and image; no audio or video input in the documented Generate endpoint
Reference capacity Up to 9 images, 3 videos, 3 audio clips; 12 files total Up to 30 images, 10 videos, and 10 audio clips Up to 3 asset images
First/last-frame control Yes Yes Yes
Native generated audio Yes Yes Yes
Open weights H3-Base released under a community license No published open weights No
Best fit 2K multimodal reference work, motion transfer, editing, and self-hosting research Longer scenes, large reference packs, and timeline-led editing Short high-fidelity clips in Google Cloud and 4K delivery workflows

Compare duration, reference capacity, audio, retries, and finished-clip cost—not maximum resolution alone. Seedance 2.5 may reduce scene stitching; H3 may simplify mixed-reference 2K work; Veo 3.1 may suit short Google-centered 4K workflows.

How to Write Better MiniMax H3 Prompts

Write the prompt as a compact production brief that separates source roles, timeline, camera direction, audio, preserved details, and exclusions.

Prompt formula: subject and reference roles + scene goal + timed actions + camera movement + lighting and visual style + dialogue, sound effects, and music + details to preserve + final frame

Ecommerce Product Video Prompt

Use Image 1 for product identity and Image 2 for lighting. Create a 10-second vertical launch video for the matte-black speaker. Begin with a macro grille shot, pull back as it rotates, then show water droplets vibrating with the bass. Preserve the logo, controls, proportions, and finish. Use cool edge light, restrained motion, low electronic ambience, and no extra text.

First-and-Last-Frame Transition Prompt

Start from the supplied closed package and end exactly on the supplied final frame with the product assembled. Use one continuous camera move: the box opens, components rise, and the product locks into place. Preserve package graphics and final geometry. Add mechanical clicks and a short resolved tone.

MiniMax H3 API FAQ

What is the MiniMax H3 API?

It provides hosted access to MiniMax's open-weight multimodal video model for 4–15-second video at 768P or 2K with native stereo sound.

Does MiniMax H3 support both text-to-video and image-to-video?

Yes. It supports prompt-only generation, first- or last-frame I2V, two-keyframe transitions, and a separate multimodal reference mode.

What duration and resolution does MiniMax H3 support?

The official API accepts integer durations from 4 to 15 seconds at 24 fps, with 768P and 2K output modes.

Does MiniMax H3 generate audio with the video?

Yes. It generates native 32 kHz stereo sound. Prompts can direct dialogue, ambience, effects, and music; reference audio can guide voice or rhythm.

How many reference assets can I use?

Up to nine images, three video clips, and three audio clips, with 12 files total. Video and audio references are each limited to 15 seconds in aggregate.

Can I use MiniMax H3 for batch ecommerce video generation?

Yes, through application-level orchestration: submit one task per variation, store task IDs, limit concurrency, make retries idempotent, and save completed files to your storage.

MiniMax H3 or Seedance 2.5: which is better?

Choose H3 for 2K, open H3-Base weights, or compact mixed-reference jobs. Choose Seedance 2.5 for clips up to 30 seconds, larger reference packs, or timeline-led editing.

MiniMax H3 vs Veo 3.1: which is more cost-effective?

Compare the live rate for the same duration and resolution, then include retries and usable-output rate. H3 offers straightforward 768P/2K cost planning; Veo 3.1 may be preferable when Google Cloud integration or 4K delivery matters.

How do I access the MiniMax H3 API on GPTProto?

Create a GPTProto key and select H3 in the playground or documentation. Use the model string shown in the live API example: MiniMax-H3.

Is MiniMax H3 the same as Hailuo 3?

MiniMax H3 is the official name. “Hailuo 3” and “MiniMax Hailuo H3” are search aliases tied to MiniMax's Hailuo product; API docs identify the model as MiniMax-H3.

Related Articles

Guides, comparisons, and updates related to this model.

All Articles
5 Best Affordable AI Video APIs in 2026: Pricing, Ecommerce, and Short Drama

5 Best Affordable AI Video APIs in 2026: Pricing, Ecommerce, and Short Drama

Compare 5 affordable AI video APIs for ecommerce and AI short drama. See current pricing, clip costs, audio fees, and the best model for each job.

6 Cheapest AI Image Generators in 2026: Real Cost per Image

6 Cheapest AI Image Generators in 2026: Real Cost per Image

Compare 6 cheapest AI image generators in 2026, from about $0.0035 per image. See batch costs, hidden fees, and the best API for startups.

How to Make an AI Story Video for Kids with GPT Image 2 and Seedance 2.5

How to Make an AI Story Video for Kids with GPT Image 2 and Seedance 2.5

Learn how to make an AI story video for kids with GPT Image 2 and Seedance 2.5, including the full prompt, captions, sound, editing, and a real test.

Seedance 2.0 vs Seedance 2.5: Same Prompt, Storyboard, and Real Results

Seedance 2.0 vs Seedance 2.5: Same Prompt, Storyboard, and Real Results

See Seedance 2.0 vs 2.5 in the same 15-second storyboard test. Compare cinematic quality, emotion, pricing, ecommerce use cases, and API features.

GPT Proto

Empowering AI Innovation with Global Scale and Stability:

With our flagship product GPT Proto, we offer a unified interface to access and combine APIs from the world's leading AI providers—spanning text, vision, speech, and beyond. We empower developers and enterprises to simplify integration and accelerate innovation without limits.

Global Infrastructure, Local Compliance:

To ensure enterprise-grade reliability and compliance, Talent Tech Global Limited operates specifically as our global Billing and Contracting Entity. Meanwhile, our core technical infrastructure and R&D teams are strategically distributed across global innovation hubs, including Silicon Valley, Singapore, and Hong Kong.

Built to Scale:

We understand that stability is paramount. Our platform is built on a robust, decentralized architecture supporting dynamic Auto-scaling. Whether you are running a pilot or handling millions of concurrent requests, our system expands instantly to meet demand—guaranteeing that your business never outgrows our infrastructure.

Navigation

  • Dashboard
  • Models
  • Create Image
  • AI Image Upscale
  • AI Background Remover
  • Create Video
  • Edit in Canvas
  • Chat
  • Features
  • Pricing
  • AI Docs
  • AI Blog
  • AI Insight
  • AI Skills

Features

  • AI Cat Dance Video Generator
  • AI Story Board Generator
  • Unrestricted AI Video Generator
  • Cute Wallpaper Generator
  • AI French Kissing Generator
  • AI Age Filter
  • AI Packaging Design Generator
  • Anime to Real Life AI
  • Anime AI Art Generator
  • AI Object Remover
  • AI Image Editor
  • AI Motion Transfer
  • AI Watermark Remover
  • AI Image Enhancer Online
  • Online Background Remover Tool
  • AI Face Swap Image
  • AI Passport Photo Maker
  • MS Paint AI Generator
  • AI Clothes Remover
  • Unrestricted AI Image Generator
Explore all features >

LLM

  • DeepSeek Flash
  • GLM 5.3
  • Claude Fable 5
  • DeepSeek v4 Pro
  • Gemini 3.7 Flash
  • Grok 4.6
  • Hy4 Preview
  • GPT 6 Astra
  • Gemini 3.8 Flash
  • Claude Fable 5.1
  • Qwen3.8 Max 0902
  • GLM 5.3 Flash
  • DeepSeek v4 Flash Vision Exp
  • Qwen3.8 Max
  • Claude Opus 5
  • Gemini 3.6 Flash
  • Gemini 3.5 Flash Lite
  • Kimi K3
  • GPT 5.6 Luna
  • GPT 5.6 Terra
Explore all models >

Image

  • GPT Image 2.5 Sunburst
  • GPT Image 2
  • Nano Banana Pro (Gemini 3 Pro Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Midjourney
  • GPT Image 2.5 Flare
  • Grok Imagine Image 2.0
  • Seedream 5.0 Pro (Build 260628)
  • Nano Banana 2 Lite (Gemini 3.1 Flash-Lite Image)
  • Nano Banana 2 (Gemini 3.1 Flash Image)
  • Seedream 5.0 (Build 260128)
  • Doubao Seedream 5.0 (Build 260128)
  • Vidu Q2
  • Grok Imagine Image
  • Kling Image O1
  • GPT Image 1.5
  • Seedream 4.5 (Build 251128)
  • Doubao Seedream 4.5 (Build 251128)
  • Grok Imagine 0.9
  • Qwen Image Lora
Explore all models >

Video

  • Minimax H3
  • Seedance 2.5 (Build 260628)
  • Seedance 2.0 (Build 260128)
  • Seedance 2.0 Mini (Build 260615)
  • Kling v3.0 4k
  • Vidu Q3 Turbo
  • Wan 3.0
  • Kling v3 Omni 4k
  • Seedance 2.0 Fast (Build 260128)
  • Vidu 2.0
  • Doubao Seedance 2.0 (Build 260128)
  • Doubao Seedance 2.0 Fast (Build 260128)
  • Kling v3 Omni Pro
  • Kling v3 Omni Std
  • Kling v3.0 Pro
  • Kling v3.0 Std
  • Vidu Q3 Pro
  • Kling v2.6 Std
  • Vidu Q2 Pro
  • Vidu Q2 Turbo
Explore all models >

Contact us

Questions or feedback? Reach us through any of the channels below.

TelegramWhatsApp

© 2026 Talent Tech Global Limited (Hong Kong). All rights reserved.

Registered Address: Unit 1022a, Beverley Commercial Centre, 87-105 Chatham Road South, Tsim Sha Tsui, Hong KongCertificate No.: 79462435-000-12-25-0
  • About Us
  • Privacy Policy
  • Terms of Service
  • Sitemap
Friendslogoto.videotopostudio.cc

Input

Output

Your browser does not support the video tag.