MiniMax H3 Open Text to Video: AI Video Generator

Generate coherent AI videos with native audio using MiniMax H3 Open Text to Video. Supports 480p/768p and 5-15s clips. Try free — pay per generation.

📝

Overview

About this model

MiniMax H3 Open Text to Video generates highly coherent AI videos directly from text prompts, complete with natively generated stereo audio. Supporting both fast 480p and high-detail 768p resolution tiers, this open-weights model allows fine-grained control over duration (5 to 15 seconds), aspect ratios, and random seeds for consistent creative results. Compare with MiniMax H3 Text to Video and MiniMax H3 Reference to Video.

1Cinematic Video: Create high-fidelity scene renders with matching audio soundscapes.
2Social Content: Produce 9:16 vertical videos for TikTok, Reels, and Shorts.
3Creative Drafts: Rapidly preview text prompts at 480p before generating final 768p videos.
💰

Pricing & Value

Cost analysis

muapiapp$0.08–$0.1825 per second

Flexible resolution tiers (480p / 768p) and per-second billing with no subscription required.

Fal.ai$0.10–$0.20 per second

Higher per-second rates for similar MiniMax H3 endpoints.

ReplicateNot available

Model not hosted natively on Replicate.

* Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

Promptstring

Text description of the video scene, action, camera movement, and desired soundtrack.

Default ValueA cinematic ocean wave at sunrise, highly detailed
Aspect RatioEnum (7 options)

-

Default Value16:9
ResolutionEnum (2 options)

-

Default Value480p
DurationEnum (11 options)

Output duration in seconds.

Default Value5
📖

Implementation Guide

Developer documentation

How to Use MiniMax H3 Open Text to Video API

  1. Prepare your prompt: Write a descriptive text prompt explaining scene, action, camera movement, and audio context.
  2. Configure options: Choose aspect ratio (16:9, 9:16, etc.), duration (5 to 15 seconds), and resolution (480p or 768p).
  3. Submit request: Send a POST request to /api/v1/minimax-h3-open-text-to-video using curl or your HTTP client.
curl -X POST "https://api.muapi.ai/api/v1/minimax-h3-open-text-to-video" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{
    "prompt": "A cinematic ocean wave at sunrise, highly detailed",
    "aspect_ratio": "16:9",
    "resolution": "480p",
    "duration": 5
  }'
  1. Retrieve output: Poll the resulting request ID or receive webhooks automatically when processing completes.

Common Questions

Frequently asked

What is MiniMax H3 Open Text to Video?

MiniMax H3 Open Text to Video is an open-weights text-to-video model that generates coherent videos with native stereo audio from textual descriptions.

What resolutions and durations are supported?

The model supports 480p (faster, lower cost) and 768p (native canvas) resolutions, with video durations from 5 to 15 seconds.

Does it generate audio automatically?

Yes, native stereo audio is synthesized alongside the video frames based on your text prompt.

What aspect ratios are available?

Supported aspect ratios include 16:9, 9:16, 1:1, 4:3, 3:4, 21:9, and 9:21.

How is billing calculated for this model?

Pricing is per second of output video. 480p costs $0.08/sec and 768p costs $0.1825/sec.