Sign inGet $5 free

MiniMax H3 Max Text to Video

Makes a video from a description

Liveby Minimax

fal's H3 Max is a post-trained variant of MiniMax H3, tuned for stronger prompt adherence and better aesthetics while co-optimized with our custom inference stack for higher throughput with no compromises on output quality

Ask your agent

Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.

Make a 5-second video with MiniMax H3 Max Text to Video of a drone flying over Florianópolis at sunrise.
Make a short video of waves rolling onto a beach at sunset.

Reviews

No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.

How it works

  1. Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
  2. Ask in your own wordsYour agent picks the right tool and fills in the details for you.
  3. Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema

Video generation API by Minimax, callable through Superpowers. It costs $0.162 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.minimax-h3-max-text-to-video", or over MCP with run_tool. 89 other providers in this category take the same input and can be compared side by side.

Tool id
Price per call
$0.162 / 5s clip

Call it

Also over MCP as run_tool
curl -X POST https://superpowers.tools/v1/run \
  -H 'Authorization: Bearer $SP_KEY' \
  -d '{"tool":"fal.minimax-h3-max-text-to-video","input":{"prompt_expansion_mode":"balanced","prompt":""},"idempotency_key":"example-request-1"}'

Input

9 fields · 2 required
{
  "type": "object",
  "required": [
    "prompt",
    "prompt_expansion_mode"
  ],
  "properties": {
    "aspect_ratio": {
      "type": "string",
      "description": "The aspect ratio of the generated video.",
      "default": "16:9",
      "enum": [
        "21:9",
        "16:9",
        "4:3",
        "1:1",
        "3:4",
        "9:16"
      ]
    },
    "prompt_expansion_mode": {
      "type": "string",
      "description": "How much effort to spend rewriting the prompt before generation. 'disabled' skips prompt expansion. 'balanced' returns in about a second. 'quality' spends up to ~30s on a richer prompt.",
      "default": "balanced"
    },
    "target_audio_url": {
      "type": "string",
      "description": "Optional URL of an audio clip at least 2 seconds long (maximum 15 MB) to pin to the generated soundtrack. Longer clips are trimmed to the generated video's length, keeping the beginning. The original audio replaces the output soundtrack, padded with silence if shorter than the video, without changin"
    },
    "sync_mode": {
      "type": "boolean",
      "description": "Return the generated video as base64 instead of a CDN URL.",
      "default": false
    },
    "seed": {
      "type": "integer",
      "description": "Random seed. A random seed is selected when omitted."
    },
    "prompt": {
      "type": "string",
      "description": "Text prompt for video generation"
    },
    "duration": {
      "type": "number",
      "description": "Video length in seconds. The output can run up to about 0.7 s longer than requested. Use at least 2 s with target_audio_url.",
      "default": 5
    },
    "resolution": {
      "type": "string",
      "description": "The native generation resolution, or 1080P latent refinement from a native 768P source.",
      "default": "768P",
      "enum": [
        "480P",
        "768P",
        "1080P"
      ]
    },
    "enable_safety_checker": {
      "type": "boolean",
      "description": "If set to true, the safety checker will be enabled.",
      "default": true
    }
  }
}