
Wan 3.0 Prime
Edits a video you give it
Wan 3.0 Prime Reference-to-Video combines reference images, videos, and audio into a unified video with fast generation and strong multimodal coherence.
It follows character identity, visual style, movement, and sound cues across references to create controlled, consistent, and production-ready results.
Ask your agent
Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.
Reviews
No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.
How it works
- Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
- Ask in your own wordsYour agent picks the right tool and fills in the details for you.
- Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema
Video API by Alibaba, callable through Superpowers. It costs $0.270 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.alibaba-wan-3.0-prime-reference-to-video", or over MCP with run_tool.
- Tool id
- Price per call
- $0.270 / 5s clip
Call it
Also over MCP asrun_toolcurl -X POST https://superpowers.tools/v1/run \
-H 'Authorization: Bearer $SP_KEY' \
-d '{"tool":"fal.alibaba-wan-3.0-prime-reference-to-video","input":{"file_url":""},"idempotency_key":"example-request-1"}'Input
14 fields · all optional{
"type": "object",
"required": [],
"properties": {
"file_url": {
"type": "string",
"description": "Document URL to base the video on. Requires enable_thinking=true."
},
"audio": {
"type": "boolean",
"description": "Include generated audio.",
"default": true
},
"reference_video_urls": {
"type": "array",
"description": "Up to 5 reference video URLs totaling at most 15 seconds. Each clip must be at least 16 fps. Input video duration is billed in addition to output duration."
},
"reference_image_urls": {
"type": "array",
"description": "Up to 10 reference image URLs."
},
"enable_thinking": {
"type": "boolean",
"description": "Enable enhanced reasoning before generation.",
"default": false
},
"reference_audio_urls": {
"type": "array",
"description": "Up to 5 reference audio URLs totaling at most 15 seconds."
},
"enable_prompt_expansion": {
"type": "boolean",
"description": "Enable intelligent prompt rewriting. Disabling it can save roughly 20-60 seconds of latency but is likely to degrade generation quality.",
"default": true
},
"aspect_ratio": {
"type": "string",
"description": "Output aspect ratio, or adaptive selection.",
"default": "adaptive",
"enum": [
"adaptive",
"16:9",
"4:3",
"1:1",
"3:4",
"9:16"
]
},
"seed": {
"type": "integer"
},
"prompt": {
"type": "string",
"description": "Text prompt directing how the reference media is used. Reference media can be addressed positionally, e.g. 'the subject in Image 1 walks past Video 1'."
},
"duration": {
"type": "integer",
"description": "Output duration in seconds. Set to null for smart duration, which lets the model pick a length from the prompt and reference media.",
"default": 5
},
"enable_safety_checker": {
"type": "boolean",
"description": "Enable content moderation for input and output. Disabling it requires account authorization; unauthorized requests are always checked.",
"default": true
},
"web_url": {
"type": "string",
"description": "Public webpage URL to base the video on. Requires enable_thinking=true. Only pages that do not require login can be read."
},
"resolution": {
"type": "string",
"description": "Output video resolution tier.",
"default": "1080p",
"enum": [
"480p",
"720p",
"1080p"
]
}
}
}