
Scail 2
Edits a video you give it
Liveby fal
SCAIL-2 is an end-to-end character animation model that drives a reference character from a source video without relying on intermediate pose representations like skeleton maps.
Ask your agent
Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.
Use Scail 2 to make this clip look like it was shot at night.
Change the background of this video to a beach.
Reviews
No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.
How it works
- Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
- Ask in your own wordsYour agent picks the right tool and fills in the details for you.
- Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema
Video API by fal, callable through Superpowers. It costs $1.08 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.scail-2", or over MCP with run_tool.
- Tool id
- Price per call
- $1.08 / 5s clip
Call it
Also over MCP asrun_toolcurl -X POST https://superpowers.tools/v1/run \
-H 'Authorization: Bearer $SP_KEY' \
-d '{"tool":"fal.scail-2","input":{"image_url":"","prompt":"","video_url":""},"idempotency_key":"example-request-1"}'Input
12 fields · 3 required{
"type": "object",
"required": [
"prompt",
"image_url",
"video_url"
],
"properties": {
"driving_type": {
"type": "string",
"description": "Driving signal for animation mode. `end_to_end` (default) drives directly from the SAM3-masked driving video and is the recommended, most robust path. `pose` renders an explicit NLF/DWPose skeleton from the driving video (more control for challenging inputs).",
"default": "end_to_end",
"enum": [
"end_to_end",
"pose"
]
},
"enable_safety_checker": {
"type": "boolean",
"description": "Authorized callers may set false to request relaxed screening.",
"default": true
},
"image_url": {
"type": "string",
"description": "URL of the reference character image. The character in this image is animated (or used to replace a subject) according to the driving video."
},
"num_inference_steps": {
"type": "integer",
"description": "Number of diffusion sampling steps. Higher improves quality but is slower.",
"default": 40
},
"mode": {
"type": "string",
"description": "`animation` animates the reference character with the driving motion. `replacement` replaces the driving subject with the reference character while keeping the original scene.",
"default": "animation",
"enum": [
"animation",
"replacement"
]
},
"seed": {
"type": "integer",
"description": "Random seed. Leave empty for a random seed. The seed used is returned in the response."
},
"subject_type": {
"type": "string",
"description": "Type of subject in the driving video. `animal` switches the SAM3 detection prompt to non-human subjects.",
"default": "human",
"enum": [
"human",
"animal"
]
},
"guidance_scale": {
"type": "number",
"description": "Classifier-free guidance scale. Controls prompt adherence versus creativity.",
"default": 5
},
"shift": {
"type": "number",
"description": "Flow-matching noise schedule shift. 3.0 is recommended for the 512p tier.",
"default": 3
},
"resolution": {
"type": "string",
"description": "Output resolution. 512p outputs 896x512 (landscape) or 512x896 (portrait); 704p outputs 1280x704 or 704x1280. Orientation is chosen automatically from the reference image aspect ratio.",
"default": "512p",
"enum": [
"512p",
"704p"
]
},
"prompt": {
"type": "string",
"description": "The prompt describing the final video to generate."
},
"video_url": {
"type": "string",
"description": "URL of the driving (motion) video. The subject can be a human, multiple humans, or an animal."
}
}
}