Sign inGet $5 free

Longcat Single Avatar

Turns a recording into a video

Liveby fal

LongCat-Video-Avatar is an audio-driven video generation model that can generates super-realistic, lip-synchronized long video generation with natural dynamics and consistent identity.

Ask your agent

Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.

Make a 5-second video with Longcat Single Avatar of a drone flying over Florianópolis at sunrise.
Turn this product photo into a short video for Instagram.

Reviews

No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.

How it works

  1. Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
  2. Ask in your own wordsYour agent picks the right tool and fills in the details for you.
  3. Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema

Video API by fal, callable through Superpowers. It costs $1.62 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.longcat-single-avatar-image-audio-to-video", or over MCP with run_tool.

Tool id
Price per call
$1.62 / 5s clip

Call it

Also over MCP as run_tool
curl -X POST https://superpowers.tools/v1/run \
  -H 'Authorization: Bearer $SP_KEY' \
  -d '{"tool":"fal.longcat-single-avatar-image-audio-to-video","input":{"image_url":"","audio_url":""},"idempotency_key":"example-request-1"}'

Input

12 fields · 2 required
{
  "type": "object",
  "required": [
    "image_url",
    "audio_url"
  ],
  "properties": {
    "text_guidance_scale": {
      "type": "number",
      "description": "The text guidance scale for classifier-free guidance.",
      "default": 4
    },
    "num_segments": {
      "type": "integer",
      "description": "Number of video segments to generate. Each segment adds ~5 seconds of video. First segment is ~5.8s, additional segments are 5s each.",
      "default": 1
    },
    "enable_safety_checker": {
      "type": "boolean",
      "description": "Whether to enable safety checker. Disabling it requires account authorization; unauthorized requests are always checked.",
      "default": true
    },
    "image_url": {
      "type": "string",
      "description": "The URL of the image to animate."
    },
    "seed": {
      "type": "integer",
      "description": "The seed for the random number generator."
    },
    "prompt": {
      "type": "string",
      "description": "The prompt to guide the video generation.",
      "default": "A person is talking naturally with natural expressions and movements."
    },
    "audio_guidance_scale": {
      "type": "number",
      "description": "The audio guidance scale. Higher values may lead to exaggerated mouth movements.",
      "default": 4
    },
    "negative_prompt": {
      "type": "string",
      "description": "The negative prompt to avoid in the video generation.",
      "default": "Close-up, Bright tones, overexposed, static, blurred details, subtitles, style, works, paintings, images, static, overall gray, worst quality, low quality, JPEG compression residue, ugly, incomplete, extra fingers, poorly drawn hands, poorly drawn faces, deformed, disfigured, misshapen limbs, fused fingers, still picture, messy background, three legs, many people in the background, walking backwards"
    },
    "audio_url": {
      "type": "string",
      "description": "The URL of the audio file to drive the avatar."
    },
    "num_inference_steps": {
      "type": "integer",
      "description": "The number of inference steps to use.",
      "default": 50
    },
    "resolution": {
      "type": "string",
      "description": "Resolution of the generated video (480p or 720p). Billing is per video-second (16 frames): 480p is 1 unit per second and 720p is 4 units per second.",
      "default": "480p",
      "enum": [
        "480p",
        "720p"
      ]
    },
    "enable_prompt_expansion": {
      "type": "boolean",
      "description": "Whether to enable prompt expansion using an LLM to enhance the prompt for better video quality.",
      "default": false
    }
  }
}