Sign inGet $5 free

Kling LipSync Text-to-Video

Makes a video from a description

Liveby Kling

Kling LipSync is a text-to-video model that generates realistic lip movements from text input.

Ask your agent

Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.

Sync this video to this voice recording with Kling LipSync Text-to-Video.
Make the person in this clip say the words in this audio file.

Reviews

No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.

How it works

  1. Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
  2. Ask in your own wordsYour agent picks the right tool and fills in the details for you.
  3. Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema

Video generation API by Kling, callable through Superpowers. It costs $0.076 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.kling-video-lipsync-text-to-video", or over MCP with run_tool. 89 other providers in this category take the same input and can be compared side by side.

Tool id
Price per call
$0.076 / 5s clip

Call it

Also over MCP as run_tool
curl -X POST https://superpowers.tools/v1/run \
  -H 'Authorization: Bearer $SP_KEY' \
  -d '{"tool":"fal.kling-video-lipsync-text-to-video","input":{"voice_id":"","video_url":"","text":""},"idempotency_key":"example-request-1"}'

Input

5 fields · 3 required
{
  "type": "object",
  "required": [
    "video_url",
    "text",
    "voice_id"
  ],
  "properties": {
    "voice_id": {
      "type": "string",
      "description": "Voice ID to use for speech synthesis",
      "enum": [
        "genshin_vindi2",
        "zhinen_xuesheng",
        "AOT",
        "ai_shatang",
        "genshin_klee2",
        "genshin_kirara",
        "ai_kaiya",
        "oversea_male1",
        "ai_chenjiahao_712",
        "girlfriend_4_speech02",
        "chat1_female_new-3",
        "chat_0407_5-1",
        "cartoon-boy-07",
        "uk_boy1",
        "cartoon-girl-01",
        "PeppaPig_platform",
        "ai_huangzhong_712",
        "ai_huangyaoshi_712",
        "ai_laoguowang_712",
        "chengshu_jiejie",
        "you_pingjing",
        "calm_story1",
        "uk_man2",
        "laopopo_speech02",
        "heainainai_speech02",
        "reader_en_m-v1",
        "commercial_lady_en_f-v1",
        "tiyuxi_xuedi",
        "tiexin_nanyou",
        "girlfriend_1_speech02",
        "girlfriend_2_speech02",
        "zhuxi_speech02",
        "uk_oldman3",
        "dongbeilaotie_speech02",
        "chongqingxiaohuo_speech02",
        "chuanmeizi_speech02",
        "chaoshandashu_speech02",
        "ai_taiwan_man2_speech02",
        "xianzhanggui_speech02",
        "tianjinjiejie_speech02",
        "diyinnansang_DB_CN_M_04-v2",
        "yizhipiannan-v1",
        "guanxiaofang-v2",
        "tianmeixuemei-v1",
        "daopianyansang-v1",
        "mengwa-v1"
      ]
    },
    "video_url": {
      "type": "string",
      "description": "The URL of the video to generate the lip sync for. Supports .mp4/.mov, ≤100MB, 2-60s, 720p/1080p only, width/height 720–1920px. If validation fails, an error is returned."
    },
    "voice_language": {
      "type": "string",
      "description": "The voice language corresponding to the Voice ID",
      "default": "en",
      "enum": [
        "zh",
        "en"
      ]
    },
    "text": {
      "type": "string",
      "description": "Text content for lip-sync video generation. Max 120 characters."
    },
    "voice_speed": {
      "type": "number",
      "description": "Speech rate for Text to Video generation",
      "default": 1
    }
  }
}