
Kling LipSync Text-to-Video
Makes a video from a description
Liveby Kling
Kling LipSync is a text-to-video model that generates realistic lip movements from text input.
Ask your agent
Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.
Sync this video to this voice recording with Kling LipSync Text-to-Video.
Make the person in this clip say the words in this audio file.
Reviews
No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.
How it works
- Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
- Ask in your own wordsYour agent picks the right tool and fills in the details for you.
- Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema
Video generation API by Kling, callable through Superpowers. It costs $0.076 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.kling-video-lipsync-text-to-video", or over MCP with run_tool. 89 other providers in this category take the same input and can be compared side by side.
- Tool id
- Price per call
- $0.076 / 5s clip
Call it
Also over MCP asrun_toolcurl -X POST https://superpowers.tools/v1/run \
-H 'Authorization: Bearer $SP_KEY' \
-d '{"tool":"fal.kling-video-lipsync-text-to-video","input":{"voice_id":"","video_url":"","text":""},"idempotency_key":"example-request-1"}'Input
5 fields · 3 required{
"type": "object",
"required": [
"video_url",
"text",
"voice_id"
],
"properties": {
"voice_id": {
"type": "string",
"description": "Voice ID to use for speech synthesis",
"enum": [
"genshin_vindi2",
"zhinen_xuesheng",
"AOT",
"ai_shatang",
"genshin_klee2",
"genshin_kirara",
"ai_kaiya",
"oversea_male1",
"ai_chenjiahao_712",
"girlfriend_4_speech02",
"chat1_female_new-3",
"chat_0407_5-1",
"cartoon-boy-07",
"uk_boy1",
"cartoon-girl-01",
"PeppaPig_platform",
"ai_huangzhong_712",
"ai_huangyaoshi_712",
"ai_laoguowang_712",
"chengshu_jiejie",
"you_pingjing",
"calm_story1",
"uk_man2",
"laopopo_speech02",
"heainainai_speech02",
"reader_en_m-v1",
"commercial_lady_en_f-v1",
"tiyuxi_xuedi",
"tiexin_nanyou",
"girlfriend_1_speech02",
"girlfriend_2_speech02",
"zhuxi_speech02",
"uk_oldman3",
"dongbeilaotie_speech02",
"chongqingxiaohuo_speech02",
"chuanmeizi_speech02",
"chaoshandashu_speech02",
"ai_taiwan_man2_speech02",
"xianzhanggui_speech02",
"tianjinjiejie_speech02",
"diyinnansang_DB_CN_M_04-v2",
"yizhipiannan-v1",
"guanxiaofang-v2",
"tianmeixuemei-v1",
"daopianyansang-v1",
"mengwa-v1"
]
},
"video_url": {
"type": "string",
"description": "The URL of the video to generate the lip sync for. Supports .mp4/.mov, ≤100MB, 2-60s, 720p/1080p only, width/height 720–1920px. If validation fails, an error is returned."
},
"voice_language": {
"type": "string",
"description": "The voice language corresponding to the Voice ID",
"default": "en",
"enum": [
"zh",
"en"
]
},
"text": {
"type": "string",
"description": "Text content for lip-sync video generation. Max 120 characters."
},
"voice_speed": {
"type": "number",
"description": "Speech rate for Text to Video generation",
"default": 1
}
}
}