
ACE Step
Text to audio
Liveby fal
Generate music with lyrics from text using ACE-Step
Ask your agent
Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.
Read this paragraph aloud in a warm voice and send me the audio.
Turn my blog post into an audio file I can share.
Reviews
No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.
How it works
- Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
- Ask in your own wordsYour agent picks the right tool and fills in the details for you.
- Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema
Voice & audio API by fal, callable through Superpowers. It costs $0.00108 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.ace-step", or over MCP with run_tool.
- Tool id
- Price per call
- $0.00108 / 5s clip
Call it
Also over MCP asrun_toolcurl -X POST https://superpowers.tools/v1/run \
-H 'Authorization: Bearer $SP_KEY' \
-d '{"tool":"fal.ace-step","input":{"tags":""},"idempotency_key":"example-request-1"}'Input
14 fields · 1 required{
"type": "object",
"required": [
"tags"
],
"properties": {
"tag_guidance_scale": {
"type": "number",
"description": "Tag guidance scale for the generation.",
"default": 5
},
"guidance_interval_decay": {
"type": "number",
"description": "Guidance interval decay for the generation. Guidance scale will decay from guidance_scale to min_guidance_scale in the interval. 0.0 means no decay.",
"default": 0
},
"guidance_scale": {
"type": "number",
"description": "Guidance scale for the generation.",
"default": 15
},
"scheduler": {
"type": "string",
"description": "Scheduler to use for the generation process.",
"default": "euler",
"enum": [
"euler",
"heun"
]
},
"lyrics": {
"type": "string",
"description": "Lyrics to be sung in the audio. If not provided or if [inst] or [instrumental] is the content of this field, no lyrics will be sung. Use control structures like [verse], [chorus] and [bridge] to control the structure of the song.",
"default": ""
},
"guidance_type": {
"type": "string",
"description": "Type of CFG to use for the generation process.",
"default": "apg",
"enum": [
"cfg",
"apg",
"cfg_star"
]
},
"granularity_scale": {
"type": "integer",
"description": "Granularity scale for the generation process. Higher values can reduce artifacts.",
"default": 10
},
"tags": {
"type": "string",
"description": "Comma-separated list of genre tags to control the style of the generated audio. Can also be supplied as `prompt`."
},
"guidance_interval": {
"type": "number",
"description": "Guidance interval for the generation. 0.5 means only apply guidance in the middle steps (0.25 * infer_steps to 0.75 * infer_steps)",
"default": 0.5
},
"duration": {
"type": "number",
"description": "The duration of the generated audio in seconds.",
"default": 60
},
"seed": {
"type": "integer",
"description": "Random seed for reproducibility. If not provided, a random seed will be used."
},
"lyric_guidance_scale": {
"type": "number",
"description": "Lyric guidance scale for the generation.",
"default": 1.5
},
"number_of_steps": {
"type": "integer",
"description": "Number of steps to generate the audio.",
"default": 27
},
"minimum_guidance_scale": {
"type": "number",
"description": "Minimum guidance scale for the generation after the decay.",
"default": 3
}
}
}