Sign inGet $5 free

VibeVoice 7B

Reads text aloud

Liveby fal

Generate long, expressive multi-voice speech using Microsoft's powerful TTS

Ask your agent

Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.

Read this paragraph aloud in a warm voice and send me the audio.
Turn my blog post into an audio file I can share.

Reviews

No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.

How it works

  1. Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
  2. Ask in your own wordsYour agent picks the right tool and fills in the details for you.
  3. Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema

Text to speech API by fal, callable through Superpowers. It costs $0.216 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.vibevoice-7b", or over MCP with run_tool. 4 other providers in this category take the same input and can be compared side by side.

Tool id
Price per call
$0.216 / 5s clip

Call it

Also over MCP as run_tool
curl -X POST https://superpowers.tools/v1/run \
  -H 'Authorization: Bearer $SP_KEY' \
  -d '{"tool":"fal.vibevoice-7b","input":{"script":"","speakers":["marketing digital"]},"idempotency_key":"example-request-1"}'

Input

4 fields · 2 required
{
  "type": "object",
  "required": [
    "script",
    "speakers"
  ],
  "properties": {
    "script": {
      "type": "string",
      "description": "The script to convert to speech. Can be formatted with 'Speaker X:' prefixes for multi-speaker dialogues."
    },
    "seed": {
      "type": "integer",
      "description": "Random seed for reproducible generation."
    },
    "cfg_scale": {
      "type": "number",
      "description": "CFG (Classifier-Free Guidance) scale for generation. Higher values increase adherence to text.",
      "default": 1.3
    },
    "speakers": {
      "type": "array",
      "description": "List of speakers to use for the script. If not provided, will be inferred from the script or voice samples."
    }
  }
}