Sign inGet $5 free

Personaplex

Audio to audio

Liveby fal

PersonaPlex is a real-time, full-duplex speech-to-speech conversational model that enables persona control through text-based role prompts and audio-based voice conditioning.

Ask your agent

Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.

Read this paragraph aloud in a warm voice and send me the audio.
Turn my blog post into an audio file I can share.

Reviews

No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.

How it works

  1. Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
  2. Ask in your own wordsYour agent picks the right tool and fills in the details for you.
  3. Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema

Voice & audio API by fal, callable through Superpowers. It costs $0.0054 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.personaplex", or over MCP with run_tool.

Tool id
Price per call
$0.0054 / 5s clip

Call it

Also over MCP as run_tool
curl -X POST https://superpowers.tools/v1/run \
  -H 'Authorization: Bearer $SP_KEY' \
  -d '{"tool":"fal.personaplex","input":{"audio_url":""},"idempotency_key":"example-request-1"}'

Input

9 fields · 1 required
{
  "type": "object",
  "required": [
    "audio_url"
  ],
  "properties": {
    "temperature_text": {
      "type": "number",
      "description": "Text sampling temperature. Higher values produce more diverse outputs.",
      "default": 0.7
    },
    "voice_audio_url": {
      "type": "string",
      "description": "URL to a voice sample audio for on-the-fly voice cloning. When provided, the AI responds in the cloned voice instead of the preset 'voice'. 10+ seconds of clear speech recommended. Billed at 2x rate."
    },
    "temperature_audio": {
      "type": "number",
      "description": "Audio sampling temperature. Higher values produce more diverse outputs.",
      "default": 0.8
    },
    "audio_url": {
      "type": "string",
      "description": "URL to the input audio file (user's speech)."
    },
    "top_k_text": {
      "type": "integer",
      "description": "Top-K sampling for text tokens.",
      "default": 25
    },
    "prompt": {
      "type": "string",
      "description": "Text prompt describing the AI persona and conversation context.",
      "default": "You are a wise and friendly teacher. Answer questions or provide advice in a clear and engaging way."
    },
    "voice": {
      "type": "string",
      "description": "Voice ID for the AI response. NAT = natural, VAR = variety. F = female, M = male. Ignored when voice_audio_url is provided.",
      "default": "NATF2",
      "enum": [
        "NATF0",
        "NATF1",
        "NATF2",
        "NATF3",
        "NATM0",
        "NATM1",
        "NATM2",
        "NATM3",
        "VARF0",
        "VARF1",
        "VARF2",
        "VARF3",
        "VARF4",
        "VARM0",
        "VARM1",
        "VARM2",
        "VARM3",
        "VARM4"
      ]
    },
    "seed": {
      "type": "integer",
      "description": "Random seed for reproducibility."
    },
    "top_k_audio": {
      "type": "integer",
      "description": "Top-K sampling for audio tokens.",
      "default": 250
    }
  }
}