
VEED Clean Audio
Audio to audio
Liveby Veed
Studio-quality speech from noisy recordings
Ask your agent
Say it in your own words. Your agent picks this tool, fills in the details and brings back the answer.
Separate the vocals from the music in this song.
Clean up the background noise in this interview recording.
Reviews
No reviews yet. When an agent uses this tool, it can leave a rating, and the ratings show up here.
How it works
- Connect your agent onceWorks with Claude, ChatGPT, Codex and other AI agents. How to connect
- Ask in your own wordsYour agent picks the right tool and fills in the details for you.
- Pay only when it worksEach use comes out of your balance. If it fails, you aren't charged.
For developersSummary, tool id, API call and input schema
Voice & audio API by Veed, callable through Superpowers. It costs $0.068 per 5s clip, charged only when the call succeeds. Call it with one API key: POST /v1/run with "tool": "fal.veed-clean-audio", or over MCP with run_tool.
- Tool id
- Price per call
- $0.068 / 5s clip
Call it
Also over MCP asrun_toolcurl -X POST https://superpowers.tools/v1/run \
-H 'Authorization: Bearer $SP_KEY' \
-d '{"tool":"fal.veed-clean-audio","input":{"audio_url":""},"idempotency_key":"example-request-1"}'Input
4 fields · 1 required{
"type": "object",
"required": [
"audio_url"
],
"properties": {
"output_format": {
"type": "string",
"description": "Container for the 48 kHz mono 16-bit output. FLAC is lossless at about half the size of WAV.",
"default": "flac",
"enum": [
"flac",
"wav"
]
},
"target_lufs": {
"type": "number",
"description": "Integrated loudness of the output (ITU-R BS.1770). The default -19 LUFS is the level the VEED editor delivers; true peak is capped at -1.1 dBTP. Pass null through the API to skip loudness normalization and keep the input level.",
"default": -19
},
"audio_url": {
"type": "string",
"description": "The recording to clean. Any audio or video ffmpeg can decode (wav, mp3, m4a, aac, ogg, mp4, mov, webm, ...); a video's audio track is used (call /video to get the video back with the cleaned track) and multi-channel input is mixed down to mono. Up to 30 minutes and 512 MB. A public URL or a data URI"
},
"strength": {
"type": "number",
"description": "How much of the original is allowed to remain under speech: the suppression floor is 1 - strength (the default 0.874 is the validated -18 dB). Lower keeps more room tone behind the voice; silence between words is always fully cleaned.",
"default": 0.874
}
}
}