Try Seed Audio 1.0 in the Workbench
Run this model interactively, tune parameters, and compare outputs.
bytedance-seed-audio-1-0
ByteDance Seed Audio 1.0 is a text-to-speech and audio generation model served through Fal. It synthesizes natural speech from a text prompt with control over voice, output format, sample rate, speed, volume, and pitch. Up to three reference audio clips can be supplied for voice cloning and referenced in the prompt as @Audio1, @Audio2, @Audio3, or a single reference image can be provided instead.
Example request
- Minimal
- Basic parameters
- All parameters
Fetch model details
The models endpoint returns the full model object, including itsjson_request_schema.