Skip to main content

Try Seed Audio 1.0 in the Workbench

Run this model interactively, tune parameters, and compare outputs.
Model ID: bytedance-seed-audio-1-0 ByteDance Seed Audio 1.0 is a text-to-speech and audio generation model served through Fal. It synthesizes natural speech from a text prompt with control over voice, output format, sample rate, speed, volume, and pitch. Up to three reference audio clips can be supplied for voice cloning and referenced in the prompt as @Audio1, @Audio2, @Audio3, or a single reference image can be provided instead.

Example request

Use the Workbench as a request builder: configure parameters for this model in the UI, then open the API tab to copy the exact cURL or Python call.

Fetch model details

The models endpoint returns the full model object, including its json_request_schema.

Request parameters

Required parameters

Optional parameters