Skip to main content
Model family1

Seed Audio

ByteDance’s speech model — a script read in a voice you describe or attach

Seed Audio is ByteDance’s speech line. Seed Audio 1.0 is sold here for text-to-speech: it reads a script in a voice you describe, or in a voice from a recording you attach with the speaker’s permission, and it is priced by the length of the script.

Strengths and limits

Strengths and limits

Key Features

  • Speech in a voice from a recording you are entitled to use.
  • Priced by the length of the script.

Where it falls down

  • A recording must be of your own voice or used with the speaker’s permission.
  • The number of reference recordings per request is capped.
Model1

Model

Capability tables and prices are read from the live model catalogue.

Seed Audio — Model family
ModelToolsDurationResolutionAudioCredits
Seed Audio 1.0Text-To-Speech——from 1 credit · ≈ $0.06
Tools1

Tools

Frequently Asked Questions

ByteDance’s speech model. ByteDance describes it shaping a voice from a text description, an authorised reference sample, or both.

Ready to Get Started?

Every tool shows its price in credits before you run it.