Text To Speech

AUDIO GEN
Fish Audio

Fish Audio Text To Speech is an audio-generation model that turns text into spoken audio. It supports reference voices, prosody controls, and formats including WAV, MP3, and Opus, with adjustable sample rates, bitrates, and sp

Capabilities

Audio output

Supported output media

audio

Supported generation options

Formats

wav
pcm
mp3
opus

Supported tools

No tools enabled.

Pricing

TypeCreditsUnits
Generation199.50Credits per 1,000 characters

Variants

No variants available for this model.

Try This Model

Write a prompt and experiment with Text To Speech in the model experiments page. You can compare it with other models side by side.