Skip to main content
Synthesize text to speech in WAV or MP3. The CLI picks the format from the model: mist and mistv2 output MP3, while coda and mistv3 default to WAV. Deprecated Arcana identifiers also default to WAV. Use --format to override.
Demo of the rime tts command

Required flags

Optional flags

Deprecated Arcana compatibility flags

These flags remain for existing Arcana integrations. Do not use Arcana for new requests.

mist/mistv2/mistv3 flags

mist/mistv2 flags

Only mist and mistv2 support these flags; mistv3 does not.

Examples

Supported languages by model

The mist and mistv2 models default to MP3. coda, mistv3, and the deprecated Arcana identifiers default to WAV. Use --format to override.