Audio & Music
Last updated
Audio mode covers spoken voice and music generation.
Use fast speech for previews and higher-quality persona speech when voice character matters.
Enter the text to speak.
Choose the voice or persona.
Adjust available delivery settings.
Generate and listen.
Download or reuse the result.
Write punctuation the way you want the speaker to pause. Short paragraphs are easier to control than one long block.
Describe genre, instrumentation, tempo, emotional arc, vocal style, and lyrical theme. If lyrics are included, separate structural sections such as verse and chorus clearly.
Companions can use generated speech as part of multimodal conversation. Their identity and memory provide character context; the selected speech pipeline provides the audible performance.
Local speech and music remain on the device unless you download, share, mint, publish, or send them through a configured workflow.
Last updated