Appearance
Dialogue
Dialogue turns a script into one audio conversation, with a different voice for each character. You can redo any single line without making the whole thing again.
Make a conversation
- Open Dialogue in the sidebar. A sample script is filled in.
- Write your lines. For each line:
- choose the speaker in the first menu (choose + New speaker… to add one)
- type what they say
- choose an emotion
- choose a volume
- Choose + Add line to add more lines. Use the arrows to move a line up or down, and × to delete it.
- On the right, under Cast, choose a voice for each speaker. Choose the play button to hear it.
- Set Model, Language and Speed if you need to.
- Choose Generate conversation.
The result plays when it's ready. It shows each line with its start time.
Paste a script instead
- Choose Paste a script at the top.
- Paste lines like
Host: Welcome to the show!, one line per turn. - Choose Use this script.
How the script is read:
- Text before the first
:becomes the speaker (up to 40 characters). - A line without a speaker is added to the line before it.
- If the script starts without a speaker, that first line goes to "Narrator".
Fix one line
In the result, each line has two buttons:
- Redo line: say the same line again.
- Edit: change the text, emotion or volume, then choose Redo with changes.
Only that line is made again. The rest of the conversation stays as it was. The editor above keeps the new wording.
Options
- Emotion per line: Natural (the default), Neutral, Happy, Sad, Angry, Fearful, Disgusted or Surprised.
- Volume per line: 50%, 75%, 100%, 125%, 150% or 200%.
- Add a sound…: puts a tag like
(laughs)where your cursor last was, or at the end of the last line. Choices: Laughs, Sighs, Coughs, Clears throat, Gasps, Sniffs, Groans, Yawns. - Model: Cinara Speech (1 credit per character) or Cinara Speech HD (2 credits per character).
- Language: Detect automatically or one of 37 languages (the same list as Speech).
- Speed: 0.50× to 2.00×, for the whole conversation.
- Voices: your workspace's cloned and designed voices come first, then Cinara's voices. Each new speaker gets a different voice to start with.
- Your workspace's pronunciations apply to every line.
Subtitles
- In the result, choose SRT subtitles or VTT subtitles.
- In the page's History, choose SRT or VTT on any finished conversation.
- Subtitles come from the line timings and include speaker names.
History on the page
Your last 30 conversations are listed. Choose Open to load a conversation, its voices and settings back into the editor and the result.
Limits
- Up to 200 lines.
- Up to 2,000 characters per line.
- Up to 50,000 characters in total.
- Speaker names up to 40 characters.
- Every line needs a speaker and some text before you can generate.
Credits
- Cinara Speech: 1 credit per character of all lines.
- Cinara Speech HD: 2 credits per character.
- Redo line and Redo with changes charge again for that one line's characters.
- If generation or a redo fails, those credits are refunded.
Tips
- Keep each line to one speaker's turn. Short turns sound more natural.
- Use emotions sparingly. Natural often works best.
- Use volume to make an aside quieter or a shout louder, instead of redoing the voice.
- Preview voices before you generate so each character sounds distinct.
Related
- Speech: one voice.
- Voice Library: find voices for your cast.
- Voice Cloning and Voice Design: make your own voices.
- Dubbing: translate an existing recording.