Direct the performance,not just the words.
Tell Drama 3 how a line should feel, in your own words. Add a sound, emotion, or pause where you want it, and control exactly which words change.
Try Drama 3One line, five directions
InputI'm fine. Really. Go on without me.
Try this lineThree ways to direct, one script
Mix them freely in the same line. Every direction is plain text you can read, edit, and reuse.
[warm and relaxed, like talking to an old friend] It's been a while.
Describe it in your own words
Try natural-language directions, beyond the preset tags: emotion, intent, pacing, who they're talking to.
Alright, [sigh] let's try again.
Add a sound, emotion, or pause
Put a sigh, a laugh, a breath, or a short pause where you want it, and nowhere else.
I'm fine. Really. Go on without me.
Control selected words
Control exactly which words change
Select a few words and pick a scoped tag. Paired tags apply only to the words you select.
[像是在强忍着眼泪,却努力笑着]
[conteniendo las lágrimas, pero sonriendo]
[涙をこらえながら、無理に笑って]
I'm fine. Really. Go on without me.
Direct in your own language
Write tags and directions in English or your own language. The performance follows the direction, whatever language it's written in.
I was home. Asleep.
Auto Tag[defensive, slightly uneasy tone] I was home. [pause] Asleep.
Start with Auto Tag
Need a starting point? Auto Tag adds directions to your script. Keep what works and rewrite the rest.
Write the scene. Auto Tag directs it.
Give each speaker a voice, write plain dialogue, and let Auto Tag add the directions. Drama 3 performs the whole scene in one pass.
Detective: Adrian · Suspect: Delia · Generated in one pass
Int. Interrogation room – Night
Detective
You were there that night.
Suspect
defensive, slightly uneasy tone
I was home. pause Asleep.
Detective
skeptical, questioning tone
Then why is your coat still wet?
Suspect
It rained.
Drama 3 FAQ
Drama 3 is a text-to-speech model from Fish Audio built for direction. Alongside your script you can describe how a line should be performed, add sounds and pauses, and limit an effect to specific words.
Put a direction or a sound in square brackets where it should happen, for example [sigh] or [barely holding it together]. To change only some words, select them and pick a scoped tag such as Whisper or Slower from the Tags menu; the editor wraps them in a paired tag.
Yes. Write tags and directions in English or your own language.
Yes. Add speakers in the Text to Speech editor, give each one a voice, and Drama 3 generates the whole dialogue together.
Drama 3 is still improving. Results vary between takes, so generate a few and pick the best. We will keep updating the model while it is in preview.
Yes. Call the Text to Speech API with the model header set to drama-3-preview. The same tags work in API requests.