Creating Text-to-Speech Audio via Asset Manager

Create text-to-speech in the production Create panel or from the Audio area in Asset Manager.

Creating Text-to-Speech Audio

Create text-to-speech in the production Create panel or from the Audio area in Asset Manager.

What you need first

Before you can generate speech from text, you need at least one saved voice.

If you do not have one yet:

  1. Open Asset Manager
  2. Go to Voices
  3. Click Create Voice
  4. Design and save a voice first

Production Create panel

When working in production, open the Create panel and choose the Speech tab. It has two fields:

  • Dialogue: the exact words to generate
  • Performance Direction: delivery guidance such as pacing, emphasis, energy, or pauses

Choose a saved voice, fill in the fields, then select Generate. Existing audio is not changed automatically; generating creates a new result for review.

Asset Manager workflow

  1. Open your project
  2. Open Asset Manager
  3. Go to the story's Audio folder
  4. Click Create Audio
  5. Choose Text to Speech
  6. Choose the saved voice and model
  7. Enter the text you want spoken
  8. Generate the audio

What happens after generation

The generated speech is saved back into the story as a reusable audio asset. You can then place it on the timeline, review it, or use it in a larger voice workflow.

Changing existing generated speech

To change wording or performance direction, generate a new speech result and replace or review the clip you want to use. Do not promise an in-place rewrite of an already generated audio asset unless that control is visible in the current workflow.

Try it in Ciaro Pro

Open the app and follow along with the steps in this guide.

Your vision. Every frame.

Start free. Scale when the production is ready.

Creating Text-to-Speech Audio via Asset Manager | Ciaro Pro Help