CyberLink Logo
Free Download

How to Use Text to Speech in PowerDirector

Turn subtitles you have already written into a spoken voice-over, without a microphone.

Updated Aug. 29, 2026 · 4 min read

No microphone, or a voice-over needed in a language you do not speak: both are jobs for Text to Speech. The thing to know before you open it is that it has no box to type into. It reads text that is already on the timeline, so the writing happens first.

⚡ Quick Answer: How to Generate a Voice-over in 4 Steps
  1. Put the words on the timeline as subtitles or text objects.
  2. Select them, then click Text-to-Speech above the timeline.
  3. Set the language to match the text before anything else.
  4. Pick a voice and a mood, then click Generate.

Steps to Generate a Voice-over

Prefer to watch? Here is the whole thing in a little over two minutes.

Step 1: Get the Words on the Timeline

There are three routes, and any of them works. From the Subtitles menu, Create Subtitles Manually lets you type the lines in.

The Subtitles menu in PowerDirector with Create Subtitles Manually highlighted, above AI Speech to Text and Import Subtitles from File.
Type them in from the Subtitles menu.

If the script already exists as a file, Import Subtitles from File brings it in instead of retyping it.

The Import Subtitles from File option highlighted in the Subtitles menu in PowerDirector.
Or bring an existing script in.

Text objects count too, so a title you have already placed can be spoken without being retyped as a subtitle. The tutorial calls this the Text feature; the tab itself is labeled Titles.

The Titles tab highlighted in the PowerDirector toolbar.
Titles on the toolbar, Text in the tutorial. Same thing.

Step 2: Select the Text and Open the Tool

With every line written, select the subtitles or text objects you want spoken. Only what you select is converted.

The subtitle list in PowerDirector with several lines of text and their start times.
Select the lines you want spoken.

Then click Text-to-Speech above the timeline.

A callout in PowerDirector reading text to speech, pointing at the control above the timeline.
Above the timeline, not in a menu.

Step 3: Set the Language First

The language menu comes first for a reason: a language that does not match the text causes generation errors rather than an accent. Set it to whatever the words actually are.

The language dropdown open in the Text to Speech window in PowerDirector, listing several languages.
Match the language to the text, not to your preference.

Step 4: Pick a Voice, and a Mood

The window opens on My Voice Profile, and on a fresh install it is an empty dashed box. Clone Voice is the way in: it hands you to AI Voice Cloning, which builds a profile from a recording of your own voice.

Once a profile exists it sits here above the stock voices, so a whole series can be narrated in one consistent voice without recording every line.

The My Voice Profile section at the top of the PowerDirector Text to Speech window, empty, with the Clone Voice button inside it
Empty until you clone a voice. Clone Voice starts that.

Underneath sits Voice Profiles, the stock voices. The filters above the grid narrow it by gender, speaking style and use case, and Clear All resets them.

The grid is not simply a list of speakers: each card is a name paired with a style, so the same voice appears more than once reading in different moods. The heart icon at the end of the filter row keeps the ones you come back to.

Filter
The gender, speaking style and use case filters above the voice grid in the PowerDirector Text to Speech window.
Gender, speaking style, use case.
Choose
The Voice Profiles grid in the PowerDirector Text to Speech window, each card showing a name and its speaking style.
One speaker, several moods, one card each.

Click a card to hear it. Reading the same line in a different mood changes the result more than changing speaker does, so it is worth listening to a few before settling.

A callout in the PowerDirector Text to Speech window reading preview voice, over the grid of voice cards.
Click a card to hear it.

Parameter Settings underneath adjusts speed and pitch, which is the fix when a voice is right but reads a little fast for the pictures.

The Parameter Settings section of the PowerDirector Text to Speech window with speed and pitch sliders.
Speed and pitch, after the voice is chosen.

This one draws on credits, and the button shows the cost. Beside a Total characters count, the Generate button carries a credit badge, so the price of the run is visible before you commit to it. A Get Credits link sits at the top of the same window.

The Result

Click Generate and PowerDirector produces the audio and returns you to the editor. The speech lands on the timeline at the positions the text occupied, so it is already in sync with the subtitles it was made from.

Generated
A completion message in the PowerDirector Text to Speech window with the generated audio waveform.
It reports back when it is done.
On the timeline
The generated speech placed on the PowerDirector timeline beneath the subtitles it was made from.
Placed where the text was.

Haven't got PowerDirector yet? No microphone needed, and no box to type into either: it reads subtitles you have already written.

Download PowerDirector Free →

FAQs

You do not type it in the Text to Speech window. It reads text that is already on the timeline, so add it first as subtitles, through Create Subtitles Manually or Import Subtitles from File, or as text objects made with the Text feature.
Because a mismatch between the language setting and the actual text causes generation errors rather than just an odd accent. Set the language to match the words before choosing anything else.

Because each card is a speaker paired with a speaking style, so one name can appear several times reading in different moods. Use the filters above the grid to narrow it by gender, style or use case, and Clear All to start again.

Yes. The Generate button carries a credit badge showing what the current run will cost, next to a running Total characters count, and there is a Get Credits link in the same window. Longer scripts cost more, so the badge is worth a glance before you generate.

Yes, through a voice profile. Clone Voice, at the top of the Text to Speech window, leads to AI Voice Cloning, which builds a profile from a recording of you. After that the profile appears in the same window as the stock voices and can read any script you put on the timeline.