Long after an audience forgets the words, they remember how the performance made them feel. The same line can reassure, persuade, frighten, or inspire, not because the words change but because the delivery does. Getting that delivery right is what makes text-to-speech audio sound like it belongs in your content.
This GlobalLink Voice release puts that delivery in your hands. On top of the pitch, speed, and pronunciation controls you already use, you can now shape the expression of synthetic voices, down to individual words when necessary.
Read on for what’s new.
What Is GlobalLink Voice?
GlobalLink Voice is the leading enterprise solution for AI voice generation, turning text into multilingual audio. Users can produce natural text-to-speech (TTS) across dozens of languages, with speech-to-speech available through the API. One voice can speak several languages in the same file, so a character or brand narrator stays recognizable from one market to the next.
What’s New in GlobalLink Voice
Set the tone, emotion, and pace of any read
GlobalLink Voice now lets you direct how a voice performs, giving every word the precise performance you intend. Three complementary layers of expressive voice control work together, from setting a mood across an entire segment to directing the delivery of a single word. If you produce entertainment, e-learning, or corporate communications that need more than a flat read, this update is built for you.
Styles for whole segments
Styles set the emotional or tonal mode of a segment. Pick from a curated list of expressions, including options like Angry, Excited, Happy, Sad, and Scared. One text-to-speech block can move across various expressions, all while maintaining the same voice.
Inline styles for words and phrases
Inline styles bring custom expressions down to the word or phrase level. Highlight a piece of text, right-click, and apply a style to just that portion, without affecting the rest of the segment.
Inline prompts for free-text direction
Inline prompts offer the most precise control over expression. You can type any plain language direction and insert it as a tag right in the text, such as “pirate voice,” “whisper slowly,” or “like a news anchor.” The voice interprets your instruction and adjusts its delivery to match.
For teams producing at global scale, expressive voice control enables:
- Granular emotion control for large audio projects, without a separate post-editing pass.
- Consistent brand tone across channels and formats, so every market hears the same voice.
Get files out faster, and named your way
Two smaller changes round out the latest GlobalLink Voice update, both aimed at getting finished multilingual audio out of the platform with less manual work.
Custom naming for single-file downloads
The naming structure you set in a project now applies to every export type, including single-file downloads, so one convention carries across all of your exports.
Quick download from the project list
You can now download audio straight from the project list, in a couple of clicks, without opening the project first. It’s the quickest way to grab a finished file when that’s all you need.
Next Steps
If you’re an existing GlobalLink Voice user:
- Apply a style to a segment that needs a specific mood and hear the difference without changing the voice.
- Use an inline style on a single word or phrase to place emphasis exactly where you want it.
- Write an inline prompt and drop it into your script to direct the delivery.
- Set a file-naming convention, then download a single-file export to see it applied.
- Grab a finished file straight from the project list with quick download.
New to GlobalLink Voice? Visit the GlobalLink Voice page to see how multilingual audio production works from script to finished files, or get in touch about a pilot for your entertainment, e-learning, marketing, or corporate communications work.
Want to Learn More?
GlobalLink is TransPerfect’s language intelligence platform, orchestrating AI, human expertise, and enterprise content systems across markets, channels, and languages. Expressive voiceover is one part of that larger stack, working alongside translation management, multimedia localization, and the integrations that connect it all to the systems your teams already run.
Quick Answers
What is text-to-speech?
Text-to-speech (TTS) is technology that converts written text into spoken audio. Instead of recording a person reading a script, you type or import the text and the system generates the voiceover. GlobalLink Voice uses text-to-speech to produce natural-sounding speech in dozens of languages.
What’s the difference between styles, inline styles, and inline prompts?
Styles apply a curated tonal or emotional preset, such as Excited or Sad, to a whole segment. Inline styles apply those same presets to a selected word or phrase. Inline prompts let you type your own free-text direction and insert it as a tag for the voice to interpret.
What is a custom or cloned brand voice?
A custom voice is a voice model built specifically for you, and a cloned brand voice recreates a recognizable voice your audience already associates with your brand. GlobalLink Voice supports custom voice creation, including cloning a recognizable brand voice, so your audio can sound distinctly yours across content and markets.