Skip to main content

Voice Engine

This is where you configure how your character sounds — the actual voice it uses, how expressive or stable that voice is, how often it speaks vs. types, and how it behaves in live voice calls. How to get here:
Go to your character’s dashboard → left sidebar → click Universal → click Voice Engine
The full path is: Universal → Voice Engine This section has five sub-pages shown at the top of the Voice Engine area:
  1. Overview — details about the current voice
  2. Settings — sliders to adjust how the voice sounds
  3. Frequency — how often the character uses voice vs. text
  4. Instructions — how to use audio tags for emotion/expression
  5. Preferences — how the character behaves in voice channels
At the very top of every Voice Engine page, you’ll see your currently assigned voice displayed with its provider badge (for example, Eleven v3 from ElevenLabs) and a toggle to enable or disable it.

Overview

Full path: Universal → Voice Engine → Overview This is the information page about your current voice — think of it like the voice’s “profile page.”

Voice Details

Voice Latency

These numbers tell you how fast or slow the voice generates audio:
These numbers only measure how long it takes to create the audio itself. The total time before your character speaks also includes time for the AI to write what to say — that’s separate.

Training Samples

At the bottom of the Overview page, you’ll see a list of audio files (MPEG format). These are the original recordings the voice was cloned from. Each file shows its length and file size.

Changing Your Voice

Click the Edit Voice button on this page to switch to a different voice for your character.

Settings

Full path: Universal → Voice Engine → Settings These are sliders that let you fine-tune how the voice sounds. Moving them left or right changes the character’s audio quality:
For natural conversation: Try Stability at 0.6, Style at 0.4. This gives slight variation without sounding unstable. For dramatic characters: Lower Stability (0.3) and higher Style (0.6–0.8) makes the voice more expressive and theatrical.

Frequency

Full path: Universal → Voice Engine → Frequency This controls how often your character responds with a voice message instead of a regular text message (in Discord text channels where voice is enabled). There are two modes. You can only use one at a time:

Let Character Decide (toggle)

When this toggle is ON, the character will decide on its own whether to send a voice message or a text reply based on the conversation context. The manual slider below becomes disabled.

Manual Frequency Control (slider)

When “Let Character Decide” is OFF, you drag a slider to set the exact rate:

Instructions

Full path: Universal → Voice Engine → Instructions This is a text box (up to 1000 characters) where you write instructions for how the voice uses audio tags to show emotion.

What are Audio Tags?

Audio tags are special words you put in square brackets [like this] that tell the voice model to add a specific sound or change how it speaks. The AI includes these tags in the text before it’s turned into audio. Examples of audio tags:

How to write the Instructions

In the text box, write rules that tell the AI when to use these tags. For example:
The character will follow these rules every time it generates a voice response.

Preferences

Full path: Universal → Voice Engine → Preferences These are two simple on/off toggles that control what your character does in Discord voice channels: