Prompt naturally
Write “bright, curious, and close to the listener” instead of configuring a synthesizer.
Voice design lets you begin with a written description instead of a fixed voice list. Describe age, energy, texture, mood, pacing, or storytelling context, then refine the result by listening.
Write “bright, curious, and close to the listener” instead of configuring a synthesizer.
Combine voice identity with a separate direction for the current line.
Keep the description, seed, and controls together as a preset.
Start with the speaker’s general character, add the vocal texture, and finish with the intended setting. For example: “A thoughtful young adult with a clear, gentle tone, speaking like a late-night radio host.” Avoid imitating a real person without permission or requesting a deceptive impersonation.
The voice description defines the broad sound. The performance direction explains how this particular passage should be delivered. Keeping them separate makes a voice reusable across calm introductions, excited announcements, and reflective endings.
A fixed seed improves repeatability, but generated speech can still vary with text, punctuation, and other controls. Treat the seed as part of a complete recipe rather than a universal voice identifier.
Create fictional or properly authorized voices. Do not design speech to mislead listeners about who is speaking, fabricate endorsements, bypass identity checks, or cause harm. When a listener could reasonably mistake synthetic speech for a real person, disclose that the audio is AI-generated.
QUICK ANSWERS
No. Voice design begins from a written description.
You can save the complete settings and seed as a reusable preset.
No. Use fictional descriptions or voices you are authorized to create, and avoid deceptive impersonation.
Begin with 50 credits. No card required.