Voice design Text-to-voice ElevenLabs Commercial use
Voice Creation

Voice Creation
– Voice design
model in Phygital+

Voice Creation invents a speaker from a written brief. Describe age, gender, timbre, pace, emotion, accent, and style, and the node returns several audio candidates in one run so you can hear the options side by side before choosing.

The voice you pick becomes a reusable Voice ID that works in Sound Creation and Dialogue Creation, so a brand voice or a recurring character stays consistent across every clip on the canvas.

Voice Creation

At a glance Voice Creation

Built for

Character voices for games, brand voices for campaigns, narrators for explainers, and assistant voices

Brand, game, and localisation teams inventing a speaker instead of casting one

Powered by
ElevenLabs
Category
AI audio — voice design
Modalities
Text description → audio candidates
Description length
20 to 1000 characters
Candidates per run
Several, typically three
Default model
eleven_ttv_v3
Reuse
Voice ID works in Sound Creation and Dialogue Creation
Access
All Phygital+ plans

Why use Voice Creation

Describe a speaker, compare several candidates, and reuse the one you pick everywhere.

1

A voice from a sentence

No recording, no casting, no rights to clear. A sentence describing age, timbre, pace, and accent is enough to get a usable speaker, which changes what is worth prototyping.

2

Several candidates per run

One run returns several candidates rather than a single take, so you audition options instead of accepting the first result. Each sounds meaningfully different, and listening to all of them before choosing is worth the minute it takes.

3

Designed once, reused everywhere

The Voice ID you pick works in Sound Creation and Dialogue Creation and can be reused indefinitely. A brand voice or a recurring character stays the same speaker across a whole campaign.

4

Explore around a brief

Same brief, different seed, different candidates. It is a cheap way to explore a range around one idea without rewriting the description every time.

5

Audition on your own copy

Write the description in any language, and add your own script to hear the voice reading real copy rather than a generic sample. What you audition is what you will ship.

Built for brand, game, and localisation teams

Voice Creation fits at the start of an audio workflow, where you need a speaker who does not exist yet and a casting call is not an option.

Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation
Voice Creation

How to use Voice Creation

Describe the voice you want, generate candidates, and reuse the Voice ID downstream.

Voice Creation
Step 1

Describe the voice

Write 20 to 1000 characters covering age, gender, timbre, pace, emotion, accent, and style. Around 50 to 200 characters gives the most focused result, and style tags such as whispery or news anchor sharpen it further. Any language works.

Voice Creation
Step 2

Set the preview and parameters

Add your own script in the optional text field to hear the voice on real copy, or enable Auto-generate text and let the model write appropriate lines. Set Loudness for intensity and a seed if you want the run reproducible.

Voice Creation
Step 3

Compare, pick & reuse

Listen to every candidate before deciding, since each one sounds different. Extract the Voice ID from the one you want and paste it into Sound Creation or Dialogue Creation, where you can reuse it indefinitely.

Pricing and access

One subscription across 30+ AI models, with no per-tool credit balances or separate signups. Credit cost per run is shown live in the node before you start it.

Free

Try Voice Creation free

500 weekly credits to test the node and hear the output quality. No credit card required.

  • 500 weekly credits
  • Access to 30+ models
  • Personal use
Join as Free
Pro

Professional access

45,000 monthly credits with cheaper per-credit pricing, commercial rights, and full downloads.

  • 45,000 monthly credits
  • Commercial use license
  • 15% cheaper credits
Join as Pro
Team

Team collaboration

90,000 monthly credits, up to 10 seats, shared workspace, centralized billing, and priority support.

  • 90,000 monthly credits
  • Up to 10 seats
  • Centralized billing
Join as Team
Enterprise

Enterprise scale

210,000 monthly credits, API and integration support, unlimited seats, and dedicated management.

  • API & integration support
  • Unlimited seats
  • Dedicated manager
Join as Enterprise

Technical Specifications

Full specs for Voice Creation in Phygital+.

Powered by
ElevenLabs (text-to-voice)
Input
Voice description, 20–1000 characters
Output
List of audio candidates
Default model
eleven_ttv_v3
Loudness
-1 to 1, 0.4–0.7 typical
Downstream use
Voice ID in Sound Creation and Dialogue Creation
Category
Voice design
Optional input
Preview script, 100–1000 characters
Models
eleven_ttv_v3, multilingual TTV v2
Auto-generate text
Optional, recommended when no script is supplied
Reproducibility
Seed-controlled

Voice Creation vs Other AI models

How Voice Creation fits alongside the other audio nodes in Phygital+.

Model

Voice Creation

Invents a voice from a written brief and returns candidates to compare

Voice Clone

Clones an existing voice from audio samples

Sound Creation

Generates narration and sound effects from text

Dialogue Creation

Renders two-speaker scenes with independent voices

Best for Inventing a speaker who does not exist Keeping a real speaker's voice Voice-over and sound effects Scripted conversations
Input A written description of the voice Audio samples A block of text Two scripts
Output Several audio candidates to choose from A reusable Voice ID A single audio track A rendered dialogue track
Source material None — no recording required None needed None needed None needed

Use Voice Creation Alongside other AI models In Phygital+

Design a voice once, then reuse it across every audio node on the canvas.

voicecreationall

Chain into Sound Creation

Paste the chosen Voice ID into Sound Creation and every narration clip from then on uses the same voice.

voicecreationall

Chain into Dialogue Creation

Design two contrasting voices, then assign one to each speaker in a two-person scene.

voicecreationall

Chain with a text model

Write the voice brief with a text node when you are exploring a range of character options at once.

Chain into video

Generate the visuals in a video node and lay the finished voice track underneath.

Start Generating!

Everything you need to know about Voice Creation in Phygital+.

FAQ about Voice Creation

Voice Creation
What is Voice Creation?

Voice Creation generates new synthetic voices from a written description. You describe the speaker you want and the node returns several audio candidates in a single run, typically three, so you can compare them before committing to one voice identity.

What goes in the voice description?

Age, gender, timbre, pace, emotion, accent, and style, in a description between 20 and 1000 characters. Around 50 to 200 characters tends to give the most focused results. Style tags help: whispery, news anchor, cheerful, calm. You can write the brief in any language.

What does a good brief look like?

A professional male news anchor, 40 years old, clear and authoritative. A young female voice, 25 years old, cheerful and energetic, slight British accent. An elderly grandmother, 75 years old, warm and caring, speaks slowly. Short and specific beats long and hedged.

Can I preview the voice on my own script?

Yes. Add your own script in the optional text field, between 100 and 1000 characters, and the candidates will read it. If you leave it empty, enable Auto-generate text so the model writes contextually appropriate lines. The node tips recommend auto-generated text for the best results.

How do I reuse the voice I chose?

Extract the Voice ID from the sample you picked and paste it into the custom voice field of Sound Creation or Dialogue Creation. Voice IDs are reusable indefinitely, so a character or brand voice designed once stays consistent across every clip you generate afterwards.

What if none of the candidates work?

Change the seed and re-run. The description stays the same but the candidates differ, which is the fastest way to explore a range around the same brief. Fix the seed when you want a run to be reproducible.

What does Loudness do?

Loudness sets the intensity and volume of the generated voice, from -1 to 1. Around 0.4 to 0.7 works for most narration. Push it higher for energetic reads and lower for intimate or restrained delivery.

How much does Voice Creation cost in Phygital+?

Usage is credit-based and included in every plan, from Free up to Enterprise. The credit cost per run is shown in the node before you start it, so there is no separate ElevenLabs subscription to manage.

Can I use the voice commercially?

Outputs created on any paid plan come with commercial-use rights. Since the voice is generated from a description rather than cloned from a person, there is no source recording to clear.

Explore more AI models In Phygital+

Browse the full catalog of 30+ AI models available in Phygital+.

Dialogue Creation ElevenLabs v3 FLUX Gemini Omni GPT-Image-2 Hailuo 3.0 Ideogram 3 Kling 3.0 Kling Omni Kling Omni Image Krea AI LTX Video Luma Magnific Upscale Midjourney Nano Banana Pro OmniHuman Qwen Image Recraft V4 Reve Runway 4.5 Seedance 2.5 Seedream 5 Sound Creation Topaz Upscale Upscale Video Veo 3.1 Voice Clone Voice Creation WAN 2.5 Wan Video 2.1

Ready to work Faster with Voice Creation?

Join 100+ teams using Phygital+ – every model in one workspace

Try Phygital+ free
bg