Text-to-dialogue Two speakers ElevenLabs Commercial use
Dialogue Creation

Dialogue Creation
– Text-to-dialogue
model in Phygital+

Dialogue Creation turns two scripts into one conversation. Each speaker gets their own text field and their own voice, and the model renders the exchange as a single scene with matched prosody rather than two monologues cut together.

It runs the Eleven v3 text-to-dialogue workflow, with ten voice presets per speaker and support for custom Voice IDs built elsewhere on the canvas. Recurring characters keep the same voice across every scene you generate.

Dialogue Creation

At a glance Dialogue Creation

Built for

Scripted ads, game scenes, UX prototypes, interview-style content, and role-play audio

Ad, game, and product teams that need a conversation rather than a monologue

Powered by
ElevenLabs
Category
AI audio — multi-speaker dialogue
Modalities
Text → audio
Speakers
Two, with independent voices
Default model
eleven_v3
Voice presets
10 per speaker, plus custom Voice IDs
Prompt limit
5000 characters
Access
All Phygital+ plans

Why use Dialogue Creation

Two speakers, two voices, one rendered scene with prosody matched across the exchange.

1

A conversation, not two monologues

The model generates both sides of the exchange together, matching prosody and pacing across the turn. Reactions land where they should instead of sounding like two people recorded in separate rooms.

2

Two voices, chosen separately

Each speaker gets their own preset or their own custom Voice ID, so the two characters stay clearly distinct. Ten presets cover most casting needs before you have to build anything.

3

Characters that persist

Paste a Voice ID from Voice Creation or Voice Clone and the same character sounds the same in every scene you generate. That is what makes a series of clips feel like one production.

4

Scripts you can restructure

Split phrases with a semicolon and the dialogue alternates in the order they appear. Restructuring a scene is an edit to two text fields, not a re-record.

5

Reproducible takes

Fix the seed and the take is reproducible, so refining one line does not reshuffle the rest of the scene. Iterating on dialogue stops being a gamble.

Built for ad, game, and product teams

Dialogue Creation fits wherever a script has two voices in it: ads, game scenes, UX prototypes, and interview-style content.

Dialogue Creation
Dialogue Creation
Dialogue Creation
Dialogue Creation
Dialogue Creation
Dialogue Creation
Dialogue Creation
Dialogue Creation

How to use Dialogue Creation

Split the script between two speakers, assign a voice to each, and generate.

Dialogue Creation
Step 1

Split the script

Put the first speaker's lines in Text (1st person) and the second speaker's in Text (2nd person). Separate phrases with a semicolon, and keep the phrase count roughly balanced so the exchange alternates naturally.

Dialogue Creation
Step 2

Assign the voices

Choose a preset for each speaker, or paste a custom Voice ID to override it. Voice IDs come from Voice Creation for invented characters or Voice Clone for real ones, and can be reused across every scene in a series.

Dialogue Creation
Step 3

Generate, iterate & export

Fix a seed so re-running one line does not change the rest of the scene. Chain the audio into a video node for an ad or animatic, or export it with commercial rights on your plan.

Pricing and access

One subscription across 30+ AI models, with no per-tool credit balances or separate signups. Credit cost per generation is shown live in the node before you run it.

Free

Try Dialogue Creation free

500 weekly credits to test the node and hear the output quality. No credit card required.

  • 500 weekly credits
  • Access to 30+ models
  • Personal use
Join as Free
Pro

Professional access

45,000 monthly credits with cheaper per-credit pricing, commercial rights, and full downloads.

  • 45,000 monthly credits
  • Commercial use license
  • 15% cheaper credits
Join as Pro
Team

Team collaboration

90,000 monthly credits, up to 10 seats, shared workspace, centralized billing, and priority support.

  • 90,000 monthly credits
  • Up to 10 seats
  • Centralized billing
Join as Team
Enterprise

Enterprise scale

210,000 monthly credits, API and integration support, unlimited seats, and dedicated management.

  • API & integration support
  • Unlimited seats
  • Dedicated manager
Join as Enterprise

Technical Specifications

Full specs for Dialogue Creation in Phygital+.

Powered by
ElevenLabs
Input
Two text fields, one per speaker
Output
Single rendered dialogue track
Models
eleven_v3, default
Custom voices
Voice ID overrides the preset
Reproducibility
Seed-controlled
Category
Text-to-dialogue
Delimiter
Semicolon between phrases
Speakers
Two
Voice presets
10 per speaker
Prompt limit
5000 characters
Requirement
Both speaker texts must be filled

Dialogue Creation vs Other AI models

How Dialogue Creation fits alongside the other audio nodes in Phygital+.

Model

Dialogue Creation

Two-speaker scenes rendered as a single coherent conversation

Sound Creation

Single-voice narration and text-described sound effects

Voice Creation

Builds new synthetic voices from a written description

Voice Clone

Clones an existing voice from audio samples into a reusable Voice ID

Best for Scripted conversations, interviews, role-play Voice-over, narration, sound effects Inventing a voice that does not exist yet Keeping a real speaker's voice
Input Two scripts, one per speaker One block of text A written voice brief Audio samples
Voices Two, selected independently One Not applicable Not applicable
Delivery Prosody matched across the exchange Sequential takes Not applicable Not applicable

Use Dialogue Creation Alongside other AI models In Phygital+

Chain scripting, voice design, dialogue, and video into one repeatable workflow.

Dialogue Creation

Chain with a text model

Draft the conversation in a text node, split it between the two speaker fields, and generate the scene in one pass.

Dialogue Creation

Chain with Voice Creation

Design distinct character voices from written briefs, then paste each Voice ID into its speaker slot.

Dialogue Creation

Chain with Voice Clone

Clone two real speakers from recordings and reuse those Voice IDs for every scene in the series.

Dialogue Creation

Chain with a video model

Generate the visuals in a video node and lay the dialogue track underneath for ads, demos, or animatics.

Start Generating!

Everything you need to know about Dialogue Creation in Phygital+.

FAQ about Dialogue Creation

Dialogue Creation
What is Dialogue Creation?

Dialogue Creation renders a two-speaker conversation as a single audio track using ElevenLabs voices. It runs the Eleven v3 text-to-dialogue workflow, which matches prosody across the exchange rather than generating each line in isolation and stitching them together.

How do I write the script?

Put the first speaker's lines in Text (1st person) and the second speaker's lines in Text (2nd person). Separate individual phrases with a semicolon. The dialogue alternates between speakers in the order the phrases appear, so keeping the phrase count balanced between the two fields produces a more natural back-and-forth.

Can I use custom voices?

Both. Ten voice presets are available per speaker, covering narration, news, character, educational, and conversational reads. If you paste a custom Voice ID into a speaker's custom voice field, it overrides the preset for that speaker.

Where do custom Voice IDs come from?

Two ways. Voice Creation generates candidates from a written brief such as age, gender, timbre, pace, and accent, and returns several options to compare. Voice Clone builds a Voice ID from audio samples of a real speaker. Either way you end up with a Voice ID you can paste in and reuse.

How many speakers does it support?

Two. For scenes with more speakers, generate them in passes and assemble the result, or use a video model with native multi-character audio for scenes where the voices need to overlap.

Do I have to fill in both speakers?

Both speaker texts are required, so the node will not generate a one-sided scene. If you need a single voice, use Sound Creation instead.

Can I reproduce a take?

Fix the seed. The same seed with the same scripts and voices reproduces the same take, which matters when you are iterating on one line and do not want the rest of the scene to change underneath you. The prompt limit is 5000 characters.

How much does Dialogue Creation cost in Phygital+?

Usage is credit-based and included in every plan, from Free up to Enterprise. The credit cost per generation is shown in the node before you run it, so there is no separate ElevenLabs subscription to manage.

Can I use the audio commercially?

Outputs created on any paid plan come with commercial-use rights. Make sure you have permission for any voice cloned from a real person before publishing.

Explore more AI models In Phygital+

Browse the full catalog of 30+ AI models available in Phygital+.

Dialogue Creation ElevenLabs v3 FLUX Gemini Omni GPT-Image-2 Hailuo 3.0 Ideogram 3 Kling 3.0 Kling Omni Kling Omni Image Krea AI LTX Video Luma Magnific Upscale Midjourney Nano Banana Pro OmniHuman Qwen Image Recraft V4 Reve Runway 4.5 Seedance 2.5 Seedream 5 Sound Creation Topaz Upscale Upscale Video Veo 3.1 Voice Clone Voice Creation WAN 2.5 Wan Video 2.1

Ready to work Faster with Dialogue Creation?

Join 100+ teams using Phygital+ – every model in one workspace

Try Phygital+ free
bg