이 페이지는 아직 한국어로 번역되지 않아서 영어판을 보여 드려요.

워크플로

AI Audio Drama Generator

An audio drama is a script performed by several people. Every general-purpose text-to-speech tool gives you one performer, which means the hard part — casting, directing, keeping characters distinct — is still yours. This is the page for the tool that does the hard part.

이 방식이 여기서 통하는 이유

Casting happens automatically

Paste a scene and CastDub identifies who is speaking, in whatever format you already write — Name: prefixes, screenplay blocks, or prose with dialogue tags. Each speaker gets a proposed voice based on how you described them. You override what you disagree with, and the override sticks for the whole project.

Direction is per line, not per character

Drama lives on the moment a character's tone changes. Emotion is a property of the line, so a character can be level for eight lines and break on the ninth. That single mechanism is the difference between a read-through and a performance.

Characters stay themselves within a project

Each character in a project carries a name, a persona, a library character and a default emotion, and no two characters in the project share a voice until the language runs out of them. Adding someone later does not reshuffle anyone else, so the audio you have already generated stays valid.

Exports fit a real workflow

Take a mixed MP3 for a quick share, per-character stems for scoring and sound design, or an SRT if you are publishing with captions. Nothing forces you to finish inside one tool.

이런 작업에 어울리는 보이스

Kestrel

Kestrel

Low, weary swordswoman who has already lost this war once

여성성인미국 영어드라마틱거친 질감
Ashford

Ashford

Velvet-voiced villain who apologises before he ruins you

남성성인영국 영어빌런부드러운
Wren

Wren

Bedtime-story narrator, slow and warm, never wakes anyone up

여성성인미국 영어내레이션따뜻한
Torvald

Torvald

Knight commander who gives orders that men actually follow

남성성인영국 영어영웅판타지
Juno

Juno

Podcast host who thinks out loud and never reads a script

여성청년미국 영어팟캐스트대화체

라이브러리 전체 보기 →

이렇게 하면 돼요

  1. Bring your script

    Paste it in. Any consistent format works; you do not have to reformat a screenplay into a template first.

  2. Review the cast

    Auto-casting proposes a voice per character. Audition alternatives back to back on the same line, which is the only comparison that actually settles a casting question.

  3. Set defaults per character

    Give each character a resting emotion. Most of a scene runs on the default, so getting it right means you only direct the exceptions.

  4. Direct the turns

    Mark the specific lines where something changes — the confession, the threat, the joke that is not a joke. Leave everything else neutral so those lines have something to land against.

  5. Render and listen end to end

    Play the whole scene without stopping. Problems in an audio drama are almost always pacing problems, and pacing is invisible line by line.

  6. Export in the shape you need

    A mixed MP3 on any plan, subtitles from Plus, stems plus subtitles on Pro. If you plan to add music and effects, take the stems — a pre-mixed dialogue track is very hard to score under.

Why audio drama is having a moment

Audio fiction has grown steadily for a decade, driven by the same things that grew podcasting: commuting, chores, and the fact that audio is the only medium you can consume while doing something else. What changed recently is production cost. A drama needs several performers, and several performers means scheduling, studio time and a producer, which put the format out of reach for the people who most wanted to make it.

The result was a genre dominated either by well-funded productions or by solo narrators reading everything themselves. Neither is what audio drama is. The format's whole appeal is hearing a scene between people, and a single reader doing four voices is a compromise that listeners tolerate rather than enjoy.

Generated performance changes the economics rather than the craft. Casting, direction, pacing and writing all still matter — arguably more, because they are now the only things separating a good production from a bad one.

The four decisions that make a drama work

First, casting for separation. Listeners identify speakers primarily by pitch band and speaking rate, with no visual information at all. Two characters in the same band will blur no matter how different their personalities are on the page. Spread your cast deliberately and the scene becomes legible.

Second, restraint in direction. The temptation is to mark every line with an emotion, and the result is uniform mush. A scene with two directed lines and twenty neutral ones is far more affecting than a scene where everything is turned up.

Third, pacing. Audio has no equivalent of skimming. A scene that reads quickly on the page can run four minutes aloud, and listeners feel every second. The fix is almost always cutting, not speeding up.

Fourth, silence. Gaps between lines are where the listener does their work. Directors of radio drama spend most of their time on the gaps, and it is the single most underused tool available to someone producing alone.

A workflow that survives revisions

The reason people abandon audio drama projects is rarely the first render. It is the twentieth revision, when a line change means reassembling an episode by hand. Any workflow that does not make revision cheap will collapse under its own weight.

The structure that holds up is: one project per episode, the same library characters and personas chosen for your recurring cast in every episode's project, and the script as the single source of truth. Change a line, regenerate that line — nothing else in the project is re-rendered — then export the episode again and drop it back into the mix. Every export counts the whole project, which is exactly why an episode, not a season, is the right size for one.

The corollary is that you should not do your mixing until the script is locked. It is very tempting to add music to a scene that sounds good, and very painful to redo it after a rewrite. Get the performance right first, then score once.

What this does not do

It does not write your script. There is no story generation here, deliberately — the interesting problem in audio drama is production, not ideas, and a generated script performed by generated voices is a product nobody has asked for twice.

It does not mix. No music beds, no reverb, no spatial placement. Those belong in an audio editor where you can hear everything together, and every attempt to fold them into a rendering tool has produced something worse than either.

It also does not handle overlapping dialogue in a single render. People talking over each other is a real dramatic device and it is a mixing operation: export the stems and slide them. That is one drag in any editor and it is the most realistic thing you can do to an argument scene.

비용은 이 정도

Free

US$0 / 월

 

내 대본에 목소리가 붙으면 어떤 소리가 나는지, 돈 한 푼 안 들이고 들어 보세요.

  • 생성과 미리 듣기는 크레딧이 안 빠져요 (24시간에 30줄까지)
  • 다운로드 크레딧 4,500개, 평생 한 번 (완성 오디오 11분쯤)
  • 캐릭터 보이스 40개 전부 — 유료 장벽은 라이브러리가 아니에요
  • 프로젝트당 등장인물 2명
  • 개인·비상업적 용도만

Basic

US$5 / 월

 

라이브러리를 전부 열고, 완성본을 가져가기 시작하는 단계.

  • 캐릭터 보이스 40개 전부
  • 매달 다운로드 크레딧 54,000개 (130분쯤)
  • 프로젝트당 등장인물 5명
  • 상업적 이용 라이선스 포함

Plus

US$19 / 월

 

속도를 직접 잡는 단계. 독백은 늦추고, 말다툼은 몰아치게.

  • 매달 다운로드 크레딧 180,000개 (450분쯤)
  • 대사별 말속도 조절
  • 프로젝트당 등장인물 무제한
  • SRT 자막 내보내기

전체 비교 →

자주 묻는 질문

What exactly is an audio drama?

A story told entirely in sound: dialogue, narration, effects and music, with no picture. It is the radio-play tradition, and it is having a substantial revival through podcast distribution. The defining feature is that characters are played by different voices, which is precisely what a single-voice tool cannot give you.

How is this different from ordinary text to speech?

Ordinary text to speech converts a block of text with one voice. This takes a script with several speakers, assigns a different voice to each, lets you direct each line separately, and mixes the result. The underlying speech synthesis is one component; the casting and direction layer is the product.

Do I need a specific script format?

No. Screenplay format, Name: lines, and ordinary prose with dialogue tags all work. If a scene is ambiguous — unnamed speakers, a lot of implied attribution — you will need to label those lines yourself, which takes a minute.

How many characters can a project have?

Two on the free plan, five on Basic, unlimited on Plus and Pro. A typical audio drama scene has three to six speakers including the narrator, so Basic covers most single-creator work and Plus covers ensemble pieces.

Can I add music and sound effects?

Not inside CastDub — it renders performances, not mixes. Export the stems and score them in any audio editor. This is deliberate: dialogue and sound design are separate crafts and combining them in one tool tends to make both worse.

Can I sell what I produce?

Yes on any paid plan; commercial rights are included. The free plan is for personal, non-commercial use. Your script remains yours in all cases.

오늘 밤, 첫 장면을 캐스팅해 보세요

장면을 붙여 넣고, CastDub가 인물마다 보이스를 붙이게 하고, 완성된 오디오 드라마로 내보내세요.

무료로 만들어 보기

카드 등록 없이. 무료 요금제는 한 번만 주어지는 크레딧이에요.