€5
Script to Voice Generator - Hume AI Octave 2
Turn author scripts into fully voiced audio files with Hume AI Octave 2 TTS, acting descriptions, per-character effects, and smart merging.
## Description
Script to Voice Generator converts formatted `.txt` or `.md` script files written by you into fully voiced audio using **Hume AI Octave 2 TTS** — an expressive, emotionally nuanced cloud voice engine that produces cinematic-quality output no local model comes close to matching.
Each speaker in your script gets their own voice, acting description, pitch, speed, and audio effects. Assign Radio, Reverb, Distortion, Telephone, Robot Voice, Cheap Mic, Underwater, Megaphone, Worn Tape, Intercom, Alien Voice, or Cave effects per character at Off / Mild / Medium / Strong levels. Two bonus toggles — FMSU (brutal digital corruption) and Reverse — round out the toolkit. Combine effects freely for distinct character identities.
**Acting descriptions** are a core part of the design. Each speaker panel has a description field — a short phrase like `"Gruff, military precision"` or `"Warm, conspiratorial, slightly amused"`. Inside the script, override it per-line using `[brackets]`: `[barely containing panic]`, `[speaking to a child]`, `[tense, barely breathing]`. Multiple brackets on one line are merged automatically. Descriptions are parsed and stored correctly. *Note: Hume AI Octave 2 is currently in preview and the API does not yet accept the description field — this is expected to be resolved as the API matures.*
**The generator produces:**
- Individual audio per each spoken line — clean (TTS only) and effects-processed versions
- All (effects-enabled) clips merged into a fully edited and smartly paced audio file, not normalized (true audio, better for media/games)
- All (effects-enabled) clips merged into a fully edited and smartly paced audio file, loudness normalized (even audio, better for podcasts)
- Reference .txt file with filenames, line numbers, and spoken content for every clip
**Features:**
- 160+ Hume AI voices (plus custom voice support)
- Per-character acting description — global default + per-line `[bracket]` overrides *(preview limitation: not yet transmitted to API)*
- Per-character voice, pitch, speed, volume, and audio effects
- 12 audio effects with 4 levels each — combinable
- 2 bonus toggles: FMSU (brutal corruption) and Reverse
- Clip continuity — chains generation IDs for natural voice flow across consecutive lines
- Yell Impact mode for punchy single-word exclamations
- Inner thoughts filter (Whisper, Dreamlike, Dissociated presets)
- Sound effect events (play/stop/loop) placed in the merge timeline
- Smart merged audio with configurable punctuation-based pause timing
- Loudness-normalized and raw merge outputs
- Per-line clip files (clean and effects) for editing flexibility
- Character profiles saved automatically between sessions
- Parse log with line-by-line error reporting
- Included example scripts and AI prompt templates to get started fast
**A Hume AI account and API key are required.** The free tier includes 10,000 characters per month (~20 minutes of audio). Paid plans start at $3/month. This is a premium TTS engine — use it intentionally and it delivers.
**Ideal for:**
- Audio dramas
- Game dialogue
- Character oneliners
- Audio books
- Ambient narration
- Meditation audios
- Voice banks
...and any project where you want to convert a written script into audio with genuine emotional performance.
*Includes in-depth guides covering script writing for TTS, acting descriptions, and the full audio effects pipeline.*
---
Check out everything else I do: ✨🚀
https://linktr.ee/reactorcore
€5
€5