Video & Audio

🗣️ Voicebarn

Unlimited natural-sounding text-to-speech on your own machine.

Quick facts

Voicebarn at a glance

Price$34, once — no renewal
ReplacesElevenLabs (Creator) at $22/mo — roughly $264 a year, every year
TypeDesktop app — runs offline
PlatformsWindows, macOS, Linux
CategoryVideo & Audio
What's in the box

Voicebarn features

01

🎙️ Local neural TTS via Piper

Clean, natural narration voices, 100% offline after the one-time engine + voice download.

02

📝 Per-paragraph control

Every paragraph gets its own voice + speed override (0.5x–2.0x), inheriting document defaults when left blank.

03

⏸️ Inline pause tags

Drop or [pause 1s] anywhere for a precise, sample-rate-matched silence gap.

04

▶️ Instant preview

Synthesize and play any single paragraph before you export, with results cached by content hash.

05

💾 WAV or MP3 export

MP3 via bundled ffmpeg with a bitrate picker (128/192/256/320 kbps).

06

📦 Batch mode

Point it at a folder of .txt files and get one narrated audio file per input — chapter-by-chapter audiobook workflow.

07

🌍 8 curated starter voices

English (US/UK), German, Spanish, French — medium-quality Piper voices, each downloaded once from Hugging Face.

08

🔒 100% private

No telemetry, no account, no network calls except the clearly-surfaced one-time engine and voice downloads.

Pricing

$34 once — that's the whole pricing page

ElevenLabs (Creator) charges $22/mo — roughly $264 a year, every year. Voicebarn is a single payment, and it keeps working whether or not we ever sell another copy.

$22/mo forever $34once

Or get Voicebarn inside the full suite — every app in the catalog, currently $97 for the bundle. See the bundle →

FAQ

Honest answers

Is it really free on GitHub?

Yes — the full source is MIT at github.com/bensblueprints/voicebarn-mvp, always. $34 buys the signed installer, 1-click setup and updates instead of npm i && npm start.

Can it clone my voice or sound like a specific person?

No — and ElevenLabs is genuinely better if that's what you need. Voicebarn ships clean neural narration voices, great for videos, courses, audiobooks, IVR prompts and accessibility — not celebrity impressions or voice cloning. That's a different product category.

Does it support SSML?

Only an SSML-lite pause tag: or [pause 1s]. Piper doesn't support full SSML, and Voicebarn doesn't pretend otherwise — prosody, emphasis and phoneme tags will be read aloud literally. If you need real SSML, Piper isn't the right engine.