Nobody owns this listing yet

Voicebox is live in the directory, but no verified owner controls the page.

Claim this listing

Free to claim · 2 minutes

  • Pick your launch daySchedule a launch and compete for the daily #1 spot.
  • Win trophies you can embedRank top 3 on launch day and get an embeddable trophy.
  • “Featured on AI Kaptan” badgeOwner-only badge with your live upvote count.
  • Control what the page saysEdit copy, pricing, links and screenshots as the owner.

Unclaimed listings don't stay up forever. Listings left unclaimed are removed from the directory in periodic cleanups. When a listing goes, it goes everywhere — the tool page and every alternatives comparison it appears on, along with the rankings and backlinks that page had built up.

Verify with a work email on the domain (about 2 minutes), then post about it on LinkedIn and tag us. See everything you get

Voicebox
Voice cloning
Free

Voicebox

Open-source, local AI voice studio for professional voice cloning and TTS. Powered by Qwen3-TTS with multi-voice timeline editing and full privacy.

Rating
0
Reviews
0
Upvotes
4 months ago
Listed
Voice cloning
Text-to-speech
Open source
Local AI
Podcast tools
Audio editing
Privacy focused
Transcription
Tool information
Provider
Jamie Pine
Platforms
Windows
macOS
Linux
Languages
English
Chinese
Japanese
Korean
German
French
Italian
Spanish
Arabic
Hindi
API Available
Yes
Added to directory
4 months ago
Last updated
about 2 months ago
Rating
Not available
Pricing
Free

About Voicebox

What the tool does and who it's for

Voicebox is an open-source, local-first AI voice studio designed for high-fidelity voice cloning and text-to-speech without cloud dependency. Powered by the Qwen3-TTS engine, it allows users to replicate voices from short audio samples (as little as a few seconds) while maintaining complete privacy by running all processing on the user's hardware. The application features a comprehensive suite of tools, including a multi-voice timeline editor for stories and podcasts, post-processing audio effects via Spotify's Pedalboard, and Whisper-powered transcription. It supports multiple TTS engines like Chatterbox and TADA, offering capabilities ranging from expressive paralinguistic tags (laughs, sighs) to lightweight CPU-optimized inference.

Key capabilities

Local-First Architecture: Runs entirely on your machine via CUDA or Metal acceleration for 100% privacy.
Multi-Engine Support: Switch between Qwen3-TTS, Chatterbox, TADA, and Kokoro engines for different needs.
Stories Editor: Timeline-based editor to arrange tracks, trim clips, and mix multi-voice narratives.
Expressive Tags: Use paralinguistic tags like [laugh], [sigh], or [cough] for realistic emotional delivery.
Whisper Transcription: Automatically convert reference audio to text for seamless voice profile creation.
Post-Processing: Built-in 8 audio effects including Reverb, Pitch Shift, and Delay via Spotify Pedalboard.
Local REST API: Integrate voice generation into other apps, games, or AI agents via a local server.
Voice Profile Management: Support for multi-sample imports to improve cloning accuracy and consistency.

Pricing

Free — plans below

Open Source

Most popular

Free

  • Unlimited local generation
  • No subscription required
  • Full access to all TTS engines
  • Privacy-first processing
  • Stories timeline editor
  • Local REST API access

Reviews

Be the first to review Voicebox

No reviews yet

Be the first to share your experience with Voicebox.

Featured Tools

Handpicked by our team of experts

Suggest an Update

Found outdated information? Help us keep this listing accurate.

FAQ

Voicebox FAQs

Common questions about Voicebox