back to top
HomeSoftwareAI ToolsVoicebox – Offline AI Voice Cloning & TTS Studio (Qwen3-TTS, Open Source)

Voicebox – Offline AI Voice Cloning & TTS Studio (Qwen3-TTS, Open Source)

- Advertisement -

File Information

FileDetails
NameVoicebox
Versionv0.3.0
Formats .msi.dmg
Size299MB (exe) • 330MB (dmg)
PlatformsWindows • macOS
LicenseOpen Source (MIT License)
Github RepositoryVoiceBox Github
Official Websitevoicebox
CategoryVoice AI • Speech Synthesis • Audio Tools

Description

Voicebox is a local-first, open-source voice synthesis studio designed for cloning voices, generating realistic speech, and building voice-powered applications directly on your own machine.

It keeps everything local. Your voice samples, models, and generated audio never leave your system, giving you full privacy, ownership & control.

With a DAW-like interface, multi-track editing, and an API-first design, Voicebox is built for creators, developers, and teams who want professional voice tools without usage limits or cloud dependency.


Use Cases

  • Clone voices locally for narration or dialogue
  • Create podcasts, stories, and multi-speaker conversations
  • Build game dialogue and character voice systems
  • Automate voice generation in content pipelines
  • Develop privacy-focused voice assistants
  • Generate speech for accessibility tools
  • Integrate voice synthesis into apps via API
  • Experiment with open-source TTS models safely

Screenshots

Features of VoiceBox

FeatureDescription
Local Voice CloningClone voices from short audio samples completely offline
Speech QualityNatural prosody, emotion, and realistic cadence
Studio EditorTimeline-based, multi-track audio composition
Multi-Voice SupportCreate conversations with multiple speakers
Open ModelsPowered by Qwen3-TTS, with more open models planned
API AccessFull REST API for automation and integrations
Native AppLightweight, high-performance desktop app (Tauri)
Apple Silicon BoostMLX backend delivers 4–5× faster inference
Privacy FirstNo cloud, subscriptions, limits, or internet required

System Requirements

Windows

RequirementDetails
Operating SystemWindows 10 or later
Architecture64-bit
RAM8 GB minimum
Disk Space5–10 GB
GPUOptional (CPU supported)

macOS

RequirementDetails
Operating SystemmacOS (Apple Silicon or Intel)
ArchitectureARM64 / x64
RAM8 GB minimum (16 GB recommended)
Disk Space5–10 GB (models + audio)
AccelerationMetal / MLX (Apple Silicon)

How to Install VoiceBox??

Windows (.exe)

  1. Download the Voicebox .msi installer
  2. Run the installer
  3. Follow the setup steps
  4. Launch Voicebox from the Start Menu

macOS (.dmg)

  1. Download the Voicebox .dmg file
  2. Open the DMG
  3. Drag Voicebox.app into the Applications folder
  4. Launch from Applications
    • If macOS shows a security warning, go to
      System Settings → Privacy & Security → Open Anyway

Linux

According to the developer , it is planned to launch the Linux build soon. So as soon as it will be available , we will update the page. But if you want to build it from source, follow the official guide

Recommended For You: Handy: Offline Open-Source Speech-to-Text AI App For Windows, macOS & Linux

How to Use Voicebox (Simple Steps)

Getting started with Voicebox is straightforward just follow the below steps after installation:

  1. Launch the Voicebox app on macOS or Windows
  2. On first launch, select and download a voice model
    • Progress, speed, and status are shown clearly
  3. Once the model is ready, import or record a short voice sample
  4. Voicebox automatically creates a voice profile
  5. Enter your text and generate speech locally
  6. Use the timeline editor to mix voices, trim audio, or build conversations
  7. Export your audio or reuse it later from generation history

Download Voicebox: Local Voice Cloning & Speech Synthesis Studio For Windows & macOS

Open Source & Development

Voicebox is developed as a fully open-source project, that means users and developers can:

  • Inspect and audit the source code
  • Contribute features or bug fixes
  • Experiment with new voice models
  • Build custom voice-powered tools

By using Tauri instead of Electron, Voicebox stays lightweight, fast, and memory-efficient while still offering a modern UI.

Conclusion

Voicebox delivers a powerful, privacy-first approach to voice synthesis, combining voice cloning, speech generation & audio editing into one open-source desktop application.

With local execution, native performance, and an API-driven design, it’s well-suited for creators and developers who want professional voice tools without cloud subscriptions.

If you’re exploring voice AI, audio storytelling, or voice-powered applications with control, transparency & performance, Voicebox is a very useful software.

Want more stories worth your time?

Add us to your Google favorites. We cover the tech stories, AI developments, and open-source projects that are easy to miss in the noise.

Add as a preferred source on Google

Don’t miss any Tech Story

Subscribe To Firethering NewsLetter

You Can Unsubscribe Anytime! Read more in our privacy policy

LEAVE A REPLY

Please enter your comment!
Please enter your name here

YOU MAY ALSO LIKE
hister search engine

Hister: Your Own Private Search Engine for Web Pages and Files

0
You know that page you read three months ago and somehow can’t find again? Hister is built for exactly that problem. It turns the web pages you visit and the files you keep into your own searchable index, so you can search the actual content instead of trying to remember a title, URL, or where you saved it.
Atomic Chat App

Atomic Chat: Run Open-Weight LLMs Locally on Windows, macOS & Linux

0
Want to run an AI model locally, but still use it with the tools you already have? Atomic Chat makes that possible. It lets you run open-weight LLMs from Hugging Face on your own computer, then exposes them through an OpenAI-compatible API so coding agents, CLIs, IDE plugins and other apps can use your local models too. You can run models such as Llama, Gemma, Qwen, Mistral and Phi, use Atomic Chat as a regular AI chat app, or connect it to tools such as OpenCode, Goose and Kilo Code. Your local conversations and API keys can stay on your machine, while cloud providers such as OpenAI, Anthropic, Mistral and Groq are available when you need them. Under the hood, it also supports multiple inference engines and performance features such as speculative decoding, Flash Attention and TurboQuant on supported models and hardware. So you're getting more than a local chatbot. Atomic Chat can act as the local AI layer behind the rest of your setup.
OpenNOW Open-Source GeForce NOW Gaming Client for Windows Mac and Linux

OpenNOW: Open-Source GeForce NOW Gaming Client for Windows, Mac & Linux

0
OpenNOW is an open-source desktop client for GeForce NOW that gives cloud gaming users a community-built alternative to accessing NVIDIA's game-streaming service. It provides a dedicated interface for browsing the GeForce NOW catalog, configuring streaming options, and launching gaming sessions from one place. Built as an Electron application, OpenNOW is actively developed with support extending beyond desktop platforms, including Android, iOS beta, and Nintendo Switch. It also includes experimental native streaming infrastructure for users who want to go beyond the standard web-streaming path.