a5d98cd2b350beea12ef286f217797dc033ffd9f
Updated the dev:api script to use `uv run` instead of a hardcoded virtual environment path. Changes: - replaced `.venv/bin/uvicorn` with `uv run uvicorn` Reason: The previous implementation relied on a POSIX-specific path, which breaks on Windows (where executables are located in `.venv/Scripts`). Using `uv run` ensures the command works consistently across different operating systems by resolving the environment automatically.
refactor: simplify README documentation and update API and frontend to support voice design features
refactor: simplify README documentation and update API and frontend to support voice design features
refactor: simplify README documentation and update API and frontend to support voice design features
OmniVoice Studio
OmniVoice Studio is a local, full-stack application built for fast, high-quality cinematic dubbing and custom voice generation. It wraps the 600-language zero-shot OmniVoice model into a clean workspace that works right out of the box.
Features
- Cinematic Dubbing: Drop in an MP4. The studio transcribes the speech, automatically translates it into your chosen language, and mixes the newly generated voice back into the video.
- Background Audio Mixing: Automatically utilizes
demucsto isolate vocals. When your dubbed audio is laid over the video, the original background music and sound effects are perfectly preserved in the mix. - Voice Design & Cloning: Create entirely new voices using simple tag combinations (like
female,elderly,british accent), or instantly clone an existing voice from just a 3-second audio snippet. - Runs Locally: Fully handles its own asynchronous threading, caching, and VRAM management across Mac (Apple Silicon), NVIDIA/AMD, and CPU setups.
Getting Started
- Make sure
ffmpegis installed on your system. - Install Bun if you don't have it.
- Install uv if you don't have it.
- Clone and run:
git clone https://github.com/debpalash/OmniVoice-Studio.git
cd OmniVoice-Studio
uv sync
bun install
bun dev
The interface will load at http://localhost:5173 (or 5174), and the API runs on port 8000. All necessary model weights will download automatically during your first generation.
Powered by the open-source OmniVoice diffusion model.
Description
VoiceStudio is the open-source, fully-local ElevenLabs alternative — voice cloning, voice design, video dubbing, dictation, transcription & audiobook creation in 646 languages.
https://voicestudio.sh
aiaudiobookcudadubbingelevenlabs-alternativehuggingfacelocal-firstmlxomnivoice-studiospeech-to-texttauritext-to-speechtranscriptiontranslatettsvoice-aivoice-cloningvoice-generationvoicestudioworkflow
Readme
AGPL-3.0
113 MiB
Languages
Python
51.6%
JavaScript
24.3%
TypeScript
15.9%
Rust
4.8%
CSS
1.7%
Other
1.7%