1.2 KiB
1.2 KiB
Features and engines
Engine availability depends on installed models, hardware, and configured providers.
Features
- Voice Cloning
- Voice Design
- Video Dubbing
- Dictation Widget
- Vocal Isolation
- Speaker Diarization
- Batch Queue
- MCP Server
- AI Watermark
- Local-first
- GPU Auto-Detect
- Remote Model Downloads
- Extensible
Speech generation
- VoiceStudio (default, powered by k2-fsa/OmniVoice)
- omnivoice-subprocess — Guide
- CosyVoice 3 — Guide
- KittenTTS
- MLX-Audio
- VoxCPM2
- MOSS-TTS-Nano
- gpt-sovits
- sherpa-onnx
- IndexTTS 2.5 ⚡ — Guide
- omnivoice-gguf
- supertonic3
- MOSS-TTS-v1.5 — Guide
- dots.tts — Guide
- Confucius4-TTS — Guide
- pockettts
- audiocpp — Guide
Transcription
- WhisperX (default)
- Faster-Whisper
- MLX Whisper
- PyTorch Whisper
- Parakeet TDT
- Parakeet TDT v3 (MLX)
- Moonshine
- FunASR
- sherpa-onnx (live dictation)
- OpenAI-compatible ⚠️ configured server