Transcribe audio to text and generate realistic voice clones locally on Windows. Powered by Whisper, Chatterbox, Kokoro, LuxTTS, TADA, and Qwen — no cloud required.
Zero cloud uploads. Zero account setup. Full local processing.
Download the free Windows executable. Includes local sidecar engine wrappers.
Choose Whisper for ASR or Chatterbox / Kokoro for zero-shot voice cloning.
Generate speech or transcribe audio instantly on CPU/GPU without internet.
Declarative backend architecture powering offline speech and voice capabilities.
Tiny, Base, Small, Medium, Large-v3, and Turbo variants for 24+ language transcription and real-time streaming.
Zero-shot voice cloning across 23 languages using 5-second reference audio prompts.
Fast English voice cloning with native support for paralinguistic tags like [laugh],
[cough], and [chuckle].
High-fidelity, ultra-lightweight speech synthesis with minimal memory footprint.
CPU-friendly voice generation designed for fast execution without dedicated GPU requirements.
HumeAI TADA (1B & 3B) text-acoustic dual alignment and Qwen2 Audio custom voice backends.
CincoScribe is 100% free and open-source under the MIT License.