VibeVoice on a Mac
Microsoft's long-form, multi-speaker TTS · Speech AI
VibeVoice is Microsoft's open-source frontier text-to-speech model for long-form, multi-speaker audio — expressive, podcast-style conversations with up to four distinct speakers and natural turn-taking. It pairs an LLM with a diffusion head and a low-frame-rate acoustic tokenizer to stay coherent over long durations. Run it locally from the GitHub repo / Hugging Face weights (a 1.5B model and larger variants; a GPU helps): install the package, then synthesize from a script of speaker turns. Good for narration, dialogue and synthetic voice-over. Weights are MIT-licensed and for research use.
Run VibeVoice with FrontierStack
FrontierStack lists VibeVoice in its Speech AI catalog. Install or connect it from one place, then monitor its status, ports and certificate, secure it with the firewall and Malware Audit, and back it up.
Run it all from one Mac app.
FrontierStack installs, monitors and secures the whole stack — locally and across your fleet — from a single native macOS app.
Download FrontierStack