ホームサービス › VibeVoice

VibeVoice を Mac で

Microsoft's long-form, multi-speaker TTS · Speech AI

VibeVoice is Microsoft's open-source frontier text-to-speech model for long-form, multi-speaker audio — expressive, podcast-style conversations with up to four distinct speakers and natural turn-taking. It pairs an LLM with a diffusion head and a low-frame-rate acoustic tokenizer to stay coherent over long durations. Run it locally from the GitHub repo / Hugging Face weights (a 1.5B model and larger variants; a GPU helps): install the package, then synthesize from a script of speaker turns. Good for narration, dialogue and synthetic voice-over. Weights are MIT-licensed and for research use.

VibeVoice を FrontierStack で使う

FrontierStack は VibeVoice を Speech AI カタログに収録しています。ひとつの場所でインストール/接続し、状態・ポート・証明書を監視、ファイアウォールとマルウェア監査で保護し、バックアップできます。

すべてをひとつの Mac アプリで。

FrontierStack は、ローカルでもフリート全体でも、スタック全体のインストール・監視・保護をひとつのネイティブ macOS アプリで行います。

FrontierStack をダウンロード

関連 Speech AI