Best Open-Source TTS Yet? Microsoft VibeVoice @WhatsAI
Best Open-Source TTS Yet? Microsoft VibeVoice  @WhatsAI
Uploaded September 2025 | Updated September 2026, 15 minutes ago
Microsoft just released a major breakthrough for AI voice tech — and it’s open source! 🎙️

They dropped VibeVoice, an open-source TTS framework for long-form, multi-speaker audio.

• Two models:
 – VibeVoice-1.5B: 64K context, ~90 min
 – VibeVoice-Large: ~10B params, 32K context, ~45 min

• Handles up to 4 speakers → perfect for podcasts or multi-voice audiobooks

• MIT-licensed → free for commercial use

• Outperforms baselines in realism, richness, and speaker similarity

• Still only supports English + Chinese for now

• Doesn’t handle background music or effects — pure speech only

• A smaller streaming model is also coming soon!

It’s not replacing things like GPT’s real-time voice API, but it’s a seriously powerful alternative for creating high-quality long-form audio.

Would you use this for podcasts or audiobooks? 👀

I’m Louis-François — PhD dropout, now CTO & co-founder at Towards AI. Follow me for tomorrow’s no-BS AI roundup 🚀 #short
Best Open-Source TTS Yet? Microsoft VibeVoiceMy AI Coding Stack as a CTOProprietary vs Open-Weight vs Open-Source AI ModelsSora 2 update — by Sora 2When to Use a Workflow Instead of an AI AgentShould you have a plan B?The truth about working for yourself in AIHow it is to write a book in AIWhy AI “Forgets” Your ConversationWhat to Do When Your AI Coding Limit Resets TomorrowAnthropic vs OpenAI: The Government AI Deal TimelineAI Model Distillation, Explained
Whats AI by Louis-François Bouchard |

Best Open-Source TTS Yet? Microsoft VibeVoice

SHARE TO X SHARE TO REDDIT SHARE TO FACEBOOK WALLPAPER