VibeVoice is an open-source frontier voice AI model developed by Microsoft, which includes both Text-to-Speech (TTS) and Automatic Speech Recognition (ASR) models. The model can handle long-form audio and supports multiple speakers, languages, and streaming text input. However, the original TTS model was removed from the repository due to security and safety concerns. The current models available include VibeVoice-ASR, VibeVoice-TTS, and VibeVoice-Streaming.