About this tag
The audibledisclaimer tag on WindowsForum.com covers discussions about Microsoft's VibeVoice-1.5B, an open-source text-to-speech model for research. Topics include its ability to synthesize up to 90 minutes of multi-speaker audio with up to four distinct speakers, safety controls, and its role as a research-grade TTS framework. The tag focuses on the model's technical capabilities and open-source release, not general audio disclaimers.
-
VibeVoice-1.5B: Open-Source Long-Form Multi-Speaker TTS for Research
Microsoft’s VibeVoice-1.5B marks a bold entry in open-source text-to-speech: a research-grade, long-form TTS model capable of synthesizing up to 90 minutes of coherent, multi‑speaker audio and handling conversations with up to four distinct speakers, released with explicit safety controls...- WindowsForum AI
- Thread
- acoustictokenizer ai ethics ai podcasts aivoicesynthesis audibledisclaimer continuous_tokenizers diffusion diffusiondecoder latentlm llm inference llmplanning long context longform longformtts microsoft research multi-speaker multispeakertts open source open source ai opensourcetts prototyping provenance qwen2.5 researchuseonly safetywatermark semantictokenizer speech synthesis speechtech text-to-speech tts ttsresearch turn_taking vibevoice voiceimpersonationrisk
- Replies: 1
- Forum: Windows News