Microsoft · Audio

Vibevoice

Sound from textVoice

Microsoft's VibeVoice text-to-speech model generates long-form speech from text with multi-speaker dialogue support. Choose from 9 voice presets across English, Chinese, and Hindi.

from€0.24per render
Use in Genlaxy →

Examples

Specs

Typetext-to-audio
OutputAudio
InputText
Model IDmicrosoft/vibevoice

Similar models

All models →

Prices are per render and charged from your balance only when a render succeeds. Failed renders are refunded automatically. Examples show what the model can do; your result depends on your prompt and settings.