Speech 2.6 HD
Sound from textVoice
Minimax Speech 2.6 HD: Ultra-human, low-latency (< 250ms) TTS with voice cloning, text normalization and support for 40+ languages.
from€0.20per render
Examples
Specs
Typetext-to-audio
OutputAudio
InputText
Model IDminimax/speech-2.6-hd
Similar models
All models →Prices are per render and charged from your balance only when a render succeeds. Failed renders are refunded automatically. Examples show what the model can do; your result depends on your prompt and settings.