Models

MAI Voice 2

mai-voice-2Microsoft's expressive text-to-speech model, ranked #1 for audio realism. Four natural voices across English, Spanish, French, and German.

Per 1K characters

$0.0220

Uptime · 24h

—

No traffic yet

Details

Type
Text to speech
Input
text
Output
audio
Voices
en-US-Harper:MAI-Voice-2, es-MX-Valeria:MAI-Voice-2, fr-FR-Soleil:MAI-Voice-2, de-DE-Klaus:MAI-Voice-2
Max characters
4,096

Supported parameters

inputvoiceresponse_format

Last 24 hours

Requests per hour; red = provider errors.

Avg latency
—
Avg first token
—
Throughput
—

Use it

curl https://api.ultragpt.pro/v1/audio/speech \
  -H "Authorization: Bearer $ULTRAGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "model": "mai-voice-2", "voice": "en-US-Harper:MAI-Voice-2", "input": "Hello from UltraGPT!" }' \
  --output hello.mp3