Speech & transcription

POST /audio/speech returns audio (mp3, opus, aac, flac, wav, pcm) in one of the model's voices, priced per character. POST /audio/transcriptions (and /audio/translations) take a multipart file up to 25MB and return json, text, verbose_json, srt or vtt.

curl https://api.ultragpt.pro/v1/audio/speech \
  -H "Authorization: Bearer $ULTRAGPT_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "model": "ai-voice", "voice": "default", "input": "Hello from UltraGPT!" }' \
  --output hello.mp3
curl https://api.ultragpt.pro/v1/audio/transcriptions \
  -H "Authorization: Bearer $ULTRAGPT_API_KEY" \
  -F model="audio-transcribe" \
  -F file="@meeting.mp3"