docs
models / Inworld / Inworld TTS 2

Inworld TTS 2

audioinworld/tts-2ZDR supported

Inworld's most powerful and expressive text-to-speech model. Builds on TTS 1.5 with rich expressive speech, real-time latency, natural language steering (e.g. [whisper], [say excitedly]), and stronger multilingual support across 15 production languages plus 90+ experimental languages.

$0.025list priceper 1K characters
pricing
unitlist price
per 1K characters$0.025

Everyone pays the list price. Reaching Elite (hold or stake 1,337 $XAVA, or hold 1,337,000 $XINF) unlocks the Elite price and a bigger buyback of your token. How it works

curl
curl https://www.zinf.ai/v1/media/generations \
  -H "Authorization: Bearer $XINF_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{"model":"inworld/tts-2","input":{"text":"Hello from one key."}}'
limits
providerInworld
typeaudio
context window—
max output—
prompt storagenone
provider retentionnone ZDR supported
livesoon
requests 24h—
p50 latency—
uptime 7d—