tts-1.5-max — AI Audio Model
inworld
tts-1.5-max
inworld/tts-1.5-max is a high-quality text-to-speech (TTS) model designed to generate natural, expressive, and human-like voice outputs from text. Built by Inworld AI, this model focuses on delivering realistic speech with emotional tone, clarity, and consistency—making it ideal for immersive experiences, interactive applications, and professional audio content. With support for nuanced voice modulation and low-latency performance, tts-1.5-max enables developers and creators to produce engaging voiceovers, character dialogues, and real-time conversational audio.
Pricing: 4.33 joules / generation
Output will be displayed here
Fill in the inputs on the left and hit RUN to preview the results.
Examples
Explore different use cases and parameter configurations