Inworld TTS 1.5

Audio·Inworld·by Inworld

Inworld-TTS-1.5 is an advanced text-to-speech (TTS) model that converts written text into natural, expressive, and human-like speech. Designed for low latency and real-time performance, it supports high-quality voice output for applications such as voice assistants, games, interactive experiences, and content creation.

Runtime (p50)
1s
Estimated price
From $0.000005
Call the API
prediction.sh
sh
curl -X POST \
  -H "Authorization: Bearer $EACHLABS_API_KEY" \
  -H "Content-Type: application/json" \
  --data '{
    "model": "inworld-tts-1-5",
    "version": "0.0.1",
    "input": {
        "apply_text_normalization": "APPLY_TEXT_NORMALIZATION_UNSPECIFIED",
        "audio_encoding": "MP3",
        "bit_rate": 40000,
        "model_id": "inworld-tts-1.5-mini",
        "text": "Hey everyone, welcome to Eachlabs AI! Eachlabs is an advanced AI platform that offers powerful tools for text, image, and voice generation. It’s built to help creators, developers, and businesses produce high-quality content quickly and easily. With a focus on realism, speed, and flexibility, Eachlabs supports a wide range of creative and commercial use cases, making AI more accessible and impactful for everyone.",
        "voice_id": "Alex",
        "speaking_rate": 1,
        "temperature": 1.1
    },
    "webhook_url": ""
}' \
  https://api.eachlabs.ai/v1/prediction/

Related models

4 models
* FAQ

About Inworld TTS 1.5

01 / 03

What is Inworld TTS-1.5

Inworld TTS-1.5 is a real-time text-to-speech model ranked #1 on the Artificial Analysis TTS Leaderboard. It delivers 30% greater expressiveness and 40% fewer word errors than its predecessor, making it ideal for voice agents, interactive media, and accessibility applications.