realtime-tts-2

Coming soon
inworld-ai

Realtime TTS 2.0 is a low-latency text-to-speech model with natural language steering, allowing you to control tone and emotion directly in the prompt (e.g., “[be happy and upbeat] Hello!”). It supports cross-lingual voices and multiple languages, enabling the same voice to speak consistently across different languages. This is an early access preview ahead of full launch, with ongoing improvements to voice quality and steering.

Modality

Chat

Region

US

Published

Jul 21, 2026

Use it via the API

realtime-tts-2 works with any OpenAI-compatible SDK — point the base URL at https://preprod-backend.sovereigneg.com/v1 and use your SovereignEG API key.

Run inference

OpenAI-compatible — POST /v1/chat/completions — drop-in for any OpenAI SDK.

from openai import OpenAI

client = OpenAI(
    base_url="https://preprod-backend.sovereigneg.com/v1",
    api_key="YOUR_API_KEY",
)

response = client.chat.completions.create(
    model="realtime-tts-2",
    messages=[
        {"role": "user", "content": "What are some fun things to do in Cairo?"}
    ],
)

print(response.choices[0].message.content)

Frequently asked questions

Where is realtime-tts-2 hosted?

realtime-tts-2 is served from US (US-hosted inference).

How do I use realtime-tts-2 via the API?

realtime-tts-2 is available through the OpenAI-compatible SovereignEG API: point your SDK's base URL at https://preprod-backend.sovereigneg.com/v1 and call /v1/chat/completions with model "realtime-tts-2" and your API key.