Choose Auto, TokenLab Verified, or Official for each request, with prices shown up front.See what's new

Inworld AI: inworld-realtime-tts-2

Inworld Realtime TTS-2 is a conversational text-to-speech model built for realtime voice interaction rather than static narration. It supports free-form voice direction, carries tone and pacing forward from prior audio in a session, preserves one voice identity across 100+ languages, and is designed for expressive, low-latency speech in assistants, characters, support agents, and interactive products.
Compare models
inworld-realtime-tts-2
AvailableInworld AIText to speechSync
Price
From $0.000035
Modalities
Audio

Getting started

  1. Create API key

    Create a key in Console, then use it with every model on the platform.

  2. Send your first request

    Copy the example for your language and run it against the endpoint.

    Text to speechTokenLab endpoint
    POST/v1/audio/speech
    curl -X POST "https://api.tokenlab.sh/v1/audio/speech" \
      -H "Authorization: Bearer sk-xxx" \
      -H "Content-Type: application/json" \
      -d '{
      "operation": "tts",
      "model": "inworld-realtime-tts-2",
      "input": "Hello from TokenLab.",
      "voice": "alloy"
    }'

Pricing

Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.

Text to speech

Per 1M characters
Official price
$0.000035
Official
$0.000035
Discount
-

Usage & activity

Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.

Usage & availability

Last 24 hours

No data yet

Model performance
Metrics appear once privacy and data volume thresholds are met.

Data is based on aggregate user requests, excluding status checks.

Open in Console

Open Inworld Realtime TTS-2 in Console with a prompt ready to edit or send.

Help me try inworld-realtime-tts-2 with a short audio request at /v1/audio/speech. Show the result and cost.

Use cases

01

Voice UX

Generate narration, onboarding audio, and spoken UI responses.

02

Side-by-side test

Compare response quality, latency, and price side by side.

Prompt examples

Generate calm onboarding narration for a developer tool.

Send the smallest possible request and show me the result and cost.

Compare this model with a cheaper one in the same category — when is each worth it?

FAQ

How much does Inworld Realtime TTS-2 cost?

On TokenLab, Inworld Realtime TTS-2 costs - . The pricing table above shows the full breakdown. Rates depend on the billing unit, specification, and usage. Compare matching conditions in the model's detailed pricing; a single rate does not determine the total cost.

What is Inworld Realtime TTS-2 best for?

Inworld Realtime TTS-2 supports Text to speech. You can open it directly in Create.

How do I test Inworld Realtime TTS-2?

Open Inworld Realtime TTS-2 in Create. A sample for /v1/audio/speech will be ready to try.

Which endpoint should Inworld Realtime TTS-2 use?

Use https://api.tokenlab.sh/v1/audio/speech for Inworld Realtime TTS-2. The request example below shows the matching code shape.

Can I test Inworld Realtime TTS-2 before integrating it?

Yes. Open Console starts a ready draft for Inworld Realtime TTS-2 and keeps your prompt after sign-in, so you don’t lose context.

Which operations does Inworld Realtime TTS-2 support?

Inworld Realtime TTS-2 supports Text to speech. Select an operation above to see its endpoint and request example.

More from Inworld AI

Related models