AI GlossaryㄹWords you meet while using AI
Realtime TTS-2
A real-time text-to-speech model from Inworld AI that lets you adjust voice tone and emotion on the fly using natural-language instructions.
In plain words
Realtime TTS-2 is a text-to-speech program made by Inworld AI that turns written text into a human-sounding voice.
Most programs like this speak in only one preset tone, but this one can change tone and emotion in real time based on instructions like "sound brighter" or "a little sadder," almost like directing an actor's performance. It can also switch between more than 500 dialects while keeping the same voice identity, and according to the company, this processing happens within 150 milliseconds, making it flow as naturally as a real conversation.
It's a foundational technology that can be used in various services where a computer needs to speak in a human voice, such as call center bots or audiobook narration.
How it shows up in the news
In the article, it says "Inworld AI officially launched Realtime TTS-2 and took the top spot on the Artificial Analysis leaderboard." However, this No. 1 ranking is based on the company's own citation of results measured by Artificial Analysis, an independent evaluator that scores various voice models—and it's worth noting that rankings shift often, with Cartesia's Sonic-3.6 having held the top spot on the same leaderboard just two weeks earlier.
Try it yourself
If you can try a text-to-speech service that lets you adjust tone through instructions like Realtime TTS-2, compare the results by changing only the instruction for the same sentence.
- Pick one short sentence. Example: "Today's meeting starts at 3 PM."
- Have it read with Instruction A: "Read this in a calm, businesslike tone."
- Have it read the same sentence again with Instruction B: "Read this in an excited, upbeat tone."
- Compare how differently the same words sound depending on the instruction.
See also
Stories using this term
- Inworld AI launches Realtime TTS-2, reigniting the race for the top speech synthesis spotAI · 2026.09.03
- Google's ADK adds automated evaluation for live voice agentsAI · 2026.08.25
- Google unveils SL2T, converts sign language to text, rolling out with Pixel 11AI · 2026.08.13
- LTX achieves 'stutter-free avatars' with frame-by-frame real-time streamingAI · 2026.08.15
- GPT-5.6 Sol Improved, Free Users Get LunaAI · 2026.08.08
- Cartesia Sonic-3.6 tops both voice synthesis leaderboardsAI · 2026.08.19
