Skip to content
Oct 1Thu
Sep 23Wed
Aug 31Mon
  1. Inworld AI Blog57

    Inworld AI releases Realtime TTS-2 realtime conversational voice model

    Inworld AI has released Realtime TTS-2, a new-generation realtime conversational voice model, now fully available in the Inworld API and the Inworld Realtime API.

    Why it matters: The original post covers multi-turn audio context, natural-language voice instructions and a first-audio latency of under 200ms, which readers can use to judge whether realtime voice conversation fits their own projects.

May 14Thu
  1. Inworld AI Blog49

    Inworld pairs Realtime TTS-2 with Stream Vision Agents for a realtime voice agent that can see and listen

    Inworld teamed up with Stream to build a reference implementation, Crashout Buddy, combining the latest speech model Realtime TTS-2 with Stream's open-source Vision Agents framework, so it can see facial expressions, hear speech and adjust its delivery in real time.

    Why it matters: The post shows how to turn visual signals and emotional context into speech delivery instructions, which readers can use to judge whether this realtime voice pipeline can be plugged into their own projects.