v2.0,

Streaming v2: first audio 40% faster

A rebuilt streaming pipeline that starts speaking sooner, holds intonation across chunks and reconnects on its own.

Streaming v2 is live for every workspace. Nothing to migrate: existing WebSocket connections pick it up automatically.

What changed:

  • Faster first audio: median time to first byte dropped from 290 ms to 170 ms.
  • Smoother joins: phrases streamed separately now share intonation, so long answers no longer sound stitched together.
  • Automatic reconnects: dropped connections resume mid-turn without repeating what was already spoken.
  • Flush on demand: a new flush message finishes the current turn immediately, handy for interruptions.

If you were padding text to get better intonation, you can stop. Send it as it arrives.