Stream Athena Talk replies incrementally
This commit is contained in:
@@ -9,6 +9,9 @@ Athena speech stack:
|
||||
tools,
|
||||
4. Athena XTTS/Piper returns PCM audio to the Talk client.
|
||||
|
||||
Long replies are synthesized incrementally. The first short phrase starts
|
||||
playing as soon as it is ready while the next phrase is generated in parallel.
|
||||
|
||||
The provider intentionally uses half-duplex audio: microphone input is paused
|
||||
while a response is being transcribed, generated, synthesized, or played. This
|
||||
prevents speaker feedback from aborting TTS. Spoken interruption (barge-in) is
|
||||
|
||||
Reference in New Issue
Block a user