diff --git a/docs/LIVE_STATE.md b/docs/LIVE_STATE.md index 0043c65..bc06156 100644 --- a/docs/LIVE_STATE.md +++ b/docs/LIVE_STATE.md @@ -67,7 +67,9 @@ nicht mehr den aktuellen Containerzustand. Transkriptionspfad des Plugins. Plugin 1.3.0 zerlegt längere Browser-Diktate während der Aufnahme in überlappende Sechs-Sekunden-Abschnitte und hält so den Abschluss innerhalb von OpenClaws festem Fünf-Sekunden-Fenster. OpenClaw - selbst wurde dafür nicht gepatcht. OpenClaw liefert den fertigen Agententext + selbst wurde dafür nicht gepatcht. Die lokale Sicherheitsgrenze + `maxSpeechSeconds` steht produktiv auf 180 Sekunden. OpenClaw liefert den + fertigen Agententext an die Brücke; das Sprechen beginnt daher erst nach Abschluss der Agentenantwort. Details und Grenzen stehen in der [Voice-Doku](../services/athena-realtime-voice/README.md). - `mike-ai-mikes-applio-ui` läuft gesund und ohne GPU. Die Quelle liegt im diff --git a/integrations/openclaw-athena-talk/README.md b/integrations/openclaw-athena-talk/README.md index 1534da2..3c32583 100644 --- a/integrations/openclaw-athena-talk/README.md +++ b/integrations/openclaw-athena-talk/README.md @@ -43,7 +43,7 @@ Recommended `talk.realtime` configuration: "vadThreshold": 0.018, "silenceDurationMs": 750, "prefixPaddingMs": 300, - "maxSpeechSeconds": 45 + "maxSpeechSeconds": 180 } } } @@ -84,6 +84,11 @@ provider for its URL/key. If that model provider has no key, it reuses origin. No second credential is needed. In `talk.catalog`, it appears under `transcription.providers`. +The production provider allows up to 180 seconds per recording. This limit is +a local safety cap shared by dictation and Talk, not an OpenClaw or Whisper +restriction. Incremental segmentation keeps long dictation bounded while it +is being recorded. + ### Voice-note file attachments An M4A voice note uploaded as a chat attachment does **not** use the realtime