Raise Athena dictation limit to three minutes

This commit is contained in:
Mikei386 committed 2026-09-21 16:30:47 +02:00
1 parent 46e5bbdf7f
commit 33c04150e3
2 files changed
+9 -2

No files matched your search

+3 -1
View File
@@ -67,7 +67,9 @@ nicht mehr den aktuellen Containerzustand.
Transkriptionspfad des Plugins. Plugin 1.3.0 zerlegt längere Browser-Diktate Transkriptionspfad des Plugins. Plugin 1.3.0 zerlegt längere Browser-Diktate
während der Aufnahme in überlappende Sechs-Sekunden-Abschnitte und hält so während der Aufnahme in überlappende Sechs-Sekunden-Abschnitte und hält so
den Abschluss innerhalb von OpenClaws festem Fünf-Sekunden-Fenster. OpenClaw den Abschluss innerhalb von OpenClaws festem Fünf-Sekunden-Fenster. OpenClaw
selbst wurde dafür nicht gepatcht. OpenClaw liefert den fertigen Agententext selbst wurde dafür nicht gepatcht. Die lokale Sicherheitsgrenze
`maxSpeechSeconds` steht produktiv auf 180 Sekunden. OpenClaw liefert den
fertigen Agententext
an die Brücke; das Sprechen beginnt daher erst nach Abschluss der an die Brücke; das Sprechen beginnt daher erst nach Abschluss der
Agentenantwort. Details und Grenzen stehen in der [Voice-Doku](../services/athena-realtime-voice/README.md). Agentenantwort. Details und Grenzen stehen in der [Voice-Doku](../services/athena-realtime-voice/README.md).
- `mike-ai-mikes-applio-ui` läuft gesund und ohne GPU. Die Quelle liegt im - `mike-ai-mikes-applio-ui` läuft gesund und ohne GPU. Die Quelle liegt im
+6 -1
View File
@@ -43,7 +43,7 @@ Recommended `talk.realtime` configuration:
"vadThreshold": 0.018, "vadThreshold": 0.018,
"silenceDurationMs": 750, "silenceDurationMs": 750,
"prefixPaddingMs": 300, "prefixPaddingMs": 300,
"maxSpeechSeconds": 45 "maxSpeechSeconds": 180
} }
} }
} }
@@ -84,6 +84,11 @@ provider for its URL/key. If that model provider has no key, it reuses
origin. No second credential is needed. In `talk.catalog`, it appears under origin. No second credential is needed. In `talk.catalog`, it appears under
`transcription.providers`. `transcription.providers`.
The production provider allows up to 180 seconds per recording. This limit is
a local safety cap shared by dictation and Talk, not an OpenClaw or Whisper
restriction. Incremental segmentation keeps long dictation bounded while it
is being recorded.
### Voice-note file attachments ### Voice-note file attachments
An M4A voice note uploaded as a chat attachment does **not** use the realtime An M4A voice note uploaded as a chat attachment does **not** use the realtime