Raise Athena dictation limit to three minutes
This commit is contained in:
+3
-1
@@ -67,7 +67,9 @@ nicht mehr den aktuellen Containerzustand.
|
|||||||
Transkriptionspfad des Plugins. Plugin 1.3.0 zerlegt längere Browser-Diktate
|
Transkriptionspfad des Plugins. Plugin 1.3.0 zerlegt längere Browser-Diktate
|
||||||
während der Aufnahme in überlappende Sechs-Sekunden-Abschnitte und hält so
|
während der Aufnahme in überlappende Sechs-Sekunden-Abschnitte und hält so
|
||||||
den Abschluss innerhalb von OpenClaws festem Fünf-Sekunden-Fenster. OpenClaw
|
den Abschluss innerhalb von OpenClaws festem Fünf-Sekunden-Fenster. OpenClaw
|
||||||
selbst wurde dafür nicht gepatcht. OpenClaw liefert den fertigen Agententext
|
selbst wurde dafür nicht gepatcht. Die lokale Sicherheitsgrenze
|
||||||
|
`maxSpeechSeconds` steht produktiv auf 180 Sekunden. OpenClaw liefert den
|
||||||
|
fertigen Agententext
|
||||||
an die Brücke; das Sprechen beginnt daher erst nach Abschluss der
|
an die Brücke; das Sprechen beginnt daher erst nach Abschluss der
|
||||||
Agentenantwort. Details und Grenzen stehen in der [Voice-Doku](../services/athena-realtime-voice/README.md).
|
Agentenantwort. Details und Grenzen stehen in der [Voice-Doku](../services/athena-realtime-voice/README.md).
|
||||||
- `mike-ai-mikes-applio-ui` läuft gesund und ohne GPU. Die Quelle liegt im
|
- `mike-ai-mikes-applio-ui` läuft gesund und ohne GPU. Die Quelle liegt im
|
||||||
|
|||||||
@@ -43,7 +43,7 @@ Recommended `talk.realtime` configuration:
|
|||||||
"vadThreshold": 0.018,
|
"vadThreshold": 0.018,
|
||||||
"silenceDurationMs": 750,
|
"silenceDurationMs": 750,
|
||||||
"prefixPaddingMs": 300,
|
"prefixPaddingMs": 300,
|
||||||
"maxSpeechSeconds": 45
|
"maxSpeechSeconds": 180
|
||||||
}
|
}
|
||||||
}
|
}
|
||||||
}
|
}
|
||||||
@@ -84,6 +84,11 @@ provider for its URL/key. If that model provider has no key, it reuses
|
|||||||
origin. No second credential is needed. In `talk.catalog`, it appears under
|
origin. No second credential is needed. In `talk.catalog`, it appears under
|
||||||
`transcription.providers`.
|
`transcription.providers`.
|
||||||
|
|
||||||
|
The production provider allows up to 180 seconds per recording. This limit is
|
||||||
|
a local safety cap shared by dictation and Talk, not an OpenClaw or Whisper
|
||||||
|
restriction. Incremental segmentation keeps long dictation bounded while it
|
||||||
|
is being recorded.
|
||||||
|
|
||||||
### Voice-note file attachments
|
### Voice-note file attachments
|
||||||
|
|
||||||
An M4A voice note uploaded as a chat attachment does **not** use the realtime
|
An M4A voice note uploaded as a chat attachment does **not** use the realtime
|
||||||
|
|||||||
Reference in New Issue
Block a user