Dictation
Whisper Dictation records from the computer microphone and transcribes speech locally. It inserts text into the composer without sending the message.
Set up
Section titled “Set up”- Open Extensions.
- Enable Whisper Dictation.
- Download the multilingual model.
- Wait for the model to show as verified in the shared Hugging Face cache.
- Start dictation from the composer and approve microphone access when the operating system asks.
The bundled Extension engine and model weight are separate. Enabling the Extension exposes the action; downloading the model enables local inference. By default the pinned weight lives in the standard Hugging Face hub cache, so Grokship and other compatible software can reuse the same file. Use Settings → Dictation → Model storage to choose a per-app custom cache or restore the Hugging Face default. Switching libraries never moves or deletes a model.
Use dictation
Section titled “Use dictation”- Start recording.
- Speak normally in a supported language.
- Stop recording.
- Review and edit the inserted text.
- Send only when it says what you intend.
Audio remains local to the app. The model remains local to the computer but may be shared by compatible apps. Removing it can make those apps download it again. Silence, very short recordings, or missing microphone permission can prevent transcription.
macOS users can see macOS permissions for recovery steps. Windows and Linux use their native microphone and audio controls.