Voice Input
Voice Input brings GPT-4o Transcribe speech recognition directly into Obsidian. It uses a focused design: GPT-4o Transcribe for transcription and a fixed replacement dictionary for corrections.
For voice input in any macOS or Windows app, try Blitzmemo.
Installation
Install the plugin in either of these ways:
- From Obsidian: Open Settings → Community plugins → Browse, search for
Voice Input, then install and enable it. - From the Community site: Open the Voice Input Community Plugin page and select
Add to Obsidian. Obsidian opens so you can install the plugin.
Settings
Enter your OpenAI API key, then select Test Connection. A successful test means the plugin is ready to use.

How to Use
Open the command palette and select Voice Input's Open View command. The Voice Input panel appears on the right.
Voice Input Modes
Select the start button once for continuous recording; recording continues until you select it again. For push-to-talk recording, press and hold the start button and release it when finished.
Continuous mode works well for longer passages, while push-to-talk is convenient for short comments.
When recording is complete, select the red stop button to begin transcription. To stop without transcribing, select Cancel; the recording is discarded without incurring an API charge.
You can start another recording while previous audio is still being transcribed.


Replacement Dictionary
Enable dictionary correction, then add entries to the custom dictionary. Replacements are applied mechanically after transcription, making this useful for names or technical terms that are repeatedly misrecognized.
