Voice Input

Voice Input brings GPT-4o Transcribe speech recognition directly into Obsidian. It uses a focused design: GPT-4o Transcribe for transcription and a fixed replacement dictionary for corrections.

For voice input in any macOS or Windows app, try Blitzmemo.

Installation

Install the plugin in either of these ways:

  1. From Obsidian: Open Settings → Community plugins → Browse, search for Voice Input, then install and enable it.
  2. From the Community site: Open the Voice Input Community Plugin page and select Add to Obsidian. Obsidian opens so you can install the plugin.

Settings

Enter your OpenAI API key, then select Test Connection. A successful test means the plugin is ready to use.

Voice Input settings in Obsidian

How to Use

Open the command palette and select Voice Input's Open View command. The Voice Input panel appears on the right.

Voice Input Modes

Select the start button once for continuous recording; recording continues until you select it again. For push-to-talk recording, press and hold the start button and release it when finished.
Continuous mode works well for longer passages, while push-to-talk is convenient for short comments.

When recording is complete, select the red stop button to begin transcription. To stop without transcribing, select Cancel; the recording is discarded without incurring an API charge.
You can start another recording while previous audio is still being transcribed.

Voice Input panel ready to record
Voice Input panel while recording

Replacement Dictionary

Enable dictionary correction, then add entries to the custom dictionary. Replacements are applied mechanically after transcription, making this useful for names or technical terms that are repeatedly misrecognized.

Voice Input replacement dictionary settings