
Free, 100% local macOS menu-bar app that turns speech into structured markdown using whisper.cpp and any local LLM.
Free, 100% local macOS menu-bar app that turns speech into structured markdown using whisper.cpp and any local LLM.
Speech-to-Markdown is a free, open-source macOS menu-bar app (with a companion iOS app for iPhone 15 Pro+ / iOS 26+) that streams your voice into clean, structured markdown in real time. It runs entirely on-device — whisper.cpp handles transcription and any local LLM server you already run (via Ollama or similar) handles structuring — so no cloud, no API keys, and nothing leaves your Mac. It offers two modes: Global Dictation (⌘⌥]) types the transcript straight into any app at your cursor (Terminal, browser, Slack), while Agent Mode opens a floating capsule that streams your speech through the LLM into a live Markdown/plain-text/HTML document. Installation is a single curl-piped shell script that pulls whisper-cpp, ffmpeg, and xcodegen via Homebrew and drops the app in /Applications. The iOS build is fully offline using Apple Intelligence.

Speech To Markdown is a free, 100% local macOS menu-bar app that converts spoken language into structured Markdown formatting. Utilizing whisper.cpp and any local large language model (LLM), it allows users to efficiently transcribe and format their speech into easily editable text.
Speech To Markdown is designed for macOS users who need a reliable tool to transcribe speech into text formatted in Markdown. Markdown is widely used for note-taking, documentation, and web publishing due to its simplicity and versatility. With this app, users can easily convert their spoken words into structured text, making it perfect for content creation, coding, and writing.
The app leverages whisper.cpp, an efficient implementation of OpenAI's Whisper model, known for its high accuracy in speech recognition. Users can customize their experience by integrating any local large language model that fits their needs, ensuring that transcription aligns with specific terminologies or styles.
By utilizing Speech To Markdown, users can significantly streamline their writing process, saving time and enhancing productivity while ensuring their speech is accurately captured and formatted.
Speech To Markdown works by utilizing a fully local pipeline, which combines speech recognition and language model processing to convert spoken language into structured text. It features a global dictation hotkey, real-time structuring, and an offline iOS app for private note-taking, ensuring complete data security without cloud reliance.
The Speech To Markdown tool leverages advanced technologies to transform voice into text efficiently. Here’s how it works:
Local Pipeline: The application operates entirely on your device, using whisper.cpp for speech-to-text conversion and a local language model (LLM) server for document structuring. This ensures privacy and security, as no data is sent to the cloud.
Global Dictation Hotkey: Users can activate dictation in any application by pressing ⌘⌥]. This feature allows for streamlined text entry directly into your cursor position, making it ideal for quick notes or long replies in platforms like Slack or email without needing any additional software.
Agent Mode Live Structuring: This feature streams your spoken words into a live document. Users can dictate freely, and the tool will format the transcribed text into Markdown, plain text, or HTML instantly, enhancing productivity during meetings or brainstorming sessions.
One-Line Install: Installation is simplified with a single curl-piped script that automatically installs necessary components like xcodegen, whisper-cpp, and ffmpeg via Homebrew, then builds the app directly into your Applications folder.
iOS Companion App: The offline app for iPhone and iPad utilizes Apple Intelligence, ensuring you can capture notes securely during meetings without internet access. This is particularly useful for sensitive content that must remain on the device.
Voice-Driven Coding Comments: Developers can dictate function documentation or pull request descriptions directly into their code editor, enhancing workflow efficiency.
Structured Journaling and Field Notes: Users can ramble freely in Agent Mode, producing well-structured Markdown documents. Offline functionality on iOS allows for capturing voice notes without a network connection.
Speech To Markdown features a fully local pipeline for speech-to-text conversion, global dictation hotkey, real-time document structuring, a one-line installation process, and an offline companion app for iOS devices. These features make it a powerful tool for efficient transcription and document creation without relying on the cloud.
Speech To Markdown utilizes a local processing system that runs on your device, eliminating the need for cloud services. The key features include:
100% Local Pipeline: This feature uses whisper.cpp for speech-to-text conversion and connects to a local LLM (Large Language Model) server for structuring the output. This ensures complete privacy and faster performance since no internet connection is required.
Global Dictation Hotkey: Users can activate transcription by pressing ⌘⌥] from any application, enabling seamless integration with tools like Terminal, web browsers, and Slack. This hotkey allows you to dictate notes, messages, or code directly where you want them.
Agent Mode Live Structuring: This unique feature provides a floating capsule interface that processes your speech in real-time, converting it into Markdown, plain text, or HTML formats. It allows users to see the transcription unfold as they speak, making it ideal for creating documentation, articles, or coding tasks.
One-Line Install: Installation is streamlined with a single curl-piped script that sets up necessary dependencies like xcodegen, whisper-cpp, and ffmpeg using Homebrew. This one-step process simplifies the initial setup, allowing users to get started quickly.
iOS Companion: The iOS app, compatible with iOS 26+ (iPhone 15 Pro and later), functions offline by leveraging Apple’s AI capabilities. This allows users to dictate notes or transcribe voice memos directly on their iPhones and iPads.
Speech To Markdown is ideal for professionals and creatives who need efficient voice-to-text solutions. It serves users looking for private meeting notes, voice-driven coding comments, structured journaling, offline field notes on iOS, and composing long-form messages in Slack or Email without needing extra transcription tools.
Speech To Markdown caters to a diverse range of users by offering functionalities that streamline note-taking and content creation. Here's how different groups can benefit:
Private Meeting Notes: For professionals needing to document sensitive meeting discussions, Speech To Markdown allows users to dictate notes directly on their Mac, ensuring that confidential information remains secure and doesn't leave the device. This is particularly beneficial for legal, medical, or corporate environments where data privacy is crucial.
Voice-Driven Coding Comments: Developers can use this tool to enhance their coding workflow. By utilizing Global Dictation, users can easily speak function documentation or pull request (PR) descriptions directly into their code editor. This feature saves time and allows for seamless coding without interrupting the flow of work.
Structured Journaling: With Agent Mode, users can freely express their thoughts while the tool structures the content into a clean Markdown document in real time. This is perfect for writers, bloggers, or anyone engaged in reflective journaling who prefers speaking over typing.
Offline Field Notes on iOS: Users can capture voice notes on their iPhone, even without signal. The app uses on-device Apple Intelligence to structure these notes into Markdown, making it a reliable choice for field researchers or creatives working in remote locations.
Slack / Email Long-Form: Dictating lengthy responses directly into Slack or Email can significantly boost productivity. Users can compose detailed replies without switching between applications, making communication more efficient.
Speech To Markdown is entirely free to use. Users can access this powerful tool without any associated costs, making it an excellent choice for individuals and businesses looking to convert speech into text effortlessly and economically.
Speech To Markdown is a versatile tool that converts spoken language into written text. This free-to-use software is particularly beneficial for content creators, journalists, and professionals who need to transcribe interviews, meetings, or lectures quickly.
For example, a journalist can use Speech To Markdown during an interview to capture quotes accurately, saving time on manual transcription. Similarly, educators can transcribe lectures for students who may need to review them later.
Common pitfalls include relying solely on the transcription without verification and neglecting to check for punctuation, which can alter meaning.
To get started with Speech To Markdown, visit https://voice-to-md.xajik0.workers.dev/ to sign up for an account. Once registered, you can explore the features and capabilities of the Speech To Markdown tool, allowing you to efficiently convert speech into text format.
Getting started with Speech To Markdown involves a straightforward process:
Browse by use case: Voice & Audio
Compare Speech To Markdown: vs FluentDB · vs ReExplain · vs YC Has It · vs OpenCode Superapp