macOS 14+ · Apple Silicon

Talk to your computer,
don't type.

See your mic go live. Speak, polish, and paste. Dictate locally on your Mac and talk to AI agents. No account. No subscription.

Download Free

100 free dictations · 50 notes · 10 meeting recordings · No credit card

or install via brew install muvon/tap/vext

macOS 14+ · Apple Silicon · No account required

How it works

One shortcut. One pill from microphone to paste.

1

Hold the hotkey

Hold your hotkey. Speak when the pill turns red: your mic is live.

2

Speak naturally

Speak naturally. Watch your words appear in the pill with local Parakeet.

3

Release to paste

Release. Follow Transcribing, then Polishing. Esc skips AI. Failed take? Keep the audio and Retry.

See every step before the paste.

Follow one pill from mic startup to pasted text. Meet the Press-to-Paste HUD in Vext 1.5.1.

  • Speak when the pill turns red. Audio ducking no longer delays capture.
  • Keep live text in view through transcription and polishing. Follow the elapsed clock and menu-bar processing icon.
  • Skip slow polishing or model loading with Esc. Paste as spoken.

Retry If a take fails, its audio is kept. Retry directly in the pill.

Also new in 1.5.1

  • Use Right-Cmd hands-free and Fn meeting toggles without triggering them through Right-Cmd+C or Fn+Delete.
  • See your own bindings in menu hints.
  • Keep your paid access offline.
  • Download the local AI model only when a style or translation needs it.
  • Use native pickers, localized counts and durations, and sidebar settings navigation.
  • Export dictations, notes, and meetings — TXT and Markdown, plus SRT and VTT for meetings.
  • Open System Settings from onboarding after a denied permission.
  • See the price in the Buy button.
Read the 1.5.1 release notes

Three modes. One app.

Dictate, record meetings, and capture notes. Keep it all local by default.

Dictation

Hold a hotkey, speak, release — text appears at your cursor. Works in any app, any text field.

Meetings

Record with speaker labels, get full transcripts, screenshots, and AI-generated summaries.

Notes

Quick voice remarks — transcribed, cleaned up, and stored locally in the app for later.

Choose your style.
Paste what you need.

Choose how your words read. Keep them As spoken, Clean up the wording, or choose a style for each app.

  • As spoken or Clean up
  • Email, Message, Bullet points, Formal, or Custom
  • Set a different dictation style for each app
  • Combine a style with translation; Esc skips AI while polishing
What you said

"so um I was thinking we should uh probably move the deadline to like next Friday because ah the team needs more time to finish the uh the integration tests"

What gets pasted

"We should move the deadline to next Friday — the team needs more time to finish the integration tests."

Dictation styles.

Choose As spoken, Clean up, Email, Message, Bullet points, Formal, or Custom. Set styles per app. Esc skips AI polishing.

As spoken

Original wording

Paste the raw transcript without AI reshaping.

Clean up

Light editing

Tidy grammar, punctuation, and filler words.

Email · Message

Per-app styles

Shape your dictation for correspondence or a concise chat reply.

Bullet points · Formal · Custom

Your choice

Organize a list, use a formal tone, or write your own instruction.

Your voice stays on your Mac.

Local by default. Speech and AI models run on Apple Silicon. Nothing is uploaded unless you add your own API key.

Local by default

Dictate with Parakeet TDT v3 through FluidAudio, or choose Apple Speech. Polish with a local MLX model.

Works offline

Download the models once, then dictate and polish offline. License validation connects to Vext’s server.

No account, no telemetry

Start without an account. Vext has no telemetry SDKs or usage tracking.

Try free. Pay once. Use forever.

100 free dictations, 50 notes, and 10 meeting recordings to start. Then one price, unlimited use.

Best Value

Vext

$49 once

$0 in year 2

  • Runs locally on-device
  • No account required
  • Meeting transcription
  • Lifetime access
  • Free updates within version
Buy Vext

Cloud voice tools

$10–30 /month

$120–360/year

  • Cloud-dependent
  • Account required
  • Usage caps
  • Subscription only
  • Privacy trade-offs

Unlock from within the app when you're ready. Free updates included. Major new versions at 50% off for existing owners.

How Vext compares.

Feature-by-feature against the leading voice and meeting tools.

Vext $49 onceWispr Flow $12–15/moGranola $14–35/moOtter.ai $8–17/mo
Dictation (paste at cursor)
Meeting transcription
Voice notes
Speaker labels
Cross-meeting voice recognitionN/A
AI text cleanup
Meeting summaries
Translation
Screenshot capture (any mode)
Screenshots auto-paste to AIN/AN/A
YOLO mode (auto-submit)N/AN/A
Live transcript while you speak
Per-app dictation styles
Fully local, no cloud required
Works offline
No bot joins your callN/A
Shows when the mic is live · Esc skips AI mid-takeN/AN/AN/A
Cost after 2 years$49$288–360$336–840$200–408

N/A: not applicable or not verified.
Updated 5 September 2026 for Vext 1.5.1 and the no-bot row. Other competitor features and prices reflect April 2026; they may change.

See your words as you speak.

Watch live text appear with local Parakeet. Follow transcription and polishing in the same pill, with an elapsed clock.

150x realtime

60 seconds of audio transcribed in ~400ms. On-device.

Parakeet local
150x
Apple local
25x
Gemini cloud
23x
OpenAI cloud
22x
AssemblyAI cloud
20x
Alex 00:12

Let's review the Q3 roadmap and figure out priorities.

Sarah 00:28

I think we should focus on the API redesign first. It's blocking three other teams.

Alex 00:45

Agreed. Can we have a draft by end of next week?

Meeting transcription.
And the summary.

Record any meeting — Zoom, Google Meet, FaceTime, or in-person — and get a full transcript with speaker identification. Turn on Summarize to extract key points and action items. Both versions are always saved.

  • Timestamps and per-speaker breakdowns
  • System audio + microphone capture
  • AI-powered key points and action items
  • Raw transcript always preserved
  • Get recording prompts with Zoom, Microsoft Teams, Slack, FaceTime, Discord, Webex, Skype, or Loom when the mic activates.

Label speakers once.
Recognized forever.

Vext detects every distinct voice in your meeting automatically. Name them once — and from your next call onward, the same person is identified, labeled, and color-coded without lifting a finger.

  • Automatic speaker detection in every recording
  • Label with custom names — saved to your library
  • Same voice auto-labeled in future meetings
  • Color-coded chips for fast transcript scanning
Meeting #1 Speakers
Them Sarah
Me John
Speaker 1 Jack
Meeting #2 Auto-labeled
Sarah
John
Jack

Voice + vision,
hands-free.

Capture any region of your screen during hands-free dictation. The screenshot pastes alongside your transcribed prompt — straight into Claude Code, Cursor, or any AI tool. Fully hands-free coding.

  • Drag to capture during hands-free dictation or meeting recording
  • Screenshot auto-pastes with your transcript — Claude Code, Cursor, ChatGPT
  • Combine voice + image without touching the keyboard
Vext website hero showing its voice-to-text introduction
2 min ago

Check the API rate limits before we push the new integration. Sarah mentioned the sandbox has different thresholds than production.

18 min ago

The onboarding flow needs a skip option on the second screen. Users are dropping off because they think the setup is mandatory.

1 hour ago

Try SwiftUI NavigationSplitView for the sidebar instead of the custom implementation. It handles state restoration automatically.

Capture a thought
before it's gone.

Press a key, say what's on your mind, and move on. Vext transcribes, cleans up, and stores your note locally — ready when you need it.

  • Optional AI cleanup and translation before saving
  • All notes stored locally in the app
  • No app switching — works from anywhere on your Mac

Speak one language.
Type another.

Speak, then translate before pasting. Choose from 29 target languages. Apply a dictation style in the same AI pass.

  • Translation after transcription
  • 29 target languages in Settings
  • Source-language support depends on your speech engine
  • Same hotkey workflow; choose your target language
English

"Let's schedule a meeting for next Tuesday to discuss the project roadmap and assign tasks to the team."

Russian

"Давайте назначим встречу на следующий вторник, чтобы обсудить план проекта и распределить задачи в команде."

Works everywhere you type.

Vext is a system-level service. It works in any text field, in any app — browsers, editors, terminals, email, chat.

AI Tools

Claude CodeChatGPTClaude.aiCursorCodex

Browsers

SafariChromeFirefoxArc

Editors

VS CodeXcodeSublime TextVim

Terminals

TerminaliTerm2WarpGhostty

Communication

SlackDiscordTelegramMessages

Productivity

NotionObsidianNotesGmail
New in 1.4

Your voice, in your agent’s hands.

Connect Claude Code, Codex CLI, OpenCode, Cursor/Windsurf, Claude Desktop, or Octomind through MCP. Dictate, transcribe files, and read meetings. Manage notes, speakers, and vocabulary.

  1. 1Enable the MCP server in Settings
  2. 2Copy the config for your client
  3. 3Ask your agent to use Vext
  • Dictate: start, stop, get the text — or paste it
  • Transcribe any audio file
  • Meetings: start, stop, transcript with speaker names, summary, export
  • Notes and dictation history
  • Speakers, vocabulary and per-app styles
  • Settings and microphone

Listens on 127.0.0.1 only. Every request needs the token Vext generates for you. Off by default.

Terminal
claude mcp add --transport http vext http://127.0.0.1:58398/mcp --header "Authorization: Bearer TOKEN"

Replace TOKEN with the token from Settings — or use the copy button there, which fills it in.

Go hands-free.

Press a key once to start dictation. Press again to stop. No holding required — just talk as long as you need. Perfect for longer passages or when your hands are busy. Bind it to a mouse thumb button, or let a quick tap of your push-to-talk key toggle it.

Standard Hold key → speak → release
Hands-free Press key → speak freely → press key

Capture screenshots mid-dictation — they auto-paste with your transcript. See how →

⌘
Press once
Speaking freely...
⌘
Press again

YOLO Mode.

Turn it on and Vext automatically presses Return after pasting your transcription. Speak, release, and your prompt is already running.

Stop editing. Stop polishing. Just talk. LLMs know what you mean, even when your words aren't perfect.

YOLO Mode
Speak
Review transcript
Edit / fix mistakes
Press Return
Done

Audio ducking.

When you start recording, Vext automatically fades your system audio so your voice comes through clearly. Release the hotkey and volume returns to normal. No manual adjustment needed.

Playing music
Recording
Resumed

Choose your engines.

Run speech and AI locally. Add your own API key if you want cloud models. Off by default.

Speech-to-Text

ModelTypeSpeedSize
Apple Dictation Built-in macOS speech recognition. Zero download.LocalRealtimeBuilt-in
OpenAI-compatible Choose an OpenAI-compatible speech API with your own key.APIVaries—

AI Processing Dictation styles · Translate · Summarize

ModelTypeSize
Gemma 3 1B Ultra-lightweight. Fastest local option, lower accuracy.Local~1 GB
Qwen 3 4B Strong multilingual support. Good for translation tasks.Local~2.5 GB
LLaMA 3.2 3B Meta LLaMA. Strong general-purpose performance.Local~2.4 GB
Phi-3.5 Mini Microsoft Phi-3.5. Compact, strong reasoning.Local~2.8 GB
OpenAI-compatible Choose an OpenAI-compatible AI model with your own key.API—

Built by Muvon.

Vext is built by Muvon Un Limited, a small product studio. We build tools we use every day — designed to work locally, respect your privacy, and get out of your way.

Questions & answers.