Hyprwhspr
Free, open-source dictation for Linux with local AI models
Hyprwhspr is the strongest open-source dictation tool for Linux. If you're comfortable with CLI setup and want full control, it's a top pick. However, if you need a graphical interface or out-of-the-box installation, consider Talon or cloud-based tools.
Verified 2d ago · liveness 69/100 · cite: rightaichoice.com/tools/hyprwhspr
- Developers who dictate code on Linux
- Writers and journalists wanting hands-free typing
- Linux power users on Wayland or X11
- Privacy-conscious users requiring local processing
- Users who need a graphical configuration interface
- Beginners uncomfortable with CLI setup
- Users seeking a precompiled binary for quick install
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip Hyprwhspr if you need a graphical interface or a precompiled binary; it's CLI-focused and built from source.
No hidden costs—it's free and open source, but you may need to pay for cloud API usage if you opt-in to streaming for higher accuracy.
Hyprwhspr is free for everyone, making it a cost-effective choice compared to cloud dictation services that charge per-minute or subscription fees. For teams needing support, you might consider paid alternatives like Talon.
In short
Hyprwhspr — Free, open-source dictation for Linux with local AI models. Best for Developers who dictate code on Linux, Writers and journalists wanting hands-free typing, Linux power users on Wayland or X11. Free to use.
What people actually say about Hyprwhspr — is it worth it?
We ran a structured research pass across product reviews, community discussions, and post-purchase forum threads to surface the patterns vendors won't publish themselves. Below: the recurring strengths, the hidden costs people mention most, and the cohort that consistently regrets adopting this tool.
13 mentions across 2 sources (Hacker News, GitHub) · researched Jul 6, 2026.
- +Fast, accurate local transcription using cutting-edge models like Cohere Transcribe.
- +Privacy-first: all inference runs locally, no data leaves your machine.
- +Deep Wayland integration: visualizer overlay, Waybar status, per-app paste rules.
- +GPU acceleration auto-detected for NVIDIA, AMD, and Intel Vulkan.
- +Multiple recording modes: toggle, push-to-talk, continuous, long-form, auto.
- −Audio device changes require manually restarting the systemd service.
- −Terminal escape sequence fragments appear after paste injection in some setups.
- −Requires a GPU for optimal performance with large models.
- −No progress indicator or spinner for long transcription tasks.
- −First recording start may trigger pulseaudio timeout errors.
- • Requires a GPU for best performance; may need to purchase one.
- • Optional cloud API usage (Gemini, OpenAI) may incur costs from those providers.
Viability Score
How well maintained and how widely used is Hyprwhspr? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: September 2026
How we score →Key Features
- Local speech-to-text with Parakeet TDT V3, Cohere Transcribe, Whisper family
- GPU acceleration (CUDA, Vulkan) with CPU fallback
- Models held hot in memory for instant transcription
- Five recording modes: toggle, push-to-talk, auto, continuous, long-form
- Audio ducking lowers system volume while recording
- Paste text via wtype or ydotool with per-app paste rules
- Themed visualizer overlay for Hyprland, Sway, KDE; notifications on GNOME
- Waybar live status indicator (idle, recording, processing, error)
- Multilingual support with translate-to-English mode
- Text processing: word overrides, filler word removal, symbol replacements, custom prompts
- WebSocket streaming via Google Gemini, ElevenLabs, OpenAI, or similar
- Works on Wayland and X11 (Hyprland, GNOME, KDE, Sway)
- Self-healing from suspend/resume, mic unplug, keyboard hotplug
- Post-transcription hook to pipe output through custom shell commands
- Scriptable CLI for start/stop/toggle/cancel, plus evdev hotkeys
About Hyprwhspr
Hyprwhspr is a free, open-source (MIT) dictation tool for Linux that works system-wide on Wayland and X11. It lets you speak into any application, with transcription handled locally by default—your audio never leaves your machine. You can choose from top local models like Parakeet TDT V3, Cohere Transcribe, and the full Whisper family, with GPU acceleration auto-detecting NVIDIA CUDA, AMD/Intel Vulkan, or falling back to fast CPU inference via onnx-asr. Models stay hot in memory for instant transcription. The tool offers five recording modes (toggle, push-to-talk, auto, continuous, long-form), audio ducking, per-app paste rules, a themed visualizer overlay, and Waybar integration. It's fully scriptable via CLI and supports WebSocket streaming to cloud providers like Gemini, OpenAI, and ElevenLabs if you want maximum accuracy. Built for developers, writers, and privacy-conscious users, Hyprwhspr is hackable and free forever with no telemetry.
Behind the Verdict
Hyprwhspr excels in privacy and performance on Linux. Its local-first approach means your audio never leaves your machine by default, a major win for privacy-conscious users. The support for multiple local model families (Parakeet, Cohere, Whisper) and GPU auto-detection (CUDA, Vulkan) ensures fast transcription on capable hardware, with a CPU path via onnx-asr. The five recording modes and audio ducking are thoughtful touches for daily use. However, the lack of a precompiled binary and graphical configuration tools means it's not for beginners; you'll need comfort with the command line and building from source. For developers who dictate code, the per-app paste rules and CLI scripting are powerful. The option to stream to cloud providers like Gemini or OpenAI is a flexible add-on, but requires your own API keys. Compared to Talon, which is more polished with a GUI, Hyprwhspr is more hackable and free. Overall, it's a fantastic tool for Linux power users who value privacy and customization.
Researching Hyprwhspr? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas Hyprwhspr actually fits — and what changes day-one when you adopt it.
You want to dictate code into your IDE without lifting your hands from the keyboard.
Outcome: Install Hyprwhspr, set up a push-to-talk hotkey, and you can speak code snippets that get pasted directly into your editor with per-app rules.
You need to transcribe meeting notes into a document quickly.
Outcome: Use automatic mode to capture speech, with audio ducking to reduce background noise, and transcriptions appear in your notes app.
You want speech-to-text without sending audio to the cloud.
Outcome: Hyprwhspr processes everything locally by default, so your audio stays on your machine, and you can still use cloud providers if you choose.
Use Cases
- Dictate code and commands hands-free in your IDE or terminal
- Transcribe meeting notes or journal entries into any text field
- Control your system with voice via custom scripts and hotkeys
- Translate spoken input to English in real-time for multilingual work
- Automate text input in applications using continuous mode
Models Under the Hood
as of 2026-09-01
Limitations
- Hyprwhspr is a system-wide speech-to-text tool designed for Linux, supporting Wayland and X11 across Hyprland, GNOME, KDE Plasma, and Sway.
- It operates primarily via CLI and hotkeys, with no graphical settings interface mentioned beyond themed overlays and Waybar tray indicators.
- GPU acceleration relies on compatible drivers (NVIDIA CUDA, AMD/Intel Vulkan) and falls back to CPU.
- Cloud features require you to supply your own API keys (e.g., Gemini, OpenAI, ElevenLabs), which are stored securely and never touch config files.
as of 2026-08-31
Verification history
We have re-verified Hyprwhspr 7 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-checked, vendor evidence unchanged
- — re-checked, vendor evidence unchanged
- — re-checked, vendor evidence unchanged
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
Showing the 6 most recent of 7 verification passes.
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published Hyprwhspr tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Free & Open Source
$0/mo
Ideal for
Anyone who wants free, private, system-wide dictation on Linux and is comfortable with CLI setup.
What this tier adds
Starting tier: includes all features, local models, and optional cloud streaming with your own API key.
Where the pricing makes sense
The company stage and team size where Hyprwhspr's pricing actually pencils out — and where peers do it cheaper.
Hyprwhspr is free for everyone, making it a cost-effective choice compared to cloud dictation services that charge per-minute or subscription fees. For teams needing support, you might consider paid alternatives like Talon.
Setup time & first value
How long it actually takes to get something useful out of Hyprwhspr — broken out by persona, not the marketing-page minute.
For developers: ~15 minutes from installing dependencies to first transcription on Arch via AUR. For beginners: up to 1 hour if building from source and configuring Wayland/X11.
Switching to or from Hyprwhspr
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- →From Talon: You can replace Talon with Hyprwhspr if you prefer local models and CLI control; you'll need to adapt your voice commands.
- ↗To Talon: If you need a graphical configuration and more polished experience, you can switch to Talon, which offers a GUI and broader app support.
Integrations
Resources & Guides
Tutorials & Learning
Official links
Tools that pair well with Hyprwhspr
Common stack mates teams adopt alongside Hyprwhspr, with the specific reason each pairing earns its keep.
OmniVoice Studio
Free, open-source, local-first voice cloning, design, dubbing, and dictation for 646 languages.
Whisper
Open-source speech-to-text that transcribes 99+ languages and translates to English, free to run locally or via API.
Pyvideotrans
Free open-source video translation and AI dubbing, 30+ languages, offline-ready
Featured Head-to-Head Comparisons
Hyprwhspr vs Soniox
Hyprwhspr is a free, private, Linux-only dictation tool for Wayland users who want local processing and GPU acceleration. Soniox is a paid cloud API for developers building real-time multilingual voice apps with low latency and enterprise compliance. Choose Hyprwhspr if you're on Linux and need hands-free typing; choose Soniox if you need a scalable, language-rich speech API.
Hyprwhspr vs Retell Ai
Hyprwhspr and Retell AI serve completely different niches: Hyprwhspr is a free, privacy-first dictation tool for Linux power users, while Retell AI is a paid enterprise platform for automating phone conversations. Choose Hyprwhspr if you need offline, system-wide speech-to-text on Wayland. Choose Retell AI if your goal is scaling call center operations with AI agents.
Hyprwhspr vs Voiceitt
Choose Hyprwhspr if you're a Linux Wayland user needing fast, private, local dictation with flexible modes and GPU acceleration—it's free and developer-friendly. Choose Voiceitt if you or your users have non-standard speech patterns (e.g., cerebral palsy, ALS, accents) and require a trained, inclusive voice AI with meeting captioning integrations. They solve completely different problems.
Alternatives to Hyprwhspr
View allOmniVoice Studio
Free, open-source, local-first voice cloning, design, dubbing, and dictation for 646 languages.
Whisper
Open-source speech-to-text that transcribes 99+ languages and translates to English, free to run locally or via API.
Pyvideotrans
Free open-source video translation and AI dubbing, 30+ languages, offline-ready
Frequently Asked Questions
Used Hyprwhspr? Help shape our editorial sentiment research.


