Vocalinux
Free, open-source offline voice dictation for Linux with GPU acceleration
For Linux users who want private, offline voice dictation, Vocalinux is the strongest free option—GPU-accelerated, system-wide, and genuinely local. Setup needs a few terminal commands, but the interactive installer and Flatpak packaging lower the barrier. Pick it over cloud services if privacy matters more than smart NLP features. It's also a solid alternative to Nerd Dictation for a more integrated desktop experience.
Verified 6d ago · liveness 69/100 · cite: rightaichoice.com/tools/vocalinux
- Linux users seeking private, offline voice dictation
- Developers needing hands-free text input in terminals and IDEs
- Privacy-conscious users avoiding any cloud processing
- Users with RSI or accessibility needs requiring local speech-to-text
- Windows or macOS users (no support for those platforms)
- Beginners uncomfortable with the command line or installation steps
- Users needing cloud-based advanced NLP features (smart replies, translation)
We scan live Reddit threads, YouTube comments, X posts, G2 reviews and other communities — and hand you an honest verdict in under a minute.
- Honest verdict, not marketing
- Real pros & cons from real users
- Attributed quotes with receipts
3 free scans · no card needed
Skip Vocalinux if you're on Windows or macOS, need a mobile/web app, or require cloud-grade NLP like smart replies and translation—it's strictly a local Linux dictation tool.
No usage-based fees, but you must have at least 4GB RAM and ~200MB disk; larger models require more memory (8GB+).
Vocalinux is free (AGPL-3.0) with no premium tiers. It's ideal for individuals and teams on Linux who want offline dictation without subscription costs, undercutting cloud services like Otter.ai that charge monthly. On a budget, Vocalinux costs nothing but your time to set up.
In short
Vocalinux — Free, open-source offline voice dictation for Linux with GPU acceleration. Best for Linux users seeking private, offline voice dictation, Developers needing hands-free text input in terminals and IDEs, Privacy-conscious users avoiding any cloud processing. Free to use.
What's new in Vocalinux
Checked 6 days agoAcross the latest 5 updates: 3 feature updates and 2 changelog entries.
v0.15.0 Stable: Searchable sidebar settings, AppImage, expanded languages
Adds searchable sidebar settings with live search, AppImage packages, ~33 language catalog with auto-detect, auto-capitalize, auto-pause competing apps, and Vulkan device selection.
v0.14.2 Stable: IBus and GNOME Wayland fixes
Restores IBus engine process launch after Flatpak XDG path import and fixes first dictation drop on GNOME Wayland.
v0.14.1 Stable: Flatpak packaging, AUR package, layout-aware hotkeys
Introduces Flatpak packaging for universal distribution, an AUR package for Arch Linux, and layout-aware combo keys for non-US layouts.
v0.14.0-beta: Configurable hotkeys and FunASR/SenseVoice support
Adds configurable modifier+key hotkeys and FunASR/SenseVoice models support via Remote API, plus various Wayland/IBus fixes.
v0.13.0-beta: Guided whisper.cpp model selection
Introduces guided model selection with size and specialization dropdowns, preserving spaces between dictated segments.
What people actually say about Vocalinux — is it worth it?
We ran a structured research pass across product reviews, community discussions, and post-purchase forum threads to surface the patterns vendors won't publish themselves. Below: the recurring strengths, the hidden costs people mention most, and the cohort that consistently regrets adopting this tool.
36 mentions across 4 sources (Hacker News, YouTube, Bluesky, GitHub) · researched Jul 6, 2026.
- +100% offline and private — no data ever leaves your machine.
- +Supports multiple STT engines: whisper.cpp, VOSK, and OpenAI Whisper.
- +GPU acceleration via Vulkan works on AMD, Intel, and NVIDIA GPUs.
- +Compatible with both X11 and Wayland display servers.
- +One-command installer supports Ubuntu, Fedora, Debian, Arch, and openSUSE.
- −Frequent ghost/hallucination text after silence (GitHub issue reported).
- −Non-US keyboard layouts cause incorrect character injection with ydotool.
- −Installation often completes but the app fails to launch or respond.
- −Audio recording may start but never generate transcription (common bug).
- −38 open GitHub issues — many core functionality bugs unresolved.
- • No hidden costs, but requires time to troubleshoot bugs
Viability Score
How well maintained and how widely used is Vocalinux? Built from what the vendor actually publishes (docs, changelog, tutorials, integrations, pricing), whether the site is live, and how much real users discuss it. How we calculate this
Last calculated: August 2026
How we score →Key Features
- Offline voice dictation in any focused app
- whisper.cpp engine with Vulkan GPU acceleration
- OpenAI Whisper via PyTorch/CUDA
- VOSK engine for older machines
- Remote API engine (OpenAI-compatible or whisper.cpp server)
- FunASR/SenseVoice support via Remote API (v0.14.0+)
- Toggle mode and push-to-talk activation
- Configurable hotkeys with modifier+key combos (v0.14.1+)
- Silero VAD neural silence filter with amplitude fallback
- X11 and Wayland display support
- Flatpak packaging (v0.14.1+)
- AUR package for Arch (v0.14.1+)
- Suspend/resume recovery (v0.10.1+)
- Guided whisper.cpp model selection (v0.13.0+)
- Searchable sidebar settings (v0.15.0+)
About Vocalinux
Vocalinux is a free, open-source (AGPL-3.0) system tray app that brings system-wide voice dictation to Linux. It turns your voice into text in any focused application—terminals, browsers, IDEs, and office suites—using local speech engines that process audio on-device. That means no cloud upload, no telemetry, and no account required. You get toggle or push-to-talk activation, configurable hotkeys with modifier+key combos, neural voice activity detection (Silero VAD), and reliable desktop integration for both X11 and Wayland. The one-command interactive installer supports Ubuntu, Debian, Fedora, Arch (also via AUR), and openSUSE, and recent versions add Flatpak packaging. Vocalinux is engineered around hardware-aware engines: the default is whisper.cpp with Vulkan GPU acceleration for AMD, Intel, and NVIDIA, using a ~74MB Tiny model for fast local transcription. You can also switch to the original OpenAI Whisper (PyTorch/CUDA), VOSK for older machines, or a Remote API that talks to an OpenAI-compatible or whisper.cpp server you control—while keeping local VAD and injection intact. Recent updates (v0.14.2) add layout-aware combo keys, FunASR/SenseVoice support via Remote API, Flatpak packaging, and an AUR package, plus fixes for IBus and GNOME Wayland. With under 200MB disk footprint and a 4GB minimum RAM, Vocalinux is a practical choice for developers, writers, and anyone with RSI who wants hands-free, private dictation. It's the full desktop path Linux users never had: a tray icon, hotkeys, reliable injection, and engine options that match your hardware—not a cloud dashboard with a mic icon. Unlike cloud services like Otter.ai, Vocalinux is free, offline, and open—though you trade away advanced NLP like smart replies and translation. Compared to Nerd Dictation, it offers a more integrated desktop experience with a tray and hotkeys. It won't work on Windows or macOS (though sibling apps exist for those platforms).
Behind the Verdict
Vocalinux stands out in the Linux dictation space because it's not just a wrapper around a cloud API. It's a full desktop application with local processing, a tray icon, hotkeys, and reliable text injection. The hardware-aware engine selection is a genuine strength: whisper.cpp with Vulkan works across AMD, Intel, and NVIDIA GPUs, and you can switch to Whisper (CUDA) or VOSK for older machines. The Remote API option is a thoughtful addition for those who want to offload to their own server while keeping local VAD and injection. Recent updates show active development: Flatpak packaging, an AUR package, layout-aware combo keys, and FunASR/SenseVoice support via Remote API. The changelog also demonstrates attention to Wayland reliability, with multiple fixes for IBus and compositor edge cases. On the downside, vocalinux is Linux-only, and installation is command-line based, though the interactive installer helps. Non-English language support varies by model and engine. For privacy-conscious users and developers with RSI, Vocalinux is a compelling, free choice. Its main limitation is that it doesn't offer cloud-grade NLP features like smart replies or translation, so if you need those, you'd look at Otter.ai or similar. But for pure dictation into any app with local processing, it's hard to beat at $0.
Researching Vocalinux? Get your full AI stack in 60 seconds.
Free, no signup — tell us your goal and get tools matched to your budget & existing stack.
Real-world workflow fit
Concrete scenarios for the personas Vocalinux actually fits — and what changes day-one when you adopt it.
You want to dictate code comments and documentation into your IDE without typing.
Outcome: Install Vocalinux, choose whisper.cpp with Vulkan, and use push-to-talk to insert text into VS Code or a terminal with low latency.
You need to transcribe meeting notes or draft articles but refuse to send audio to the cloud.
Outcome: Run Vocalinux with the default whisper.cpp engine, ensuring all processing is local; use toggle mode to dictate into any text editor.
You have an older machine with limited RAM and want voice input without GPU acceleration.
Outcome: Select the VOSK engine during install to get a ~40MB model that runs on CPU, then use standard hotkeys to dictate.
Use Cases
- Dictate code comments and documentation into your IDE hands-free.
- Compose emails and reports without typing to reduce RSI strain.
- Transcribe meeting notes directly into a text editor.
- Control text input in any application while multitasking.
- Enable accessible computing for users with limited mobility.
- Use voice input in terminals and office applications on Linux.
Models Under the Hood
as of 2026-08-19
Limitations
- Vocalinux is designed for Linux desktops, supporting X11 and Wayland, and is not available on mobile, web, Windows, or macOS.
- Installation is command-line based, though an interactive installer and Flatpak packages are offered.
- Language support varies by engine and model availability, with about 33 languages plus auto-detect.
- It focuses on offline dictation and local processing, with no cloud-based NLP features like smart replies or translation.
as of 2026-08-17
Verification history
We have re-verified Vocalinux 5 times since . Each pass re-reads the vendor's own pages and re-checks every listed field against that evidence; passes where nothing had changed are marked as such.
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-checked, vendor evidence unchanged
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
- — re-verified summary, description, our verdict, our analysis, pricing model, pricing tiers, features, integrations, who it suits, who should skip it
Free to cite with attribution — this page re-verifies continuously.
12-month cost
Project the real annual outlay, including the implied monthly cost when only an annual tier is published.
Vendor list price only. Add-on usage, seat overages, and contract minimums are surfaced under Hidden costs & gotchas.
Plans compared
For each published Vocalinux tier: who it actually fits, and what it adds vs. the previous tier. Cross-reference the cost calculator above for projected annual outlay.
Open Source
$0
Ideal for
Anyone on Linux who wants free, offline, privacy-first voice dictation without subscription costs. Great for developers, writers, and users with RSI.
What this tier adds
Starting tier: free forever, AGPL-3.0 license, includes all local engines (whisper.cpp, Whisper, VOSK), Remote API, and system-wide injection on X11/Wayland.
Where the pricing makes sense
The company stage and team size where Vocalinux's pricing actually pencils out — and where peers do it cheaper.
Vocalinux is free (AGPL-3.0) with no premium tiers. It's ideal for individuals and teams on Linux who want offline dictation without subscription costs, undercutting cloud services like Otter.ai that charge monthly. On a budget, Vocalinux costs nothing but your time to set up.
Setup time & first value
How long it actually takes to get something useful out of Vocalinux — broken out by persona, not the marketing-page minute.
Most users can run the one-command interactive installer and be dictating within 5–10 minutes. The installer pulls dependencies and models automatically. For Flatpak or AUR, setup is even quicker, around 5 minutes. Larger models or custom engines may add a few minutes.
Switching to or from Vocalinux
How to bring data in from common predecessors and how to get it back out — written for the switcher, not the buyer.
- ↗To Nerd Dictation: If you want a simpler, lighter tool, you can use Nerd Dictation for basic offline dictation, but you'll lose the tray icon, hotkeys, and multi-engine support.
Resources & Guides
Tutorials & Learning
Official links
Tools that pair well with Vocalinux
Common stack mates teams adopt alongside Vocalinux, with the specific reason each pairing earns its keep.
Featured Head-to-Head Comparisons
Vocalinux vs Poke Interaction Co
Vocalinux and Poke serve completely different needs. Vocalinux is the right choice if you're a Linux user who wants private, offline voice dictation into any app—free and open-source. Poke is ideal if you want an AI assistant that lives in your messaging apps and manages email, calendar, tasks, and health data, with paid tiers for advanced automation. Your pick depends on whether you need pure voice input or a full personal assistant.
Vocalinux vs Guesty
Vocalinux and Guesty serve completely different needs. Vocalinux is a free, open-source voice dictation tool for Linux users requiring privacy and offline operation. Guesty is a paid property management platform for vacation rental hosts needing AI-driven automation. Choose based on your task: dictation vs. hospitality management.
Vocalinux vs Gem
Vocalinux and Gem serve entirely different purposes: one is a free, offline voice dictation tool for Linux, the other is a paid, AI-powered recruiting platform. Choose Vocalinux if you value privacy and need local speech-to-text on Linux; choose Gem if you're a recruiter seeking an all-in-one ATS/CRM with AI automation.
Alternatives to Vocalinux
View allOpen Wispr
Free, open-source, 100% local voice dictation for macOS — no cloud, no accounts, no telemetry.
OmniVoice Studio
Free, open-source local voice cloning, dubbing, and design for 600+ languages.
Frequently Asked Questions
Categories
Topics
Used Vocalinux? Help shape our editorial sentiment research.


