Voice Recognition

Vocalinux – voice dictation application

Vocalinux is a voice dictation application that lets you enter text into virtually any Linux application by speaking into a microphone.

It supports both X11 and Wayland and performs speech recognition locally, so audio and transcribed text do not need to leave your machine.

The application supports whisper.cpp, OpenAI Whisper, and VOSK speech-recognition engines. whisper.cpp is used by default and offers Vulkan acceleration on AMD, Intel, and NVIDIA graphics hardware. A GTK-based settings interface provides access to recognition, audio, shortcut, and advanced options.

This is free and open source software.

Key Features

  • Enters dictated text into Linux applications.
  • Supports X11 and Wayland sessions.
  • Provides fully offline speech recognition.
  • Offers whisper.cpp, OpenAI Whisper, and VOSK engines.
  • Supports Vulkan acceleration with AMD, Intel, and NVIDIA graphics hardware.
  • Includes toggle and push-to-talk activation modes.
  • Offers configurable global keyboard shortcuts.
  • Provides spoken commands for punctuation, line breaks, capitalisation, and deleting text.
  • Uses neural voice activity detection when ONNX Runtime is available.
  • Includes a system tray indicator and graphical settings interface.
  • Supports automatic startup through XDG autostart.
  • Offers command-line options and multiple recognition models.

Website: github.com/jatinkrmalik/vocalinux
Support:
Developer: Jatin K Malik
License: GNU General Public License v3.0

Vocalinux in action

Vocalinux is written in Python. Learn Python with our recommended free books and free tutorials.


Related Software

Speech Recognition Tools
WhisperAutomatic speech recognition (system trained on 680,000 hours of data
FlashlightFast, flexible machine learning library written entirely in C++.
Coqui STTDeep-learning toolkit for training and deploying speech-to-text models
KaldiC++ toolkit designed for speech recognition researchers.
SpeechBrainAll-in-one conversational AI toolkit based on PyTorch
HandyOffline speech-to-text application
ESPnetEnd-to-End speech processing toolkit
deepspeech.pytorchImplementation of DeepSpeech2 using Baidu Warp-CTC.
WhisperingTranscription application with global speech-to-text functionality
JuliusTwo-pass large vocabulary continuous speech recognition engine
CMUSphinxSpeech recognition system for mobile and server applications
SimonFlexible speech recognition software
hyprwhsprNative speech-to-text designed for Arch / Omarchy
osttOpen Speech-to-Text
DeepSpeechTensorFlow implementation of Baidu's DeepSpeech architecture.
OpenSeq2SeqTensorFlow-based toolkit for sequence-to-sequence models
EesenEnd-to-End Speech Recognition

Read our verdict in the software roundup.


Best Free and Open Source Software Explore our comprehensive directory of recommended free and open source software. Our carefully curated collection spans every major software category.

This directory is part of our ongoing series of informative articles for Linux enthusiasts. It features hundreds of detailed reviews, along with open source alternatives to proprietary solutions from major corporations such as Google, Microsoft, Apple, Adobe, IBM, Cisco, Oracle, and Autodesk.

You’ll also find interesting projects to try, hardware coverage, free programming books and tutorials, and much more.

Discovered a useful open source Linux program that we haven’t covered yet? Let us know by completing this form.
Subscribe

Please read our Comment Policy before commenting.

Notify of
guest
0 Comments
Oldest
Newest Most Voted