goodroot/hyprwhspr
Native speech-to-text for Linux - Fast, accurate, private, and hackable system-wide dictation
What it solves
hyprwhspr provides a native, system-wide speech-to-text (STT) dictation tool for Linux users. It eliminates the need for cloud-dependent dictation by offering high-performance local transcription that is private, fast, and integrates directly into any active text buffer on the OS.
How it works
The tool captures audio via the microphone and processes it using a variety of configurable backends. It supports local in-memory models for offline privacy and speed, as well as cloud APIs for those who prefer them. It utilizes ONNX-ASR for CPU optimization and supports CUDA or Vulkan acceleration for GPU-powered transcription. Once transcribed, the text is automatically pasted into the active window using clipboard and window tools like wl-clipboard or xclip.
Who it’s for
Linux users (Arch, Debian, Ubuntu, Fedora, openSUSE) who want a fast, private, and hackable dictation system that works across both Wayland and X11 sessions.
Highlights
- Multiple Backend Support: Compatible with Cohere, Parakeet, Whisper, Qwen3-ASR, Gemini, and ElevenLabs.
- Privacy First: Offers strictly offline transcription so data never leaves the machine.
- Linux Native: Built specifically for Linux with native AUR packages and systemd integration.
- High Performance: Optimized for both high-end NVIDIA GPUs and CPU-only hardware via ONNX.
- User Experience: Includes a themed voice visualizer, audio ducking to reduce background noise during recording, and auto-paste functionality.
- Multilingual: Strong support for CJK (Chinese, Japanese, Korean) and built-in translation capabilities.
Related
- Project
- Project
- Project
- Project
- Project