Desktop voice
System-wide speech-to-text, selected-text TTS, global hotkeys, VAD, tray controls, conversation HUD, and barge-in.
The open-source front door
Install system-wide dictation, selected-text speech, spoken LLM conversations, Piper, Faster Whisper, the bridge CLI, and the warm daemon under the MIT License.
Not a crippled demo
The free edition contains the core desktop and developer workflows. Pro is a focused upgrade for people who want richer voice, automation, persistence, and management.
System-wide speech-to-text, selected-text TTS, global hotkeys, VAD, tray controls, conversation HUD, and barge-in.
Streaming spoken conversations with any compatible endpoint, multiple profiles, sentence-level TTS, session headers, and failure detection.
Synthesis, transcription, JSON capability discovery, streaming PCM, cancellation, a warm daemon, and integration installers.
Naming made explicit
Version 0.5.5 also adds the fully spelled alias so users can install or launch the product using the name they heard.
pip install avrvxpip install avervox → installs avrvxavrvxavervoxInteraction modes
Use only the parts you need. Dictation and selected-text speech do not require an LLM endpoint.
Ctrl+Alt+Space
Speak into the focused application.
Ctrl+Alt+S
Hear the current text selection through Piper.
Ctrl+Alt+C
Hold a streaming spoken LLM conversation.
avrvx …
Use local speech in scripts and host applications.
OSS versus Pro
The same core architecture and integration contract remain in place when Pro is installed.
OSS questions
Start with the workflow
Use OSS to validate audio, hotkeys, text insertion, endpoint compatibility, and the voice workflow on your own Linux system.