**voconly is a free, open-source, local-first AI voice input assistant that runs entirely on your device.**
It turns speech into text locally and uses AI to polish, translate, organize, and transform your words into usable content, letting you replace keyboard input with your voice in any application. No audio uploads. No cloud dependency. Your voice stays on your device.
Unlike cloud-based voice input software, voconly keeps all audio data local—no uploads, no internet required, complete privacy.
Powered by local AI, it combines real-time speech recognition with on-device LLM post-processing.
Supports multiple ASR models (Qwen-ASR, SenseVoice, Whisper, Parakeet) and offers customizable refinement modes: auto-polishing, professional formatting, cross-language translation.
Perfect for privacy-conscious users, developers, writers, and professionals handling sensitive content.
If you need reliable Windows voice input without cloud dependency, voconly is your solution
Features
- Fully Offline Operation — No internet connection required. The entire speech-to-text pipeline runs on your local device.
- Privacy-First Architecture — No audio data is ever uploaded to external servers. Your voice stays on your device at all times. Real-Time Transcription — Words appear on screen as you speak, with live progress tracking.
- Multiple ASR Models — Supports Qwen-ASR, SenseVoice, Whisper, and Parakeet. Choose the model that fits your needs.
- AI-Powered Refinement — On-device LLM post-processing for auto-polishing, professional formatting, and cross-language translation. One-Key Mode Switching — Seamlessly switch between polish, formatting, and translation workflows.
- Free & Open Source (MIT) — No subscription fees, no usage limits, no hidden costs.