Voconly is a free, open-source offline speech to text tool for Windows that runs entirely on your device.
Unlike cloud-based voice input software, Voconly keeps all audio data local—no uploads, no internet required, complete privacy.
Powered by local AI, it combines real-time speech recognition with on-device LLM post-processing.
Supports multiple ASR models (Qwen-ASR, SenseVoice, Whisper, Parakeet) and offers customizable refinement modes: auto-polishing, professional formatting, cross-language translation.
Perfect for privacy-conscious users, developers, writers, and professionals handling sensitive content.
If you need reliable Windows voice input without cloud dependency, Voconly is your solution.
Features
- Fully Offline Operation — No internet connection required. The entire speech-to-text pipeline runs on your local device.
- Privacy-First Architecture — No audio data is ever uploaded to external servers. Your voice stays on your device at all times. Real-Time Transcription — Words appear on screen as you speak, with live progress tracking.
- Multiple ASR Models — Supports Qwen-ASR, SenseVoice, Whisper, and Parakeet. Choose the model that fits your needs.
- AI-Powered Refinement — On-device LLM post-processing for auto-polishing, professional formatting, and cross-language translation. One-Key Mode Switching — Seamlessly switch between polish, formatting, and translation workflows.
- Free & Open Source (MIT) — No subscription fees, no usage limits, no hidden costs.
License
MIT LicenseFollow Voconly
Other Useful Business Software
Build Agents and Models on One Platform
Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
Rate This Project
Login To Rate This Project
User Reviews
Be the first to post a review of Voconly!