voice-input-src documents and exposes the generated source workflow for a macOS menu-bar speech input application. Holding the Fn key starts recording, and releasing it transcribes and pastes text into the currently focused field. It uses Apple's streaming Speech Recognition framework and requires macOS 14 or newer. The application supports English, Simplified Chinese, Traditional Chinese, Japanese, and Korean, with the chosen language stored locally. A floating capsule displays a live waveform and expanding transcription text while recording. Clipboard restoration and temporary input-method switching make text injection safer for CJK users. An optional OpenAI-compatible model can conservatively correct obvious recognition errors before the final text is inserted.

Features

  • Hold-to-record Fn key workflow
  • Streaming Apple speech recognition
  • Five selectable recognition languages
  • Live waveform and transcription overlay
  • Clipboard-safe text injection
  • Optional LLM transcription correction

Project Samples

Project Activity

See All Activity >

Categories

Libraries

License

MIT License

Follow voice-input-src

voice-input-src Web Site

Other Useful Business Software
Build Agents and Models on One Platform Icon
Build Agents and Models on One Platform

Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
Try It Free
Rate This Project
Login To Rate This Project

User Reviews

Be the first to post a review of voice-input-src!

Additional Project Details

Registered

2026-07-23