Showing 166 open source projects for "microphone"

View related business solutions
  • Veeam Data Platform v13.1 Icon
    Veeam Data Platform v13.1

    Move workloads across hypervisors and clouds with no vendor lock-in. Try VDP free today.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try Now
  • Ship Agents Faster Icon
    Ship Agents Faster

    Transform your applications and workflows into powerful agentic systems at global scale.

    Gemini Enterprise Agent Platform lets you rapidly build, scale, govern and optimize production-ready agents grounded in your organization's data. The platform enables developers to build custom or pre-built agents for virtually any use case. New customers get $300 in free credits.
    Start Free
  • 1
    HaishinKit

    HaishinKit

    Camera and Microphone streaming library via RTMP and SRT for iOS, Mac

    Camera and Microphone streaming library via RTMP and SRT for iOS, macOS, tvOS and visionOS.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 2
    Buzz

    Buzz

    Buzz transcribes and translates audio offline

    Buzz is a desktop application for transcribing and translating audio locally with speech recognition models based on Whisper. It can process audio files, video files, and YouTube links without requiring cloud transcription. Live microphone transcription supports real-time captions and a presentation view for accessible events. Speech separation can improve results on noisy recordings, while speaker identification distinguishes voices within transcribed media. Multiple Whisper backends, Transformer models, and GPU acceleration options provide flexibility across different computers. ...
    Downloads: 617 This Week
    Last Update:
    See Project
  • 3
    SpeechRecognition

    SpeechRecognition

    Speech recognition module for Python

    ...The first software requirement is Python 2.6, 2.7, or Python 3.3+. This is required to use the library. PyAudio is required if and only if you want to use microphone input (Microphone). PyAudio version 0.2.11+ is required, as earlier versions have known memory management bugs when recording from microphones in certain situations. To hack on this library, first make sure you have all the requirements listed in the "Requirements" section.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 4
    Hyprnote

    Hyprnote

    Local-first AI Notepad for Private Meetings

    Hyprnote is an open-source, privacy-first AI notepad app designed for taking notes during meetings—transcribing audio (microphone and system) and generating context-rich summaries using on-device AI models like Whisper and HyprLLM, all without any data leaving your machine.(turn0search7, turn0search1). Listens to your meetings while you write. Crafts smart summaries based on your quick notes. Runs completely offline using open-source models like Whisper or HyprLLM.
    Downloads: 11 This Week
    Last Update:
    See Project
  • Paessler - Monitor Your Whole Network in Minutes Icon
    Paessler - Monitor Your Whole Network in Minutes

    Auto-discovery finds your devices and deploys pre-configured sensors instantly. No project plan required, just visibility from day one.

    Waiting weeks for a monitoring rollout isn't an option when infrastructure doesn't stop running. PRTG's auto-discovery scans your network and suggests from over 200 pre-configured sensor types, so you're watching servers, applications and devices within minutes, not after a multi-week deployment. Enterprise-strength monitoring, without the enterprise complexity. Start your free trial today.
    Start Free 30-Day Trial
  • 5
    quill macOS

    quill macOS

    Ultraminimalist macOS recording + transcription

    Quill is a minimalist macOS meeting recorder and transcription utility that runs from the menu bar. It captures microphone input and system audio as separate tracks, which naturally distinguishes the user from other speakers. When recording stops, both tracks are transcribed locally and merged into a timestamped, speaker-labeled transcript. The app uses an on-device Parakeet model, so recordings and text never need to leave the Mac. Sessions are stored as audio, metadata, transcript, and log files in organized folders. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 6
    VERBI

    VERBI

    A modular voice assistant application for experimenting

    ...The project integrates hosted services such as OpenAI, Groq, Deepgram, ElevenLabs, and Cartesia while also supporting local model workflows through Ollama and Piper. It records speech from a microphone, converts it into text, generates a response, and plays synthesized audio. Centralized configuration and environment variables manage providers, model selections, endpoints, and API credentials. The repository is intended for voice technology research, rapid prototyping, and comparison of different processing pipelines. It requires Python 3.10 or newer and includes scripts and sample assets for running the assistant.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 7
    Vision Camera

    Vision Camera

    The Camera library that sees the vision

    VisionCamera was designed from the ground up to provide all features a camera app should have. You have full control over what device is used, and can even configure options such as frame rate, colorspace, and more. While having a lot of features, VisionCamera makes sure you don't get overwhelmed from the beginning. It provides hooks and functions to help you get started faster, and if you need full control, you can easily do that. Every functionality has been thoroughly documented and even...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 8
    Screendrop

    Screendrop

    A beautiful screenshot + screen recording + Loom alternative

    ...It can capture entire displays, windows, selected regions, or text through on-device OCR. Screenshots can be annotated with drawing tools, crops, redaction, blur, backgrounds, borders, and watermarks. Screen recordings can include separate camera, microphone, and system-audio tracks for later editing. Its editor supports zoom effects, reconstructed cursors, keystroke captions, clip trimming, speed changes, transcription, and social-media aspect ratios. Screendrop stores projects locally by default and can optionally share them through a self-hosted Cloudflare setup. It also integrates with Apple Shortcuts, Siri, Spotlight, and configurable global hotkeys.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 9
    WhisperLive

    WhisperLive

    A nearly-live implementation of OpenAI's Whisper

    ...The project supports multiple inference backends, including Faster-Whisper, NVIDIA TensorRT, and OpenVINO, allowing you to target GPUs and different CPU architectures efficiently. It can handle microphone input, pre-recorded audio files, and network streams such as RTSP and HLS, making it flexible for live events, monitoring, or accessibility workflows. Configuration options let you control the number of clients, maximum connection time, and threading behavior so the server can be tuned for different deployment environments. On the client side, you can set the language, whether to translate into English, model size, voice activity detection, and output recording behavior.
    Downloads: 8 This Week
    Last Update:
    See Project
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • 10
    Kooha

    Kooha

    Elegantly record your screen

    Capture your screen in an intuitive and straightforward way without distractions. Kooha is a simple screen recorder with a minimal interface. You can simply click the record button without having to configure a bunch of settings.
    Downloads: 15 This Week
    Last Update:
    See Project
  • 11
    WebCord

    WebCord

    A Discord and SpaceBar :electron:-based client

    ...WebCord does a lot to improve the privacy of the users. It blocks known tracing and fingerprinting methods, but it does not end on it. It also manages the permissions to sensitive APIs like camera or microphone, sets its own user agent to the one present in Chromium browsers and spoof web API modifications in order to prevent distinguishing it from the real Chrome/Chromium browsers.
    Downloads: 53 This Week
    Last Update:
    See Project
  • 12
    GladiaFlow

    GladiaFlow

    A desktop app for real-time voice dictation

    GladiaFlow is an open-source desktop app for turning spoken language into text in virtually any text field. It captures microphone audio through a global hotkey and streams it to Gladia’s Live Transcription API. Partial and final results arrive in real time, then the app cleans punctuation, capitalization, and duplicate text before pasting the transcript into the active application. Users can choose push-to-talk or toggle activation and configure languages, code switching, vocabulary, and pronunciations. ...
    Downloads: 14 This Week
    Last Update:
    See Project
  • 13
    Skill Recorder

    Skill Recorder

    Desktop app that records your on-screen work session

    Skill Recorder is a Microsoft desktop application that turns a recorded work session into reusable instructions for AI agents. It captures screen activity, clicks, application changes, visited pages, and optional spoken narration while a user performs a task. GitHub Copilot CLI analyzes the recording and reconstructs the overall intent plus an ordered sequence of steps. Users can review and edit that analysis before generating a reusable Skill or scheduled Automation. The generated procedure...
    Downloads: 20 This Week
    Last Update:
    See Project
  • 14
    CAVA

    CAVA

    Cross-platform Audio Visualizer

    ...Choose from several preset settings of incredible colors or create your own. CAVA is a bar spectrum audio viewer based on my own open source project with the same name. Take the audio from the device's microphone and visualize the amplitude of the different frequencies as bars on the screen. Each bar represents a certain bandwidth of low to high frequencies. The leftmost bar starts at 50 Hz and the rightmost bar ends at 10 kHz. Although the frequencies outside this spectrum are audible, they do not contribute much to the overall sound image. ...
    Downloads: 18 This Week
    Last Update:
    See Project
  • 15
    Mumble

    Mumble

    Mumble is an open-source, low-latency, high quality voice chat

    Mumble is an open-source, low-latency, high-quality voice chat software. There are two modules in Mumble; the client (mumble) and the server (murmur). The client works on Windows, Linux, FreeBSD, OpenBSD, and macOS, while the server should work on anything Qt can be installed on. Low-latency and high-quality voice-chat program written on top of Qt and Opus. Administrators appreciate Mumble for being able to self-host and have control over data security and privacy. Some make use of the...
    Downloads: 24 This Week
    Last Update:
    See Project
  • 16
    WO Mic

    WO Mic

    Transform your smartphone into a PC microphone

    WO Mic is a free utility that turns your smartphone into a functional microphone for your Windows PC. It eliminates the need to buy a separate microphone, offering a convenient and cost-effective solution for voice chat, recording, or wireless voice control. The app supports multiple connection types including Wi-Fi, Bluetooth, and USB, giving users flexible options to suit their setup. Setup involves installing the mobile app and the PC client with drivers, which is straightforward and fast. ...
    Downloads: 1,291 This Week
    Last Update:
    See Project
  • 17
    ScreenPipe

    ScreenPipe

    AI app store powered by 24/7 desktop history. open source

    Screenpipe is an AI app store powered by continuous desktop history recording. It operates entirely locally, offering developers a platform to build, distribute, and monetize AI applications that leverage comprehensive contextual data from users' desktop activities. ​
    Downloads: 2 This Week
    Last Update:
    See Project
  • 18
    Whisper-WebUI

    Whisper-WebUI

    A Web UI for easy subtitle using whisper model

    ...The platform integrates optimized implementations such as faster-whisper, significantly improving transcription speed and reducing memory usage compared to standard models. It supports multiple input sources including local files, YouTube content, and microphone input, making it versatile for different workflows. Whisper WebUI also includes advanced preprocessing and postprocessing features such as voice activity detection, background music separation, and speaker diarization, enabling more accurate and structured outputs.
    Downloads: 14 This Week
    Last Update:
    See Project
  • 19
    RealtimeSTT

    RealtimeSTT

    A robust, efficient, low-latency speech-to-text library

    RealtimeSTT is a Python-based realtime speech-to-text engine emphasizing low latency, wake-word detection, voice activity detection, and automatic speech segmentation. It provides asynchronous callbacks, nanosecond-precision timestamps, and CLI tools, suitable for building voice assistants, meeting transcribers, or live caption systems.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 20
    Snapcast

    Snapcast

    Synchronous multiroom audio player

    Snapcast is a multiroom client-server audio player, where all clients are time synchronized with the server to play perfectly synced audio. It's not a standalone player, but an extension that turns your existing audio player into a Sonos-like multiroom solution. Audio is captured by the server and routed to the connected clients. Several players can feed audio to the server in parallel and clients can be grouped to play the same audio stream. One of the most generic ways to use Snapcast is...
    Downloads: 28 This Week
    Last Update:
    See Project
  • 21
    Background Music

    Background Music

    Automatically pause your music, set individual apps' volumes, etc.

    ...With Background Music running, launch QuickTime Player and select File > New Audio Recording (or New Screen Recording, New Movie Recording). Then click the dropdown menu next to the record button and select Background Music as the input device. You can record system audio and a microphone together by creating an aggregate device that combines your input device (usually Built-in Input) with the Background Music device.
    Downloads: 10 This Week
    Last Update:
    See Project
  • 22
    Real-Time AI Voice Chat

    Real-Time AI Voice Chat

    Have a natural, spoken conversation with AI

    Real-Time AI Voice Chat is a client-server application for holding low-latency spoken conversations with large language models. The browser captures microphone audio and streams it over WebSockets to a Python backend. RealtimeSTT converts speech into text, which is sent to an LLM such as Ollama or OpenAI for response generation. RealtimeTTS then synthesizes the answer and streams speech back to the browser. Dynamic silence detection improves turn taking, while interruption handling lets users speak over an ongoing response. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    Note67

    Note67

    A private, local meeting notes assistant

    ...Built with a cross-platform architecture using Rust (via Tauri) for backend logic and a TypeScript/React frontend, it prioritizes privacy by performing audio transcription locally with Whisper models and generating summaries with locally-hosted AI, eliminating the need to send sensitive meeting content to external servers. Users can record meetings directly from their microphone, view live transcriptions, filter by speaker, and export structured summaries, making it useful for professionals who need searchable, organized records of discussions. It also features thoughtful signal processing such as voice activity detection and echo deduplication to improve transcription accuracy, and provides standard note-taking features.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 24
    ESP32-CAM_MJPEG2SD

    ESP32-CAM_MJPEG2SD

    ESP32 Camera motion capture application to record JPEGs to SD card

    Application for ESP32 / ESP32S3 with OV2640 / OV5640 camera to record JPEGs to SD card as AVI files and playback to the browser as an MJPEG stream. The AVI format allows recordings to replay at the correct frame rate on media players. If a microphone is installed then a WAV file is also created and stored in the AVI file. The ESP32 cannot support all of the features as it will run out of heap space. For better functionality and performance, use one of the new ESP32S3 camera boards, eg Freenove ESP32S3 Cam, and ESP32S3 XIAO Sense, but avoid no-name boards marked ESPS3 RE:1.0.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 25
    Recorder

    Recorder

    HTML5 js recording mp3 wav ogg webm amr format

    ​Supports microphone recording and real-time processing in most of the implemented getUserMediamobile and PC browsers, mainly including Chrome, Firefox, Safari, iOS 14.3+, Android WebView, Tencent Android X5 kernel (QQ, WeChat, Mini Program WebView) , uni-app (App, H5), and most Android phones updated after 2021 have their own browsers; do not support: UC-based kernel (typical Alipay), most of the old domestic mobile phones that have not been updated have their own browsers and any other form of browser (including PWA, WebClip, any App) on low-version iOS (11.0-14.2) except Safari inside page). ...
    Downloads: 4 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • 2
  • 3
  • 4
  • 5
  • Next