Showing 18 open source projects for "microphone"

View related business solutions
  • One Monitoring Tool for IT, OT and Cloud | Free Trial Icon
    One Monitoring Tool for IT, OT and Cloud | Free Trial

    Vendor-agnostic monitoring across on-prem servers, cloud platforms and OT devices, all in one dashboard. No more tool sprawl.

    Modern infrastructure spans data centers, cloud platforms and factory floors, and every blind spot between them is a risk. PRTG supports SNMP, WMI, SSH and other standard protocols to monitor IT, OT and hybrid environments through one customizable dashboard. Build the views your team needs, from network health to application performance, without switching tools. Try PRTG free for 30 days now.
    Try PRTG Free
  • Build Data Resilience - Take the Assessment Today Icon
    Build Data Resilience - Take the Assessment Today

    Can you recover when it matters most? Take this quick assessment to identify gaps and build greater recovery confidence.

    Is your recovery strategy as strong as you think? Take this quick self-assessment to check your recovery readiness and gain tailored insights. In only 2 minutes, you'll learn where you fall on the recovery readiness scale.
    Take the Assessment
  • 1
    Buzz

    Buzz

    Buzz transcribes and translates audio offline

    Buzz is a desktop application for transcribing and translating audio locally with speech recognition models based on Whisper. It can process audio files, video files, and YouTube links without requiring cloud transcription. Live microphone transcription supports real-time captions and a presentation view for accessible events. Speech separation can improve results on noisy recordings, while speaker identification distinguishes voices within transcribed media. Multiple Whisper backends, Transformer models, and GPU acceleration options provide flexibility across different computers. ...
    Downloads: 617 This Week
    Last Update:
    See Project
  • 2
    SpeechRecognition

    SpeechRecognition

    Speech recognition module for Python

    ...The first software requirement is Python 2.6, 2.7, or Python 3.3+. This is required to use the library. PyAudio is required if and only if you want to use microphone input (Microphone). PyAudio version 0.2.11+ is required, as earlier versions have known memory management bugs when recording from microphones in certain situations. To hack on this library, first make sure you have all the requirements listed in the "Requirements" section.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 3
    VERBI

    VERBI

    A modular voice assistant application for experimenting

    ...The project integrates hosted services such as OpenAI, Groq, Deepgram, ElevenLabs, and Cartesia while also supporting local model workflows through Ollama and Piper. It records speech from a microphone, converts it into text, generates a response, and plays synthesized audio. Centralized configuration and environment variables manage providers, model selections, endpoints, and API credentials. The repository is intended for voice technology research, rapid prototyping, and comparison of different processing pipelines. It requires Python 3.10 or newer and includes scripts and sample assets for running the assistant.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 4
    WhisperLive

    WhisperLive

    A nearly-live implementation of OpenAI's Whisper

    ...The project supports multiple inference backends, including Faster-Whisper, NVIDIA TensorRT, and OpenVINO, allowing you to target GPUs and different CPU architectures efficiently. It can handle microphone input, pre-recorded audio files, and network streams such as RTSP and HLS, making it flexible for live events, monitoring, or accessibility workflows. Configuration options let you control the number of clients, maximum connection time, and threading behavior so the server can be tuned for different deployment environments. On the client side, you can set the language, whether to translate into English, model size, voice activity detection, and output recording behavior.
    Downloads: 8 This Week
    Last Update:
    See Project
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • 5
    Whisper-WebUI

    Whisper-WebUI

    A Web UI for easy subtitle using whisper model

    ...The platform integrates optimized implementations such as faster-whisper, significantly improving transcription speed and reducing memory usage compared to standard models. It supports multiple input sources including local files, YouTube content, and microphone input, making it versatile for different workflows. Whisper WebUI also includes advanced preprocessing and postprocessing features such as voice activity detection, background music separation, and speaker diarization, enabling more accurate and structured outputs.
    Downloads: 14 This Week
    Last Update:
    See Project
  • 6
    RealtimeSTT

    RealtimeSTT

    A robust, efficient, low-latency speech-to-text library

    RealtimeSTT is a Python-based realtime speech-to-text engine emphasizing low latency, wake-word detection, voice activity detection, and automatic speech segmentation. It provides asynchronous callbacks, nanosecond-precision timestamps, and CLI tools, suitable for building voice assistants, meeting transcribers, or live caption systems.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    Real-Time AI Voice Chat

    Real-Time AI Voice Chat

    Have a natural, spoken conversation with AI

    Real-Time AI Voice Chat is a client-server application for holding low-latency spoken conversations with large language models. The browser captures microphone audio and streams it over WebSockets to a Python backend. RealtimeSTT converts speech into text, which is sent to an LLM such as Ollama or OpenAI for response generation. RealtimeTTS then synthesizes the answer and streams speech back to the browser. Dynamic silence detection improves turn taking, while interruption handling lets users speak over an ongoing response. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    Screen Recorder Studio

    Screen Recorder Studio

    Screen Recorder Studio Screen Recorder Studio is a modern, lightweigh

    ...Key Features 🎥 High-quality screen recording 🎮 Gameplay recording with minimal performance impact 🖥️ Full-screen, window, and custom-region capture 🎙️ Microphone and system audio recording ⚡ Fast startup and lightweight resource usage 📁 Easy video export and file management 🎨 Modern, clean, user-friendly interface Vision join our discord in order to suggest or to communicate: https://discord.gg/YRhcgEGVjN
    Downloads: 4 This Week
    Last Update:
    See Project
  • 9
    clone-voice

    clone-voice

    A sound cloning tool with a web interface, using your voice

    ...The tool supports around sixteen languages, including Chinese, English, Japanese, Korean, French, German, Italian, and others, and can capture reference voices directly from a microphone or from uploaded audio.
    Downloads: 8 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 10
    SonicCTRL

    SonicCTRL

    DJ Sound Mixer

    ...It offers real-time audio mixing with crossfader (Power/Linear curve) and auto-fade, a 15-band graphic equalizer with waterfall spectrum display, BPM analysis with beat sync, loop and cue points, pitch control, an 8-pad sampler, and a master recorder (export as WAV, FLAC or MP3). SonicCTRL also supports microphone talkover with automatic music ducking, line-in pass-through for external devices, and up to four simultaneous extra outputs (e.g. multiple Bluetooth speakers at once) with per-output latency compensation. The built-in music library manages playlists and multiple searchable music archives (USB sticks, internal/external drives), imports tracks via M3U8, and converts WAV/FLAC/OGG/AIFF files to MP3. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 11
    Internet DJ Console

    Internet DJ Console

    A feature packed DJ console and internet radio client for Linux users

    Conceived as an internet radio Shoutcast/Icecast client and DJ console IDJC has two main media players, a background track player, effects buttons, crossfader, webm, aac, ogg, and mp3 streaming, stream automation timers, aux input, voice and VoIP integration. Media file formats include: mp3, ogg, flac, wma, wav, m4a, m3u, xspf, pls, and cue sheet support, IRC track and station announcements, uses jack audio connection kit to provide a flexible audio chain. This list of features is by no...
    Downloads: 14 This Week
    Last Update:
    See Project
  • 12
    Kalliope

    Kalliope

    Kalliope is a framework to create your own personal assistant

    ...But, if you need a particular module, you can write it by yourself, add it to your project, and propose it to the community. Kalliope can run on all Linux Debian-based distributions including a Raspberry Pi and its multi-lang. The only thing you need is a microphone.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 13
    tom_core

    tom_core

    tom_core - a tool for automating events on a computer

    tom_core is a software tool used for the automation of everything that happens on your computer. By using this application, you can easily record your activity on your computer, starting the recording at any moment that you choose. The application repeats all your clicks or drags, keystrokes, hotkeys, etc. All in exactly the timing and number of repetitions you need. The toolbox such as the optical recognition and voice control enables to branch out the recordings into complex forms, with...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    Green Recorder

    Green Recorder

    A simple screen recorder for Linux desktop

    Green Recorder is a desktop screen recording application designed for Linux systems, providing a simple interface for capturing screen activity and audio. It supports recording in multiple formats by leveraging FFmpeg and other backend tools to encode output efficiently. The application allows users to record full screens or specific areas, making it suitable for tutorials and demonstrations. It includes options for selecting audio sources and controlling frame rates to balance quality and...
    Downloads: 3 This Week
    Last Update:
    See Project
  • 15

    Distant Speech Recognition

    Beamforming and Speech Recognition Toolkit

    BTK contains C++ and Python libraries that implement speech processing and microphone array techniques such as speech feature extraction, speech enhancement, speaker tracking, beamforming, dereverberation and echo cancellation algorithms. The Millennium ASR provides C++ and python libraries for automatic speech recognition. The Millennium ASR implements a weighted finite state transducer (WFST) decoder, training and adaptation methods.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    PyEPL (Python Experiment-Programming Library) is a library for coding psychology experiments in Python. It supports presentation of both visual and auditory stimuli, and supports both manual (keyboard/joystick) and sound (microphone) input as responses.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 17
    Two programs for playing board games over the internet. A board viewer, where all clients see the same image with moveable tokens. And a chat client with calculator and dice roller.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18
    A little more than the command line, a lot less than Midnight Commander, and a home-row-based interface (I don't want to say vi-like, but then again, I just did). This is the file mangler I use, and I thought you might like it. Written in Python.
    Downloads: 0 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • Next