Showing 10 open source projects for "audio speaker software"

View related business solutions
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 1
    Open-XiaoAI

    Open-XiaoAI

    Let the Xiaoai speaker "hear your voice"

    Open-XiaoAI is an open-source framework that takes deeper control of compatible Xiaomi XiaoAI smart speakers by accessing their audio input and output paths. It extends the earlier MiGPT concept beyond cloud API responses and enables developers to build custom voice experiences directly around the speaker hardware. Demonstrations include custom wake words, Xiaozhi AI integration, MiGPT integration, Gemini Live API access, and stereo pairing. The architecture uses separate client and server components, with the client running on modified speaker firmware. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 2
    Note67

    Note67

    A private, local meeting notes assistant

    note67 is a private, local meeting notes assistant application that combines audio capture, transcription, and AI-powered summarization to help users document conversations and meetings on their own devices without relying on cloud services. Built with a cross-platform architecture using Rust (via Tauri) for backend logic and a TypeScript/React frontend, it prioritizes privacy by performing audio transcription locally with Whisper models and generating summaries with locally-hosted AI, eliminating the need to send sensitive meeting content to external servers. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    MusicGPT

    MusicGPT

    Generate music based on natural language prompts using LLMs

    ...Users can describe a musical style, mood, or instrumentation using text prompts, and the system produces original audio samples based on those instructions. The application currently integrates with models such as MusicGen and is designed to support additional models transparently in the future. In addition to a command-line interface, the project includes a web-based interface that enables conversational interaction with the AI model.
    Downloads: 18 This Week
    Last Update:
    See Project
  • 4
    Ruma

    Ruma

    A set of Rust crates for interacting with the Matrix chat network

    Matrix is an open specification for an online communication protocol. It includes all the features you'd expect from a modern chat platform including instant messaging, group chats, audio and video calls, searchable message history, synchronization across all your devices, and end-to-end encryption. Matrix is federated, so no single company controls the system or your data. You can use an existing server you trust or run your own, and the servers synchronize messages seamlessly. Learn more...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Paessler - Monitor Your Whole Network in Minutes Icon
    Paessler - Monitor Your Whole Network in Minutes

    Auto-discovery finds your devices and deploys pre-configured sensors instantly. No project plan required, just visibility from day one.

    Waiting weeks for a monitoring rollout isn't an option when infrastructure doesn't stop running. PRTG's auto-discovery scans your network and suggests from over 200 pre-configured sensor types, so you're watching servers, applications and devices within minutes, not after a multi-week deployment. Enterprise-strength monitoring, without the enterprise complexity. Start your free trial today.
    Start Free 30-Day Trial
  • 5
    Rig

    Rig

    Rust framework for building modular and scalable LLM-powered apps

    ...Rig includes built-in support for agent workflows, allowing systems to perform multi-turn reasoning, tool calling, and retrieval-based tasks within structured pipelines. It also supports capabilities such as text generation, embeddings, transcription, image generation, and audio generation depending on the provider used. Developers can integrate language models into their software with minimal boilerplate while maintaining flexibility for complex AI workflows.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    auricle

    auricle

    Native Windows music player with search, library and audio caching.

    ...Free software under GNU GPLv3.
    Downloads: 18 This Week
    Last Update:
    See Project
  • 7
    Mididash

    Mididash

    MIDI router with a node-based interface and Lua scripting

    Mididash is an open source MIDI routing software with a node-based interface and Lua scripting. A modern take on programs like MIDI-OX.
    Downloads: 13 This Week
    Last Update:
    See Project
  • 8
    OmniPull

    OmniPull

    Just pull anything

    OmniPull is a powerful, cross-platform download manager built with Python and PySide6. It provides a modern, intuitive interface for managing downloads with advanced features like multi-threading, queue management, and media extraction.
    Downloads: 6 This Week
    Last Update:
    See Project
  • 9
    Voconly

    Voconly

    Free, open-source, offline speech-to-text tool for Windows and MacOS.

    ....** It turns speech into text locally and uses AI to polish, translate, organize, and transform your words into usable content, letting you replace keyboard input with your voice in any application. No audio uploads. No cloud dependency. Your voice stays on your device. Unlike cloud-based voice input software, voconly keeps all audio data local—no uploads, no internet required, complete privacy. Powered by local AI, it combines real-time speech recognition with on-device LLM post-processing. Supports multiple ASR models (Qwen-ASR, SenseVoice, Whisper, Parakeet) and offers customizable refinement modes: auto-polishing, professional formatting, cross-language translation. ...
    Downloads: 14 This Week
    Last Update:
    See Project
  • Build Securely on Azure with Proven Frameworks Icon
    Build Securely on Azure with Proven Frameworks

    Lay a foundation for success with Tested Reference Architectures developed by Fortinet’s experts. Learn more in this white paper.

    Moving to the cloud brings new challenges. How can you manage a larger attack surface while ensuring great network performance? Turn to Fortinet’s Tested Reference Architectures, blueprints for designing and securing cloud environments built by cybersecurity experts. Learn more and explore use cases in this white paper.
    Download Now
  • 10
    Glicol

    Glicol

    Graph-oriented live coding language and music/audio DSP library

    Glicol is a graph-oriented live coding language and audio engine designed for real-time music creation and digital signal processing, written entirely in Rust. It introduces a unique paradigm where audio synthesis and sequencing are represented as interconnected nodes, allowing developers and musicians to construct complex sound pipelines through declarative code. The language is designed to be accessible to beginners while still offering powerful capabilities for advanced users, enabling...
    Downloads: 1 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • Next