Showing 8 open source projects for "speak"

View related business solutions
  • $300 Free Credits to Build on Google Cloud Icon
    $300 Free Credits to Build on Google Cloud

    New customers can spin up VMs, build with AI, and query data at no cost.

    Put your $300 in credit toward real workloads, then keep building with free monthly usage for 20+ products. No commitment and no charge until you upgrade.
    Start Free
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • 1
    OpenMAIC

    OpenMAIC

    Open Multi-Agent Interactive Classroom

    ...The platform generates multiple learning scenes rather than a single static output, including slides, quizzes, interactive simulations, and project-based activities, which makes it feel closer to a guided lesson than a simple content generator. It also supports whiteboard-style visual explanation and text-to-speech delivery, allowing agents to draw, explain formulas, and speak aloud during instruction. OpenMAIC is built for flexible deployment, with support for direct use through its web experience as well as integration with OpenClaw so classrooms can be generated from messaging platforms such as Slack, Telegram, and Feishu.
    Downloads: 10 This Week
    Last Update:
    See Project
  • 2
    Polyglot

    Polyglot

    Cross-platform AI language practice app

    ...Users can define custom AI personas, choose languages, and configure their own OpenAI and Azure keys so they retain control over which backends they use. The app supports speech recognition with quick keyboard shortcuts, allowing learners to hold down a key to speak and release it to submit for recognition and response. It includes translation features, dark mode, playback of the user’s own recorded speech, and word highlighting that tracks the progress of synthesized audio to make following along easier. Polyglot also integrates additional AI providers, supports configurable conversation scenarios, and lets users personalize avatars, making the experience more engaging and flexible.
    Downloads: 7 This Week
    Last Update:
    See Project
  • 3
    ElatoAI

    ElatoAI

    Realtime AI Voice Agents with SoTA Multimodal AI models on Arduino ESP

    ...The system integrates voice synthesis and recognition by connecting an ESP32 device through secure WebSockets to edge server functions written in Deno, allowing users to speak naturally with AI agents hosted through cloud APIs including OpenAI’s Realtime API, Gemini’s Live API, xAI’s Grok Voice Agent API, and others. It includes a web client (built with Next.js) for managing devices, controlling volume, and viewing conversation transcripts, while the hardware runs optimized firmware to deliver responses in near real time — even supporting >15-minute uninterrupted conversations.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    GetProfile

    GetProfile

    User profile and long-term memory for your AI agent

    ...The goal is to make “memory” and user understanding an infrastructure concern rather than an app-by-app feature, so teams can add continuity with minimal code changes. Because it behaves like an OpenAI-compatible gateway, it can work with multiple providers and tools that already speak that API shape.
    Downloads: 0 This Week
    Last Update:
    See Project
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 5
    Ring

    Ring

    Unofficial packages for Ring Doorbells, Cameras, Alarm System

    ...With Ring you can control your home from your smartphone, tablet or PC. Each Ring device includes a camera, speakers, and an integrated microphone so you can view, listen, and speak to anyone on your property from anywhere. Ring's customizable motion sensors allow you to focus on the most important areas of your home. You will receive instant warnings as soon as your Ring device detects movement, so you are always the first to know if someone has gotten too close to your property. Ring allows you to monitor every corner of your property.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 6
    Actionhero

    Actionhero

    Actionhero is a realtime multi-transport nodejs API server

    ...Actionhero can work in a cluster to handle all the clients you can throw at it. Actionhero was built to serve the same APIs across multiple protocols. Do your games speak both HTTP and Websockets? Actionhero has got you covered. Actionhero was built from the ground up to include all the features you expect from a modern API framework.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7
    Voconly

    Voconly

    Free, open-source, fully offline speech-to-text tool for Windows.

    Voconly is a free, open-source offline speech to text tool for Windows that runs entirely on your device. Unlike cloud-based voice input software, Voconly keeps all audio data local—no uploads, no internet required, complete privacy. Powered by local AI, it combines real-time speech recognition with on-device LLM post-processing. Supports multiple ASR models (Qwen-ASR, SenseVoice, Whisper, Parakeet) and offers customizable refinement modes: auto-polishing, professional formatting,...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    Amica

    Amica

    Amica is an open source interface for interactive communication

    Amica is an open source interface for interacting with fully animated 3D characters that combine voice chat, vision, and an emotion engine into a single experience. It lets you hold natural conversations with AI characters that can see, listen, and speak, while expressing emotional states through facial expressions and body language. Users can import VRM character models, adjust their appearance, tune the voice to match the character, and define behavior using different large language models and TTS backends. Under the hood, Amica leverages modern web and desktop technologies: three.js and three-vrm for 3D rendering, Transformers.js for running models in the browser, Whisper and Silero VAD for speech recognition and voice-activity detection, and a variety of LLM backends such as llama.cpp servers, ChatGPT-compatible APIs, Ollama, KoboldCpp, and others. ...
    Downloads: 12 This Week
    Last Update:
    See Project
  • Previous
  • You're on page 1
  • Next