Showing 134 open source projects for "speak"

View related business solutions
  • Custom VMs From 1 to 96 vCPUs With 99.95% Uptime Icon
    Custom VMs From 1 to 96 vCPUs With 99.95% Uptime

    General-purpose, compute-optimized, or GPU/TPU-accelerated. Built to your exact specs.

    Live migration and automatic failover keep workloads online through maintenance. One free e2-micro VM every month.
    Try Free
  • Train ML Models With SQL You Already Know Icon
    Train ML Models With SQL You Already Know

    BigQuery automates data prep, analysis, and predictions with built-in AI assistance.

    Build and deploy ML models using familiar SQL. Automate data prep with built-in Gemini. Query 1 TB and store 10 GB free monthly.
    Try Free
  • 1
    ChatterBot

    ChatterBot

    Machine learning, conversational dialog engine for creating chat bots

    ...For more details about the ideas and concepts behind ChatterBot see the process flow diagram. The language independent design of ChatterBot allows it to be trained to speak any language. Additionally, the machine-learning nature of ChatterBot allows an agent instance to improve it’s own knowledge of possible responses as it interacts with humans and other sources of informative data. An untrained instance of ChatterBot starts off with no knowledge of how to communicate. Each time a user enters a statement, the library saves the text that they entered and the text that the statement was in response to. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 2
    Mockito

    Mockito

    Most popular Mocking framework for unit tests written in Java

    ...In particular anyone using Kotlin (which demands using mockito-inline) and PowerMock (which exacerbates the problem even more) will want to add this to all of their test classes to avoid a large memory leak. Fancy getting world-wide visibility and building up an eternal fame of an OSS contributor? Use the latest version! Hack and experiment. Speak up at the mailing list. Mockito is a mocking framework that tastes really good. It lets you write beautiful tests with a clean & simple API. Mockito doesn’t give you hangover because the tests are very readable and they produce clean verification errors.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 3
    TEN

    TEN

    Open-source framework for conversational voice AI agents

    ...It includes a full ecosystem, TEN Turn Detection, TEN Agent, and TMAN Designer, allowing developers to rapidly assemble human-like, responsive agents that can see, speak, hear, and interact. With support for languages like Python, C++, and Go, it offers flexible deployment on both edge and cloud environments. Using components like graph-based workflow design, drag-and-drop UI (via TMAN Designer), and reusable extensions such as real-time avatars, RAG (Retrieval-Augmented Generation), and image generation, TEN enables highly customizable, scalable agent development with minimal code.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    Step-Audio-EditX

    Step-Audio-EditX

    LLM-based Reinforcement Learning audio edit model

    ...Because the model is trained with a “large-margin learning” objective over many synthesized and natural speech samples, it gains robust control over expressive attributes, and can perform iterative editing: e.g. you could record a line, then ask the model to “make it sadder,” “speak slower,” or “change accent to X.”
    Downloads: 0 This Week
    Last Update:
    See Project
  • Build Agents and Models on One Platform Icon
    Build Agents and Models on One Platform

    Everything you need to build production-ready agents and models. Access 200+ Google and third-party AI models and tools.

    Gemini Enterprise Agent Platform is Google Cloud's comprehensive platform for developers to build, scale, govern, and optimize agents and models. Choose from Google's most advanced models and third-party models like Anthropic's Claude Model Family.
    Try It Free
  • 5

    emubns

    Braille 'n Speak emulator

    Downloads: 0 This Week
    Last Update:
    See Project
  • 6
    SpeakLogPSU
    SpeakLogPSU can speak chat messages with an individual voice if the NPC or player was configured or with a default one. You will never miss if someone talks to you. Voice cloning can be accomplished with Coqui in less than five minutes without GPU. The result is archived and can be used the next time in game. Some TTS projects already started to add tag support to speak text with emotions or sing it.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7

    opentypeless

    Open-source AI voice typing for macOS, Windows, and Linux.

    OpenTypeless is a free, open-source AI voice input tool for desktop. It lets users press a hotkey, speak naturally, and turn speech into polished text in any app, with support for AI rewriting, translation, custom dictionaries, BYOK providers, and 99 languages.
    Downloads: 2 This Week
    Last Update:
    See Project
  • 8
    MisterHouse:   Home Automation with Perl
    MisterHouse is a Windows/Unix home automation program written in Perl. It can respond to voice commands, web browsers, time of day, serial port and X10 data, external files, etc and can speak via Text to Speech engines. Support is on https://sourceforge.net/p/misterhouse/mailman/misterhouse-users/ and code is maintained on https://github.com/hollie/misterhouse
    Downloads: 4 This Week
    Last Update:
    See Project
  • 9
    SpeakoFlow

    SpeakoFlow

    Free voice-to-text dictation and AI assistant

    SpeakoFlow is free, open-source voice-to-text and AI dictation software for Windows, macOS, and Linux. Press a hotkey and speak to type in any app, including email, editors, chat, browsers, and terminals. Speech recognition runs on your device with Whisper or Parakeet, keeping voice transcription private. It also includes live dictation, “Hey Flow” AI writing, text cleanup, offline translation, floating voice assistant, screen vision, text-to-speech, personal memory, and profiles. ...
    Downloads: 2 This Week
    Last Update:
    See Project
  • Demo Series - Small Business Backup By Veeam Icon
    Demo Series - Small Business Backup By Veeam

    Learn how to protect your Microsoft 365 data, with simple, actionable tips today.

    Watch this on-demand demo series and learn how to protect your Microsoft 365 data with clear, simple, actionable steps that are easy to implement for businesses of all sizes.
    Watch Demo Series
  • 10

    Mumble PushToTalk Button

    Simple GUI app for LINUX with just one clickable "PushToTalk" button

    For people that use the well-known "mumble" VOIP application on desktop LINUX environments: When your desktop LINUX upgrades from "xorg" to "Wayland", mumble will no longer be able to use any of its keyboard shortcuts (due to security restrictions of Wayland). This little application solves that problem with a single, clickable button, displayed in a small, standard-style desktop LINUX application frame. The button is labeled "PushToTalk". Left-click and hold on it to talk, release to stop...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 11
    translate-gui

    translate-gui

    GUI for translate-shell, the cli tool for quick translation

    GUI for the translate-shell, aims to be easy to use translator and a helpful tool for learning new languages. Most tools do a one way translation from source to target language, do to the reverse involves choosing the source and target languages again. This tool can do a 2 way translation accompanied by speech output of the target language text. Hence it can prove to be an indispensable aid when learning new languages
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    Conversations

    Conversations

    App in java for chatting to a generative A.I. (involving tts and stt)

    Java application for chatting to generative AI Llama3. * The user can speak into the microphone (speechToText), edit the recognized text and send it to the AI. * The AI ​​responds and the server returns that response in real time, and the sentences converted to audio (textToSpeech), and the application broadcasts them through the speaker. The application is prepared so that only one user occupies the server's resources, so if the server is busy, in theory it will not let you connect. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13

    PyNotes

    An advanced Emacs-like text editor and IDE made in Python.

    PyNotes is an advanced Emacs-like text editor and IDE made in Python, but much simpler for new users.
    Downloads: 3 This Week
    Last Update:
    See Project
  • 14
    Pearl Desktop (PDE) 12

    Pearl Desktop (PDE) 12

    The Stable Solid Multimedia Workhorse Powerful OS with Eye Candy

    Pearl Linux Desktop (PDE) 12 is based on Ubuntu 24.04 LTR. This is your go to work horse daily driver for the advanced as well as the new Linux user. We say YES to APT, Flatpak and Appimages but NO to Snaps. Featuring Firefox-ESR instead of Firefox, Pulseasudio by default however Install package pearl-pipewire-config from our REPO to have pipewire as your default sound server. Very Smooth and Easy Configs. Compiz is the default Window Manager and you may switch window managers without...
    Downloads: 2 This Week
    Last Update:
    See Project
  • 15
    Crow Translate

    Crow Translate

    Lightweight translator that allows you to translate and speak text

    Crow Translate is a simple and lightweight translator written in C++ / Qt that allows you to translate and speak text using Google, Yandex, Bing, LibreTranslate and Lingva translate API. You may also be interested in my library QOnlineTranslator used in this project. Wayland does not support global shortcuts registration, but you can use D-Bus to bind actions in the system settings. For desktop environments that support additional applications actions (KDE, for example) you will see them predefined in the system shortcut settings. ...
    Downloads: 51 This Week
    Last Update:
    See Project
  • 16
    Amica

    Amica

    Amica is an open source interface for interactive communication

    Amica is an open source interface for interacting with fully animated 3D characters that combine voice chat, vision, and an emotion engine into a single experience. It lets you hold natural conversations with AI characters that can see, listen, and speak, while expressing emotional states through facial expressions and body language. Users can import VRM character models, adjust their appearance, tune the voice to match the character, and define behavior using different large language models and TTS backends. Under the hood, Amica leverages modern web and desktop technologies: three.js and three-vrm for 3D rendering, Transformers.js for running models in the browser, Whisper and Silero VAD for speech recognition and voice-activity detection, and a variety of LLM backends such as llama.cpp servers, ChatGPT-compatible APIs, Ollama, KoboldCpp, and others. ...
    Downloads: 12 This Week
    Last Update:
    See Project
  • 17

    texttalk

    Talk through typing the text

    Speak with generated voice of text input. Using Google translate web service, the audio sound of the spoken text can be extracted. Ideal for helping remote support, voice proxy for privacy, etc.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 18
    EmotiVoice

    EmotiVoice

    Multi-Voice and Prompt-Controlled TTS Engine

    ...It supports both English and Chinese and ships with over 2,000 preset voices, making it suitable for everything from characters and virtual anchors to narration and dialogue. The core idea is prompt-based emotional and style control: you can ask the engine to speak “happy,” “sad,” “excited,” or with other high-level style prompts that shape prosody, pitch, speed, and energy. EmotiVoice provides multiple ways to interact with it, including a web interface, a Docker image, an HTTP API (including an OpenAI-compatible TTS API), and Python scripts for batch synthesis. It also supports voice cloning with your own data, backed by recipes for popular datasets like DataBaker and LJSpeech, so you can train or adapt voices to custom personas.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 19
    VALL-E X

    VALL-E X

    Open source implementation of Microsoft's VALL-E X zero-shot TTS model

    ...The model attempts to match not just timbre, but also tone, pitch, emotion, and prosody of the reference audio, resulting in highly personalized output. VALL-E-X supports zero-shot cross-lingual synthesis, meaning a monolingual speaker’s voice can be used to speak other languages without additional training. It also preserves aspects of the acoustic environment, such as background noise or reverb, making the generated audio feel more like it came from the same setting as the prompt. The repository includes Python APIs, sample scripts, ready-to-use voice presets, and demos hosted on Hugging Face Spaces and Google Colab so users can try it.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 20
    HTML To Markdown for PHP

    HTML To Markdown for PHP

    Convert HTML to Markdown with PHP

    ...You want to store new content in HTML format but edit it as Markdown. You want to convert HTML email to plain text email. You know a guy who's been converting HTML to Markdown for years, and now he can speak Elvish. You'd quite like to be able to speak Elvish. You just really like Markdown. By default, HTML To Markdown preserves HTML tags without Markdown equivalents, like span and div. To strip HTML tags that don't have a Markdown equivalent while preserving the content inside them, set strip_tags to true.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21
    ChatGPT-to-API

    ChatGPT-to-API

    Scalable unofficial ChatGPT API for production

    ...This makes it possible to plug ChatGPT into automated systems, serverless functions, or backend services that expect REST or JSON RPC interfaces without needing to modify each consumer to speak a browser protocol. Developers can deploy ChatGPT-to-API on a server, send it requests with structured parameters (like messages, metadata, or configuration flags), and receive structured replies in return.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 22
    vocdoni-node

    vocdoni-node

    A set of libraries and tools for the Vocdoni decentralized backend

    ...Vocdoni is a universally verifiable, censorship-resistant, and anonymous self-sovereign governance system, designed with the scalability and ease-of-use to support either small/private and big/national elections. Our main aim is a trustless voting system, where anyone can speak their voice and where everything can be audited. We are engineering building blocks for a permissionless, private and censorship-resistant democracy. We intend the algorithms, systems, and software that we build to be a useful contribution toward making violence in these crypto networks impossible by protecting users privacy with cryptography. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23
    alfred-ai

    alfred-ai

    The development of my ai assistant, Alfred

    The development of my ai assistant, Alfred.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 24

    Flash Card Tutor

    Randomly display language quiz lines from a text file

    ...The file is organized with Quiz lines on an even numbered line, and the Answer lines on odd-numbered lines, as shown in this example for Thai language training: Hello [-- on line #0 --) Saw wat dee kup [-- one line #1--] Excuse me [-- on line #2 --] Kho thoet kup [-- on line #3 --] I speak Thai a little [-- on line #4 --] Phom phut paasaa Thai dai nitnoi kup [-- on line #5 --] (Another even-numbered quiz line) (Another odd-numbered answer line, etc.)
    Downloads: 0 This Week
    Last Update:
    See Project
  • 25

    WikiSearch

    With this program, you can search or browse any Wikipedia article.

    This simple program enables you to search Wikipedia, the free encyclopedia, without having to open your browser.
    Downloads: 0 This Week
    Last Update:
    See Project