Showing 150 open source projects for "speaker"

View related business solutions
  • Veeam Data Platform v13.1 - Get Your Free Trial Icon
    Veeam Data Platform v13.1 - Get Your Free Trial

    Secure by design, portable by default. Recover clean, fast, anywhere. Start a free trial.

    Try Veeam Data Platform today. Experience the unified platform that's secure by design, portable by default, and proven to recover clean, fast, and anywhere.
    Try it Free
  • MongoDB Atlas runs apps anywhere Icon
    MongoDB Atlas runs apps anywhere

    Deploy in 115+ regions with the modern database for every enterprise.

    MongoDB Atlas gives you the freedom to build and run modern applications anywhere—across AWS, Azure, and Google Cloud. With global availability in over 115 regions, Atlas lets you deploy close to your users, meet compliance needs, and scale with confidence across any geography.
    Start Free
  • 1
    PseudonymizeSpeech

    PseudonymizeSpeech

    Praat script to pseudonymize speech.

    A Praat script to pseudonymize speech. That is, Pseudonymize Speech tries to make it difficult to recognize a speaker while still retaining relevant (para-)linguistic features and intelligibility. There is a trade-off between the level of pseudonymization and the (para-)linguistic features retained. The approach is to manipulate the spectro-temporal structure of the speech to simulate a different length and structure of the vocal tract, as well as a different pitch and speaking rate. ...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 2
    Multilingual Speech Synthesis

    Multilingual Speech Synthesis

    An implementation of Tacotron 2 that supports multilingual experiments

    ...It presents a model combining ideas from Learning to speak fluently in a foreign language: Multilingual speech synthesis and cross-language voice cloning, End-to-End Code-Switched TTS with Mix of Monolingual Recordings, and Contextual Parameter Generation for Universal Neural Machine Translation. We provide data for comparison of three multilingual text-to-speech models. The first shares the whole encoder and uses an adversarial classifier to remove speaker-dependent information from the encoder. The second has separate encoders for each language.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 3
    GROWbox Supervisor System (GROWSS)

    GROWbox Supervisor System (GROWSS)

    Automated Plant Environment Growing System using Raspberry Pi

    ...Hi & low values are also saved. The LEDs on the case & the mobile application indicate if there is a high/low temp alarm, hi/low humidity alarm, soil moisture alarm, or smoke alarm. A speaker (buzzer) is activated on the case if there is a smoke alarm. 2 other LEDs indicate if either the exhaust fan is on or if the humidifier is on.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 4
    remark

    remark

    A simple, in-browser, markdown-driven slideshow tool

    A simple, in-browser, Markdown-driven slideshow tool targeted at people who know their way around HTML and CSS. If your ideal slideshow creation workflow contains any of the following steps, just write what's on your mind, do some basic styling, easily collaborate with others, share with and show to everyone, then remark might be perfect for your next slideshow! Focus on the content, expressing yourself in next to plain text not worrying what flashy graphics and disturbing effects to put...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Paessler: Easy to Use With Enterprise Power. Free Trial Icon
    Paessler: Easy to Use With Enterprise Power. Free Trial

    A low-code dashboard makes monitoring intuitive for any admin, while scripting and custom sensors give experts full control.

    You shouldn't have to choose between a monitoring tool that's easy to use and one that's powerful enough for a complex environment. PRTG's low-code interface lets any admin build dashboards, set alerts and monitor devices without scripting, while custom sensors and full API access are there when your team needs deeper control. One platform, no compromise. Download a free 30-day trial now.
    Get Free Download
  • 5
    Resemblyzer

    Resemblyzer

    A python package to analyze and compare voices with deep learning

    ...Its main value is making speaker representation accessible through a simple Python workflow.
    Downloads: 1 This Week
    Last Update:
    See Project
  • 6

    STEAM game

    Arduino project, game STEAM test with record

    ...But it can be connected to a microusb charger. A motion sensor will allow the system to wake up. It will have touch sensors and LCD screens. It will also have a led array and a speaker to animate the game a bit.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 7

    ESP32 Arduino Reflow Oven Controller

    ESP32 Arduino Based Reflow Oven Controller Schematics and Firmware

    This reflow oven controller was built to control a modified toaster oven for the purpose of doing reflow soldering of printed circuit boards. The controller was designed and built around an ESP32-DevKitC development board with ancillary electronics added to complete the controller. The schematic can be found in <files/hardware>. The code for the controller was modified extensively from existing Arduino sketches for reflow oven controllers and is PID-algorithm based. The firmware...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 8
    CMU Sphinx

    CMU Sphinx

    Speech Recognition Toolkit

    ...----> Maintenance and improvement work has MOVED to https://cmusphinx.github.io/ Please go there for the most recent software and documentation. <---- CMUSphinx is a speaker-independent large vocabulary continuous speech recognizer released under BSD style license. It is also a collection of open source tools and resources that allows researchers and developers to build speech recognition systems.
    Leader badge
    Downloads: 423 This Week
    Last Update:
    See Project
  • 9
     rims-arduino-library

    rims-arduino-library

    Recirculation infusion mash system library for Arduino

    This library implement RIMS controls for home brewers. For definition of a RIMS, see https://tinyurl.com/j3lyuyc For me, an Arduino micro controller + a LCD Keypad shield was cheaper and a lot more customizable than a commercial PID controller. So, with this library, a commercial PID controller is unnecessary. Automatic PID tuning toolkit is also included. Temperature can be read with a thermistor, a resistance temperature detector (RTD) or any custom temperature probe. Heater is...
    Downloads: 0 This Week
    Last Update:
    See Project
  • Host LLMs in Production With On-Demand GPUs Icon
    Host LLMs in Production With On-Demand GPUs

    NVIDIA L4 GPUs. 5-second cold starts. Scale to zero when idle.

    Deploy your model, get an endpoint, pay only for compute time. No GPU provisioning or infrastructure management required.
    Start Free
  • 10
    SoundManager

    SoundManager

    Manage multiple sound channels

    With SoundManager you are able to control multiple sound channels of one or multiple sound cards. It uses ALSA to speak to the sound cards. An example how to use SoundManager would be a shop or store with multiple rooms where every should have its own speaker. Then with SoundManager you would have a single web interface where you could control every single speaker. SoundManager also allows to create different playlists and assign them to one or multiple sound channels. It is also possible to use line-in or a microphone instead of a playlist.
    Downloads: 5 This Week
    Last Update:
    See Project
  • 11
    Deepvoice3_pytorch

    Deepvoice3_pytorch

    PyTorch implementation of convolutional neural networks

    An open source implementation of Deep Voice 3: Scaling Text-to-Speech with Convolutional Sequence Learning.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 12
    English Grammar Tips for Russian Speaker

    English Grammar Tips for Russian Speaker

    English Grammar Tips for Russian-Speakers

    Не претендуя на полный охват, представляем вам сборник кратких подсказок по употреблению грамматики английского языка.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 13
    DC-TTS

    DC-TTS

    TensorFlow Implementation of DC-TTS: yet another text-to-speech model

    ...Training scripts, data loaders, and hyperparameter configurations are provided to reproduce results on several datasets, including LJ Speech for English, a Korean single-speaker dataset, and audiobook data from Nick Offerman and Kate Winslet.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 14
    Remote Viewing Assistant

    Remote Viewing Assistant

    Provides various utilities to aid in the practice of remote viewing.

    ...The utilities include: • A random target generator for generating and pairing a random target ID and target. • An ideogram trainer to train the nervous system to respond to certain ideas in a standardized manner. • A target identifier speaker to speak aloud the identifier of a target. Requires .NET Framework 4.7.1
    Downloads: 0 This Week
    Last Update:
    See Project
  • 15
    Oasi -  Open Document Speaker

    Oasi - Open Document Speaker

    A simple Text2Audio

    Document Speaker - A simple Editor to give VOICE on Your Documents, save your doc as AudioBook or other format this app recognizes the language of the documents and converts them into audiobooks by recognizing texts in nearly 200 languages ... Open RTF & RTFD (mac format/inode directory) ODT,EPUB (unstable), PDF as plain Text to convert as MP4 or AudioBook.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 16
    Lip Reading

    Lip Reading

    Cross Audio-Visual Recognition using 3D Architectures

    ...Audio-visual recognition (AVR) has been considered as a solution for speech recognition tasks when the audio is corrupted, as well as a visual recognition method used for speaker verification in multi-speaker scenarios. The approach of AVR systems is to leverage the extracted information from one modality to improve the recognition ability of the other modality by complementing the missing information. The essential problem is to find the correspondence between the audio and visual streams, which is the goal of this work. ...
    Downloads: 1 This Week
    Last Update:
    See Project
  • 17

    Distant Speech Recognition

    Beamforming and Speech Recognition Toolkit

    BTK contains C++ and Python libraries that implement speech processing and microphone array techniques such as speech feature extraction, speech enhancement, speaker tracking, beamforming, dereverberation and echo cancellation algorithms. The Millennium ASR provides C++ and python libraries for automatic speech recognition. The Millennium ASR implements a weighted finite state transducer (WFST) decoder, training and adaptation methods. These toolkits are meant for facilitating research and development of automatic distant speech recognition.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 18

    Vietnamese Grapheme to Phoneme

    Converting any Vietnamese word in grapheme to phoneme

    Vietnamese is a language that any Vietnamese word can be correctly pronounced even if the speaker does not know its meaning and has never seen it. This tool , a grapheme-to-phoneme method, converts any Vietnamese word from grapheme-based into a phoneme-based pronunciation that integrates tone information. It is usefull to create a lexicon for deverloping a Vi LVCSR system. Detail in: http://ieeexplore.ieee.org/xpl/login.jsp?
    Downloads: 0 This Week
    Last Update:
    See Project
  • 19
    ICE Nigeria

    ICE Nigeria

    Nigerian component of the International Corpus of English

    ...For the spoken part the eaf files (ELAN files in xml format) together with the text files can be downloaded separately from the sound files. In addition, we provide the corpus manual as well as metadata (speaker age, gender, ethnic group and profession) and XML specifications.
    Downloads: 30 This Week
    Last Update:
    See Project
  • 20
    Kamus Plus

    Kamus Plus

    Kamus Interaktif Bahasa Inggris - Indonesia

    Kamus Plus merupakan kamus bahasa Inggris – Indonesia yang interaktif, dapat mencari arti kata dari bahasa Inggris ke Indonesia maupun sebaliknya. Dapat berjalan di GNU/Linux dan Windows.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 21

    Speaker Recognition System

    Speaker Recognition System - Matlab source code

    Speaker identity is correlated with the physiological and behavioral characteristics of the speaker. These characteristics exist both in the spectral envelope (vocal tract characteristics) and in the supra-segmental features (voice source characteristics and dynamic features spanning several segments). Index Terms: speaker, recognition, verification, sound, words.
    Downloads: 0 This Week
    Last Update:
    See Project
  • 22

    HangouIRC

    Application for integration of Google Hangout and IRC Chat

    This application is a fork of chatSlide and integrates IRC chat with Google Hangout, giving the possibility of the speaker showing slides of his lecture
    Downloads: 0 This Week
    Last Update:
    See Project
  • 23

    MUNLawS

    Debate manager

    MUNLawS Software README File Welcome to open source debate manager for Model united nations conference. We are offering you an official release of the application. Date of first release: 25.10.2013 Date of second release: 20.9.2014 (minor bug fixes and different ICJ countries: Marshall Islands and United Kingdom) Java is required to run this application. If your operating system recognises .jar archive as winRar archive, you need to install Java. Download URL:...
    Downloads: 0 This Week
    Last Update:
    See Project
  • 24

    kirikiri-kag

    The English-Translated Kirikiri2/KAG3 documentation...

    ...It is still being translated and hope to finish it soon...! Hope there are not many typos, and if there are, then report them to me here.... (UPDATE: Need the help of a Japanese speaker to translate some complex files... E-Mail me at: shivanshs9@gmail.com, if you are interested.) (ANOTHER UPDATE: Just uploaded the English user interface version of Kirkiri2 2.32.2. Just download the zip file, extract it to your Kirikiri2 location and overwrite(update) the old japanese files... Hope you like my gift!)
    Downloads: 18 This Week
    Last Update:
    See Project
  • 25
    Cotovía

    Cotovía

    Text-to-Speech System for Galician and Spanish

    Cotovía is a unit-selection text-to-speech system for Galician and Spanish. Cotovía is distributed under the GPL3.0+ license, but each of the avaliable speaker voices has its own license. The speakers available at sourceforge are free for commercial and non-commercial uses. Another speaker, free for non-commercial uses, is avaliable through external links (see the Blog section). Cotovia has been developed by the University de Vigo and the center 'Ramón Piñeiro' for Research in Humanities, both in Galicia, Spain. ...
    Downloads: 40 This Week
    Last Update:
    See Project