Port of OpenAI's Whisper model in C/C++
State-of-the-art TTS model under 25MB
A GPT-4o Level MLLM for Vision, Speech and Multimodal Live Streaming
AI that sees your screen and listens to conversations
Transcribe and translate audio offline on your personal computer
iOS application for Lumo
Reading book source
On-device Speech-to-Intent engine powered by deep learning
Low-latency AI inference engine optimized for mobile devices
Offline speech recognition API for Android, iOS, Raspberry Pi
In-App assistant SDK to build a multimodal conversational UX for iOS
Easy AI Softwares for Blind, Deaf, Handicapped, Disabled People
Free & Easy AI Voice Accounting Software For Blind & Speechless People
React Native Voice Recognition library for iOS and Android
SDK to build a multimodal conversational UX for Flutter apps
Build a multimodal conversational UX for apps created with React
Assistant SDK to build a multimodal conversational UX for Apache
ColdFusion SDK for the VoiceShot API.
PHP SDK for processing phone calls and SMS through the VoiceShot API.
.NET SDK for processing phone calls and SMS through the VoiceShot API.
ASP SDK for processing phone calls and SMS through the VoiceShot API.
Easy way to create conversation chats
Open source speech models for Julius in English and other languages.