Traditional Mandarin LLMs for Taiwan
Plug-n-play module turning text-to-image models into animation
Open-source choice to scale, assess and maintain natural language data
High-Resolution Image Synthesis with Latent Diffusion Models
Unlimited, private and free Speech-To-Text program
An opinionated CLI to transcribe Audio files w/ Whisper on-device
Edge TTS Desktop turns text into speech through edge-tts.
SoundTranscriber can be used to generate automatic transcription / aut
ktrain is a Python library that makes deep learning AI more accessible
A Conversational Speech Generation Model
Bypass Ai content for GPTZero and others making text Undetectable
Offline desktop app to convert EPUB to MP3 using Kokoro-82M neural TTS
Synchronized Translation for Videos
Two Integrated Text To Speech Engines uses MMS & Silero
AI-powered tool to quickly remove watermarks from images flawlessly
Audiocraft is a library for audio processing and generation
A Python application to add watermarks (text or image) to PDF files
Chinese Llama-3 LLMs) developed from Meta Llama 3
A Pioneering Open-Source Alternative to GPT-4o
Implementation of Video Diffusion Models
mice stt tts
AI-powered semantic indexing: automating the creation of book indexes
Translate English to Bangla using CSV file format and range wise.
Implementation of Make-A-Video, new SOTA text to video generator