Best practice TTS based on BERT and VITS
Unofficial Parallel WaveGAN
Implementation of DALL-E 2, OpenAI's updated text-to-image synthesis
Label, clean and enrich text datasets with LLMs
Alfred workflow using ChatGPT, DALL·E 2 and other models for chatting
The first Chinese LLaMA2 model in the open source community
Text-to-Image generation. The repo for NeurIPS 2021 paper
SoftVC VITS Singing Voice Conversion
Clarity in the current fast-paced mess of Open Source innovation
Implementation of MusicLM music generation model in Pytorch
Framework for Accelerating LLM Generation with Multiple Decoding Heads
Run the Stable Diffusion releases in a Docker container
Application that simplifies the installation of AI-related projects
Let us control diffusion models
Unified embedding model
Basaran, an open-source alternative to the OpenAI text completion API
Inference code for Llama models
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
Convert an image to text to spot intelligible words.
An open-source framework for training large multimodal models
Resources, corpora, and tools for Chinese natural language processing
Implementation of Nougat Neural Optical Understanding
Explore large language models in 512MB of RAM
Python package for easily interfacing with chat apps