Label, clean and enrich text datasets with LLMs
SoftVC VITS Singing Voice Conversion
Alfred workflow using ChatGPT, DALL·E 2 and other models for chatting
Run the Stable Diffusion releases in a Docker container
Let us control diffusion models
Application that simplifies the installation of AI-related projects
Framework for Accelerating LLM Generation with Multiple Decoding Heads
Implementation of MusicLM music generation model in Pytorch
Basaran, an open-source alternative to the OpenAI text completion API
Inference code for Llama models
Convert an image to text to spot intelligible words.
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.
Resources, corpora, and tools for Chinese natural language processing
Unified embedding model
An open-source framework for training large multimodal models
Implementation of Nougat Neural Optical Understanding
Explore large language models in 512MB of RAM
OpenMMLab Text Detection, Recognition and Understanding Toolbox
A webui for different audio related Neural Networks
se GPT or other prompt based models to get structured output
Task-oriented finetuning for better embeddings on neural search
Chinese LLaMA & Alpaca large language model + local CPU/GPU training
Chinese voice dialogue robot/smart speaker project
Python package for easily interfacing with chat apps
Video automatic transcribe and translated subtitle generator