Miso TTS is an 8 billion, highly emotive text-to-speech model
Large-language-model & vision-language-model based on Linear Attention
Capable of understanding text, audio, vision, video
Memory-efficient and performant finetuning of Mistral's models
CodeGeeX: An Open Multilingual Code Generation Model (KDD 2023)
ChatGLM-6B: An Open Bilingual Dialogue Language Model
CodeGeeX2: A More Powerful Multilingual Code Generation Model
High-Resolution Image Synthesis with Latent Diffusion Models
Easy Docker setup for Stable Diffusion with user-friendly UI
AI-powered tool to quickly remove watermarks from images flawlessly
StudioOllamaUI is a local, portable interface for Ollama
Open-source, high-performance Mixture-of-Experts large language model
AI Suite for upscaling, interpolating & restoring images/videos
ChatGPT interface with better UI
Qwen2.5-Coder is the code version of Qwen2.5, the large language model
Di♪♪Rhythm: Blazingly Fast & Simple End-to-End Song Generation
GUI shell for running local LLM on desktop
Powerful open source image generation model
A Conversational Speech Generation Model
A state-of-the-art open visual language model
Open Multilingual Multimodal Chat LMs
Stable Diffusion with Core ML on Apple Silicon
Towards Real-World Vision-Language Understanding
The ChatGPT Retrieval Plugin lets you easily find personal documents
Pushing the Limits of Mathematical Reasoning in Open Language Models