Convert Google Gemini web into OpenAI-compatible API
Fast stable diffusion on CPU and AI PC
Fast-stable-diffusion + DreamBooth
Text and image to video generation: CogVideoX and CogVideo
Tongyi Deep Research, the Leading Open-source Deep Research Agent
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Advanced language and coding AI model
Diffusion Transformer with Fine-Grained Chinese Understanding
High-Resolution Image Synthesis with Latent Diffusion Models
Pokee Deep Research Model Open Source Repo
Repo of Qwen2-Audio chat & pretrained large audio language model
Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence
Chinese and English multimodal conversational language model
Qwen2.5-VL is the multimodal large language model series
Generating Immersive, Explorable, and Interactive 3D Worlds
Qwen3-omni is a natively end-to-end, omni-modal LLM
Language modeling in a sentence representation space
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
Capable of understanding text, audio, vision, video
ChatGLM-6B: An Open Bilingual Dialogue Language Model
Easy Docker setup for Stable Diffusion with user-friendly UI
A state-of-the-art open visual language model
ChatGPT interface with better UI
StudioOllamaUI is a local, portable interface for Ollama
GUI shell for running local LLM on desktop