Open image model at the forefront of design
A Powerful Native Multimodal Model for Image Generation
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Open-Source Financial Large Language Models
CogView4, CogView3-Plus and CogView3(ECCV 2024)
A trainable PyTorch reproduction of AlphaFold 3
An experimental version of DeepSeek model
HY-Motion model for 3D character animation generation
Personalize Any Characters with a Scalable Diffusion Transformer
A Unified Framework for Text-to-3D and Image-to-3D Generation
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
GLM-4.5: Open-source LLM for intelligent agents by Z.ai
Designed for text embedding and ranking tasks
A series of math-specific large language models of our Qwen2 series
Qwen2.5-VL is the multimodal large language model series
Long-form streaming TTS system for multi-speaker dialogue generation
Sharp Monocular Metric Depth in Less Than a Second
Research code artifacts for Code World Model (CWM)
Video Object and Interaction Deletion
RGBD video generation model conditioned on camera input
Chinese and English multimodal conversational language model
Repo of Qwen2-Audio chat & pretrained large audio language model
Renderer for the harmony response format to be used with gpt-oss
Pokee Deep Research Model Open Source Repo
Audio foundation model excelling in audio understanding