One-click local MCP server installation in desktop apps
GLM-4 series: Open Multilingual Multimodal Chat LMs
Unified Multimodal Understanding and Generation Models
Sharp Monocular Metric Depth in Less Than a Second
DeepSeek Coder: Let the Code Write Itself
A series of math-specific large language models of our Qwen2 series
Audio Language Models are Few-Shot Learners
Open image model at the forefront of design
Video Object and Interaction Deletion
Foundation model for image generation
Block Diffusion for Ultra-Fast Speculative Decoding
Video understanding codebase from FAIR for reproducing video models
Stable Virtual Camera: Generative View Synthesis with Diffusion Models
The Clay Foundation Model - An open source AI model and interface
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Bidirectional token-classification model for identifiable info
Genome modeling and design across all domains of life
Project Lyra: Open Generative 3D World Models
Ultra-Efficient LLMs on End Device
Fast and Universal 3D reconstruction model for versatile tasks
A Production-ready Reinforcement Learning AI Agent Library
Official implementation of DreamCraft3D
Open-source large language model family from Tencent Hunyuan
Use ChatGPT to summarize the arXiv papers
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning