A collection of high-quality models for the MuJoCo physics engine
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
Phi-3.5 for Mac: Locally-run Vision and Language Models
Open-source image generative foundation model
Qwen3 is the large language model series developed by Qwen team
The Clay Foundation Model - An open source AI model and interface
Recovering the Visual Space from Any Views
Open-source multi-speaker long-form text-to-speech model
FAIR Sequence Modeling Toolkit 2
Qwen-Image is a powerful image generation foundation model
Renderer for the harmony response format to be used with gpt-oss
Qwen2.5-VL is the multimodal large language model series
High-Resolution Image Synthesis with Latent Diffusion Models
SOTA on-device LLMs, small yet powerful
Block Diffusion for Ultra-Fast Speculative Decoding
LTX-Video Support for ComfyUI
CogView4, CogView3-Plus and CogView3(ECCV 2024)
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Robust Speech Recognition Across Languages, Dialects
Qwen3-ASR is an open-source series of ASR models
Inference code for scalable emulation of protein equilibrium ensembles
Qwen3-Coder is the code version of Qwen3
The official repo of Qwen chat & pretrained large language model
Open-source deep-learning framework
Generate Any 3D Scene in Seconds