The official repo of Qwen chat & pretrained large language model
Visual Causal Flow
GLM-4 series: Open Multilingual Multimodal Chat LMs
Phi-3.5 for Mac: Locally-run Vision and Language Models
Block Diffusion for Ultra-Fast Speculative Decoding
A Powerful Native Multimodal Model for Image Generation
Generating Immersive, Explorable, and Interactive 3D Worlds
SOTA on-device LLMs, small yet powerful
LTX-Video Support for ComfyUI
Python SDK for Claude Agent
Inference code for scalable emulation of protein equilibrium ensembles
Pretrained time-series foundation model developed by Google Research
HY-Motion model for 3D character animation generation
A Multi-Modal World Model for Reconstructing, Generating, Simulation
Qwen-Image-Layered: Layered Decomposition for Inherent Editablity
DeepMind model for tracking arbitrary points across videos & robotics
Global weather forecasting model using graph neural networks and JAX
code for Mesh R-CNN, ICCV 2019
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Implementation of the Surya Foundation Model for Heliophysics
The Clay Foundation Model - An open source AI model and interface
A SOTA open-source image editing model
Robust Speech Recognition Across Languages, Dialects
Open-source image generative foundation model
A theoretical reconstruction of the Claude Mythos architecture