Python inference and LoRA trainer package for the LTX-2 audio–video
Agentic, Reasoning, and Coding (ARC) foundation models
High-Resolution 3D Assets Generation with Large Scale Diffusion Models
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Fast stable diffusion on CPU and AI PC
Uncommon Objects in 3D dataset
Open-source deep-learning framework
Visual Causal Flow
Qwen3-TTS is an open-source series of TTS models
A Family of Open Sourced Music Foundation Models
Reference PyTorch implementation and models for DINOv3
Programmatic access to the AlphaGenome model
Text and image to video generation: CogVideoX and CogVideo
Advanced language and coding AI model
Foundation Models for Time Series
Hackable and optimized Transformers building blocks
AlphaFold 3 inference pipeline
Sharp Monocular Metric Depth in Less Than a Second
Recovering the Visual Space from Any Views
Code for running inference with the SAM 3D Body Model 3DB
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
Qwen3-Coder is the code version of Qwen3
A Powerful Native Multimodal Model for Image Generation
Qwen3 is the large language model series developed by Qwen team
The official repo of Qwen chat & pretrained large language model