Tongyi Deep Research, the Leading Open-source Deep Research Agent
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Qwen3-omni is a natively end-to-end, omni-modal LLM
The Clay Foundation Model - An open source AI model and interface
scikit-learn compatible tabular foundation model
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Audio Language Models are Few-Shot Learners
Open image model at the forefront of design
Video Object and Interaction Deletion
Qwen3-ASR is an open-source series of ASR models
Foundation model for image generation
Video understanding codebase from FAIR for reproducing video models
Tool for exploring and debugging transformer model behaviors
Bidirectional token-classification model for identifiable info
Project Lyra: Open Generative 3D World Models
Inference script for Oasis 500M
Generate Any 3D Scene in Seconds
Fast and Universal 3D reconstruction model for versatile tasks
A collection of high-quality models for the MuJoCo physics engine
A Production-ready Reinforcement Learning AI Agent Library
PyTorch code and models for the DINOv2 self-supervised learning
Official implementation of DreamCraft3D
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
High-Fidelity and Controllable Generation of Textured 3D Assets
State-of-the-art (SoTA) text-to-video pre-trained model