Generating Immersive, Explorable, and Interactive 3D Worlds
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
Community plugin marketplace for Claude Cowork and Claude Code
Visual Causal Flow
Tiny vision language model
Open-source image generative foundation model
Video Object and Interaction Deletion
Qwen3-ASR is an open-source series of ASR models
Foundation model for image generation
Fast-stable-diffusion + DreamBooth
Z80-μLM is a 2-bit quantized language model
4M: Massively Multimodal Masked Modeling
Hackable and optimized Transformers building blocks
Official implementation of DreamCraft3D
New family of code large language models (LLMs)
A Systematic Framework for Interactive World Modeling
FAIR Sequence Modeling Toolkit 2
Renderer for the harmony response format to be used with gpt-oss
Repo of Qwen2-Audio chat & pretrained large audio language model
Tongyi Deep Research, the Leading Open-source Deep Research Agent
Stable Diffusion WebUI Forge is a platform on top of Stable Diffusion
The official PyTorch implementation of Google's Gemma models
GLM-4.6V/4.5V/4.1V-Thinking, towards versatile multimodal reasoning
MedicalGPT: Training Your Own Medical GPT Model with ChatGPT Training
Pretrained time-series foundation model developed by Google Research