LTX-Video Support for ComfyUI
gpt-oss-120b and gpt-oss-20b are two open-weight language models
Contexts Optical Compression
Text and image to video generation: CogVideoX and CogVideo
Visual Causal Flow
Programmatic access to the AlphaGenome model
Reference PyTorch implementation and models for DINOv3
Qwen3-TTS is an open-source series of TTS models
Advanced language and coding AI model
A Family of Open Sourced Music Foundation Models
Hackable and optimized Transformers building blocks
Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference
Recovering the Visual Space from Any Views
Sharp Monocular Metric Depth in Less Than a Second
Qwen3 is the large language model series developed by Qwen team
Code for running inference with the SAM 3D Body Model 3DB
A theoretical reconstruction of the Claude Mythos architecture
Open-source multi-speaker long-form text-to-speech model
Industrial-level controllable zero-shot text-to-speech system
An experimental version of DeepSeek model
A Powerful Native Multimodal Model for Image Generation
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Open-source large language model family from Tencent Hunyuan
Unified Multimodal Understanding and Generation Models
Qwen3-Coder is the code version of Qwen3