Stable Diffusion built-in to Blender
RGBD video generation model conditioned on camera input
Official SeedVR2 Video Upscaler for ComfyUI
Diffusion Transformer with Fine-Grained Chinese Understanding
Tokenizer-Free TTS for Multilingual Speech Generation
HY-Motion model for 3D character animation generation
Open-source multi-speaker long-form text-to-speech model
A unified library of SOTA model optimization techniques
Expressive Portrait Image Animation for Live Streaming
Repo for SeedVR2 & SeedVR
InvokeAI is a leading creative engine for Stable Diffusion models
Stable Diffusion web UI
Cosmos-RL is a flexible and scalable Reinforcement Learning framework
Multimodal Diffusion with Representation Alignment
Personalize Any Characters with a Scalable Diffusion Transformer
A SOTA open-source image editing model
Official Python inference and LoRA trainer package
Official inference repo for FLUX.1 models
Project Lyra: Open Generative 3D World Models
ComfyUI wrapper nodes for WanVideo and related models
Code and models for ICML 2024 paper, NExT-GPT
Inference script for Oasis 500M
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Deep learning framework
Official PyTorch Implementation