Crafting engine for artists, designers, and filmmakers
Wan2.1: Open and Advanced Large-Scale Video Generative Model
Director, Screenwriter, Producer, and Video Generator All-in-One
Wan2.2: Open and Advanced Large-Scale Video Generative Model
A Customizable Image-to-Video Model based on HunyuanVideo
Text and image to video generation: CogVideoX and CogVideo
Multimodal-Driven Architecture for Customized Video Generation
Generate high-definition story short videos with one click using AI
Foundational video generation model with 13.6B parameters
Official Python inference and LoRA trainer package
Tencent Hunyuan Multimodal diffusion transformer (MM-DiT) model
Motion-controllable Video Generation via Latent Trajectory Guidance
Open-Sora: Democratizing Efficient Video Production for All
HunyuanVideo: A Systematic Framework For Large Video Generation Model
Pushing the Frontier of Long Audio-Visual Generation
RGBD video generation model conditioned on camera input
Visual AI Workflow Builder
Implementation of Phenaki Video, which uses Mask GIT
A Customizable Image-to-Video Model based on HunyuanVideo
Overcoming Data Limitations for High-Quality Video Diffusion Models
Open-source AI video pipeline, fully automated with MCP
Implementation of Make-A-Video, new SOTA text to video generator
AI-based tool for removing hardsubs and text-like watermarks
CLIP + FFT/DWT/RGB = text to image/video
Multimodal AI Story Teller, built with Stable Diffusion, GPT, etc.