Foundational video generation model with 13.6B parameters
Give Claude the ability to watch any video
Give Claude the ability to watch and understand videos
Let Claude (or any LLM) actually watch a video
Lets make video diffusion practical
Taming Stable Diffusion for Lip Sync
Video understanding codebase from FAIR for reproducing video models
Real time face swap and one-click video deepfake
Video-based AI memory library. Store millions of text chunks in MP4
AI-assisted storyboard and video generation tool
MiniMax H3 is a general-purpose, omni-modal generative system
Write HTML. Render video. Built for agents
Official Python inference and LoRA trainer package
Make videos programmatically with React
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
MiniMax H3 inference engine for Mac computers
Interactive video and image annotation tool for computer vision
Multimodal-Driven Architecture for Customized Video Generation
Build Vision Agents quickly with any model or video provider
Video Object and Interaction Deletion
RGBD video generation model conditioned on camera input
Inference script for Oasis 500M
Pluggable SOTA multi-object tracking modules for segmentation
Project Lyra: Open Generative 3D World Models
CLI tool for removing watermarks from AI-generated videos using frame-