Foundational video generation model with 13.6B parameters
Give Claude the ability to watch any video
State-of-the-art (SoTA) text-to-video pre-trained model
Give Claude the ability to watch and understand videos
Let Claude (or any LLM) actually watch a video
Lets make video diffusion practical
Taming Stable Diffusion for Lip Sync
Video understanding codebase from FAIR for reproducing video models
Real time face swap and one-click video deepfake
Video-based AI memory library. Store millions of text chunks in MP4
AI-assisted storyboard and video generation tool
MiniMax H3 is a general-purpose, omni-modal generative system
Write HTML. Render video. Built for agents
Official Python inference and LoRA trainer package
Make videos programmatically with React
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
MiniMax H3 inference engine for Mac computers
Interactive video and image annotation tool for computer vision
Repo for SeedVR2 & SeedVR
Multimodal-Driven Architecture for Customized Video Generation
Build Vision Agents quickly with any model or video provider
Video Object and Interaction Deletion
RGBD video generation model conditioned on camera input
Inference script for Oasis 500M