Foundational video generation model with 13.6B parameters
Give Claude the ability to watch any video
State-of-the-art (SoTA) text-to-video pre-trained model
Give Claude the ability to watch and understand videos
A GUI tool for extracting hard-coded subtitle (hardsub) from videos
AI tool that removes hardcoded subtitles and text from videos locally
Let Claude (or any LLM) actually watch a video
Lets make video diffusion practical
Taming Stable Diffusion for Lip Sync
Video understanding codebase from FAIR for reproducing video models
Real time face swap and one-click video deepfake
Video-based AI memory library. Store millions of text chunks in MP4
AI-assisted storyboard and video generation tool
MiniMax H3 is a general-purpose, omni-modal generative system
Write HTML. Render video. Built for agents
Official Python inference and LoRA trainer package
Make videos programmatically with React
VGGSfM: Visual Geometry Grounded Deep Structure From Motion
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
MiniMax H3 inference engine for Mac computers
Interactive video and image annotation tool for computer vision
Repo for SeedVR2 & SeedVR
Multimodal-Driven Architecture for Customized Video Generation
Build Vision Agents quickly with any model or video provider
Video Object and Interaction Deletion