Foundational video generation model with 13.6B parameters
Give Claude the ability to watch and understand videos
Video understanding codebase from FAIR for reproducing video models
AI-assisted storyboard and video generation tool
MiniMax H3 is a general-purpose, omni-modal generative system
Official Python inference and LoRA trainer package
Make videos programmatically with React
Official Repo For "Sa2VA: Marrying SAM2 with LLaVA
MiniMax H3 inference engine for Mac computers
Multimodal-Driven Architecture for Customized Video Generation
Build Vision Agents quickly with any model or video provider
Video Object and Interaction Deletion
RGBD video generation model conditioned on camera input
Inference script for Oasis 500M
Pluggable SOTA multi-object tracking modules for segmentation
Project Lyra: Open Generative 3D World Models
CS2, Valorant, Fortnite, APEX, every game
The official pytorch implementation of our paper
Pytorch implementation of our method for high-resolution
The implementation of an algorithm presented in the CVPR18 paper
The Janelia Automated Animal Behavior Annotator
Google’s flagship dense multimodal model for coding and reasoning