PyTorch code and models for V-JEPA self-supervised learning from video
Must Reading Papers, Research Library, Open-Source Code
Official code base for LeWorldModel: Stable End-to-End Joint-Embedding
Mixture-of-Experts Vision-Language Models for Advanced Multimodal
End-to-end pipeline converting generative videos
Probabilistic Circuits from the Juice library
A 2D rigid body physics engine for the web
Software for molecular simulations and trajectory analysis
CoTracker is a model for tracking any point (pixel) on a video
Official codebase for I-JEPA
A Strong and Easy-to-use Single View 3D Hand+Body Pose Estimator
Joint Face Detection and Alignment
WaveRNN Vocoder + TTS
The official pytorch implementation of our paper
A real-time approach for mapping all human pixels of 2D RGB images
Efficient 3D human pose estimation in video using 2D keypoint
Deep Hough Voting for 3D Object Detection in Point Clouds
XML-Print: typesetting arbitrary XML documents in high quality
Semantic image segmentation method described in the ICCV 2015 paper
Enterprise Content Management & DMS- Version Control, Scan, Barcode...
the EXercise event Injection TOolkit
AMICI enables real-time execution of cyber-physical models (Simulink)