Automated translation solution for visual novels
An agentless approach to automatically solve software development
Video understanding codebase from FAIR for reproducing video models
Machine Learning Systems: Design and Implementation
Open-source Video Translation Skill
An open sourced end-to-end VLM-based GUI Agent
High-Performance Face Recognition Library on PaddlePaddle & PyTorch
Refer and Ground Anything Anywhere at Any Granularity
PyTorch code and models for VJEPA2 self-supervised learning from video
OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark
Visual localization made easy with hloc
Code release for "Detecting Twenty-thousand Classes
Official implementation for UniVL video and language training models
gradslam is an open source differentiable dense SLAM library
Web-based image segmentation tool for object detection & localization
ChainerCV: a Library for Deep Learning in Computer Vision
http://www.sciencedirect.com/science/article/pii/S1047847711003492