Video translation and dubbing tool powered by LLMs
Automated translation solution for visual novels
An agentless approach to automatically solve software development
Recognition and resolution of numbers, units, date/time, etc.
Video understanding codebase from FAIR for reproducing video models
Machine Learning Systems: Design and Implementation
The Chinese version of OpenClaw
Open-source Video Translation Skill
High-Performance Face Recognition Library on PaddlePaddle & PyTorch
Refer and Ground Anything Anywhere at Any Granularity
PyTorch code and models for VJEPA2 self-supervised learning from video
EvoBot is a Discord Music Bot built with TypeScript + Discord.js
OpenMMLab's Next Generation Video Understanding Toolbox and Benchmark
Visual localization made easy with hloc
Code release for "Detecting Twenty-thousand Classes
Fast face detection, pupil/eyes localization
Joint Face Detection and Alignment
Official implementation for UniVL video and language training models
gradslam is an open source differentiable dense SLAM library
Web-based image segmentation tool for object detection & localization
ChainerCV: a Library for Deep Learning in Computer Vision
Simple Node.js package for robust face detection and face recognition
http://www.sciencedirect.com/science/article/pii/S1047847711003492
An easy, flexible, and accurate plate recognition project