Implementation of Vision Transformer, a simple way to achieve SOTA
ICLR2024 Spotlight: curation/training code, metadata, distribution
Blazeface is a lightweight model that detects faces in images
A fast, powerful, and simple hierarchical vision transformer
CoTracker is a model for tracking any point (pixel) on a video
Class Activation Mapping
Codebase for Image Classification Research, written in PyTorch
End-to-end object detection with transformers
Latest techniques in deep learning and representation learning
Estimates the psychovisual difference between two images
Various hashing methods for image retrieval and serves as the baseline