Implementation of Vision Transformer, a simple way to achieve SOTA
ICLR2024 Spotlight: curation/training code, metadata, distribution
A fast, powerful, and simple hierarchical vision transformer
CoTracker is a model for tracking any point (pixel) on a video
Codebase for Image Classification Research, written in PyTorch
End-to-end object detection with transformers