Fast image augmentation library and an easy-to-use wrapper
The repository provides code for running inference with SAM 2
Phi-3.5 for Mac: Locally-run Vision and Language Models
[CVPR 2025 Best Paper Award] VGGT
Implementation of Vision Transformer, a simple way to achieve SOTA
ICLR2024 Spotlight: curation/training code, metadata, distribution
Hub of ready-to-use datasets for ML models
A fast, powerful, and simple hierarchical vision transformer
CoTracker is a model for tracking any point (pixel) on a video
Code release for ConvNeXt model
AI for GNU Image Manipulation Program
Fast, modular reference implementation of Instance Segmentation
ChainerCV: a Library for Deep Learning in Computer Vision