CLIP, Predict the most relevant text snippet given an image
Language modeling in a sentence representation space
Reference PyTorch implementation and models for DINOv3
PyTorch code and models for the DINOv2 self-supervised learning
Bidirectional token-classification model for identifiable info
800,000 step-level correctness labels on LLM solutions to MATH problem
Official PyTorch Implementation of "Scalable Diffusion Models"