Advanced techniques for RAG systems
Fast and Universal 3D reconstruction model for versatile tasks
Implementation of Vision Transformer, a simple way to achieve SOTA
A secure sandbox environment for malware developers and red teamers
4M: Massively Multimodal Masked Modeling
Guiding Instruction-based Image Editing via Multimodal Large Language
This repository contains the official implementation of FastVLM
Refer and Ground Anything Anywhere at Any Granularity
Foundation Models for Time Series
Set of tools to assess and improve LLM security
Fast3R: Towards 3D Reconstruction of 1000+ Images in One Forward Pass
A Production-ready Reinforcement Learning AI Agent Library
PyTorch code and models for V-JEPA self-supervised learning from video
A PyTorch library for implementing flow matching algorithms
An implementation of a deep learning recommendation model (DLRM)
Self-supervised visual learning using momentum contrast in PyTorch
ImageBind One Embedding Space to Bind Them All
PyTorch code and models for the DINOv2 self-supervised learning
ChatGLM3 series: Open Bilingual Chat LLMs | Open Source Bilingual Chat
[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences
Official implementation of DreamCraft3D
tiktoken is a fast BPE tokeniser for use with OpenAI's models
The simplest, fastest repository for training/finetuning models
Diffusion Transformer with Fine-Grained Chinese Understanding
NVIDIA Isaac GR00T N1.5 is the world's first open foundation model