Visual intelligence for your home.
Open source framework for deep learning satellite and aerial imagery
Implementation of Vision Transformer, a simple way to achieve SOTA
Automatically find issues in image datasets
Effortless data labeling with AI support from Segment Anything
Fast image augmentation library and an easy-to-use wrapper
Training data (data labeling, annotation, workflow) for all data types
Open Source Differentiable Computer Vision Library
The open-source tool for building high-quality datasets
Medical imaging toolkit for deep learning
An Open Real-time Video-Language Interaction System
A Pragmatic VLA Foundation Model
ICLR2024 Spotlight: curation/training code, metadata, distribution
NVIDIA Isaac GR00T N1.5 is the world's first open foundation model
Data integration platform for ELT pipelines from APIs, databases
Hub of ready-to-use datasets for ML models
Deep learning library
The largest collection of PyTorch image encoders / backbones
Reference PyTorch implementation and models for DINOv3
Structured data extraction and instruction calling with ML, LLM
Automate browser-based workflows with LLMs and Computer Vision
State-of-the-art Image & Video CLIP, Multimodal Large Language Models
An on-premises, OCR-free unstructured data extraction
Bridging Reasoning and Action Prediction
LLM inference server with continuous batching & SSD caching