A New Axis of Sparsity for Large Language Models
The knowledge and task management backbone for AI coding assistants
Open-source infrastructure for Computer-Use Agents. Sandboxes
"Big Model" trains a visual multimodal VLM with 26M parameters
Simplifies the local serving of AI models from any source
Collection of Gemma 3 variants that are trained for performance
Collection of reference environments, offline reinforcement learning
LLM training in simple, raw C/CUDA
PPTAgent: Generating and Evaluating Presentations
A simple, secure MCP-to-OpenAPI proxy server
Implementation of "MobileCLIP" CVPR 2024
Code release for Cut and Learn for Unsupervised Object Detection
VMZ: Model Zoo for Video Modeling
Official implementation of Watermark Anything with Localized Messages
Training Large Language Model to Reason in a Continuous Latent Space
High-resolution models for human tasks
Code for the paper "Evaluating Large Language Models Trained on Code"
Tool for exploring and debugging transformer model behaviors
CLIP, Predict the most relevant text snippet given an image
Ling is a MoE LLM provided and open-sourced by InclusionAI
A Unified Framework for Text-to-3D and Image-to-3D Generation
Multimodal Diffusion with Representation Alignment
Personalize Any Characters with a Scalable Diffusion Transformer
Talk to Your AI Agents from Anywhere
The NVIDIA AgentIQ toolkit is an open-source library