A straightforward method for training your LLM
Designed for training LLM/VLM agents via RL
Train multi-step agents for real-world tasks using GRPO
"Big Model" trains a visual multimodal VLM with 26M parameters
RF-DETR is a real-time object detection and segmentation
Faster and easier training and deployments
Learning to Reason with Search for LLMs via Reinforcement Learning
Experimental notebook pipeline for creating language models
A simple, performant and scalable Jax LLM
Deep learning optimization library: makes distributed training easy
Recipes to train reward model for RLHF
Llama Chinese community, real-time aggregation
Scalable machine learning for time series forecasting
Retrieval and Retrieval-augmented LLMs
Minimal reproduction of OneRec
MLOps tools for managing & orchestrating the ML LifeCycle
Code release for Cut and Learn for Unsupervised Object Detection
High-resolution models for human tasks
PyTorch code and models for VJEPA2 self-supervised learning from video
Learn how to develop, deploy and iterate on production-grade ML
Towards Human-Sounding Speech
Play couplet with seq2seq model
A fast TTS architecture with conditional flow matching
Sandbox for training deep learning networks
Turn expensive prompts into cheap fine-tuned models