A straightforward method for training your LLM
Designed for training LLM/VLM agents via RL
A deep matching model library for recommendations & advertising
Learning What You Want to Learn Using Programmable Gradient Info
A full-stack codebase for training and evaluating speculative decoding
Train multi-step agents for real-world tasks using GRPO
"Big Model" trains a visual multimodal VLM with 26M parameters
RF-DETR is a real-time object detection and segmentation
Faster and easier training and deployments
Learning to Reason with Search for LLMs via Reinforcement Learning
Use pretrained transformers like BERT, XLNet and GPT-2 in spaCy
Experimental notebook pipeline for creating language models
A simple, performant and scalable Jax LLM
Ongoing research training transformer models at scale
Deep learning optimization library: makes distributed training easy
Recipes to train reward model for RLHF
Llama Chinese community, real-time aggregation
Implementation and experiments of graph embedding algorithms
Scalable machine learning for time series forecasting
Retrieval and Retrieval-augmented LLMs
Minimal reproduction of OneRec
MLOps tools for managing & orchestrating the ML LifeCycle
Code release for Cut and Learn for Unsupervised Object Detection
High-resolution models for human tasks
PyTorch code and models for VJEPA2 self-supervised learning from video