A straightforward method for training your LLM
SRCP model train controller
Designed for training LLM/VLM agents via RL
A deep matching model library for recommendations & advertising
Learning What You Want to Learn Using Programmable Gradient Info
A full-stack codebase for training and evaluating speculative decoding
Train multi-step agents for real-world tasks using GRPO
"Big Model" trains a visual multimodal VLM with 26M parameters
RF-DETR is a real-time object detection and segmentation
Graphical SRCP locking table client
Faster and easier training and deployments
Learning to Reason with Search for LLMs via Reinforcement Learning
Use pretrained transformers like BERT, XLNet and GPT-2 in spaCy
Experimental notebook pipeline for creating language models
A simple, performant and scalable Jax LLM
Deep learning optimization library: makes distributed training easy
Ongoing research training transformer models at scale
Recipes to train reward model for RLHF
Llama Chinese community, real-time aggregation
Implementation and experiments of graph embedding algorithms
Scalable machine learning for time series forecasting
Retrieval and Retrieval-augmented LLMs
Minimal reproduction of OneRec
MLOps tools for managing & orchestrating the ML LifeCycle
Code release for Cut and Learn for Unsupervised Object Detection