A straightforward method for training your LLM
Train any agents simply by 'talking'
Curated list of datasets and tools for post-training
Project aimed at extracting, exporting, and analyzing chat records
Designed for training LLM/VLM agents via RL
Libre Survival Manual for Android with offline in mind
Claude Code is an agentic coding tool that lives in your terminal
Deep learning optimization library: makes distributed training easy
Faster and easier training and deployments
Learning to Reason with Search for LLMs via Reinforcement Learning
A simple, performant and scalable Jax LLM
Train multi-step agents for real-world tasks using GRPO
Retrieval and Retrieval-augmented LLMs
TextWorld is a sandbox learning environment for the training
Experimental notebook pipeline for creating language models
Roadmap to becoming an Artificial Intelligence Expert in 2022
Learning What You Want to Learn Using Programmable Gradient Info
Implementation and experiments of graph embedding algorithms
Training framework for Stable Baselines3 reinforcement learning agents
A deep matching model library for recommendations & advertising
Generates original ARC-AGI-1-style tasks distribution-matched
Reinforcement Learning for Humanoid Robot with Zero-Shot Sim2Real
A full-stack codebase for training and evaluating speculative decoding
AI Code Security Anti-Patterns distilled from 150+ sources
A Next-Generation Training Engine Built for Ultra-Large MoE Models