Deep Understanding AI Agents
TextWorld is a sandbox learning environment for the training
Learn how to develop, deploy and iterate on production-grade ML
Open-source, high-performance AI model with advanced reasoning
Machine Learning Journal for Intermediate to Advanced Topics
ML engineer that reads papers, trains models, and ships ML models
Python observability platform for tracing apps, metrics, and logs
Advanced evolutionary computation library built on top of PyTorch
SAPIEN Manipulation Skill Framework
Python library for portfolio optimization built on top of scikit-learn
Training framework for Stable Baselines3 reinforcement learning agents
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs
Powerful AI language model (MoE) optimized for efficiency/performance
Implementation of RLHF (Reinforcement Learning with Human Feedback)
C++-based high-performance parallel environment execution engine
Numerical differential equation solvers in JAX
OpenTinker is an RL-as-a-Service infrastructure for foundation models
Cosmos-RL is a flexible and scalable Reinforcement Learning framework
Build multimodal AI applications with cloud-native stack
AI-Powered Personalized Learning Assistant
Helps data scientists define testable self-documenting dataflows
Build a large language model from 0 only with Python foundation
Reference PyTorch implementation and models for DINOv3
Large Language Model Principles and Practice Tutorial from Scratch
RL implementations