Robust recipes to align language models with human and AI preferences
Uses Qwen3-ASR, local LLM, Whisper, TEN-VAD
A dataset consists of 15,140 ChatGPT prompts from Reddit
Course to get into Large Language Models (LLMs)
Recipes to train reward model for RLHF
The official repository for ERNIE 4.5 and ERNIEKit
Fully automatic censorship removal for language models
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs
Agentic, Reasoning, and Coding (ARC) foundation models
Semi-Structured Agentic Framework. Workflows build themselves
A Survey of Large Language Models
Unleashing 10,000+ Word Generation from Long Context LLMs
Constrained Value Alignment via Safe Reinforcement Learning
Chinese safety prompts for evaluating and improving the safety of LLMs
Quick guide (especially) for trending instruction finetuning dataset
Official release of InternLM series
Training Language Models to Follow Instructions with Human Feedback