Pretrain / fine-tune / RLHF
The difference between a raw language model and an assistant you can talk to. And how thin that difference is.
2 steps
110 XP
A free account is needed
Start the lesson →
Sources
- Ouyang, L. et al. 2022 · Training Language Models to Follow Instructions with Human Feedback (InstructGPT) · NeurIPS 2022
- Rafailov, R. et al. 2023 · Direct Preference Optimization (DPO) · NeurIPS 2023
- Zhou, C. et al. 2023 · LIMA: Less Is More for Alignment · NeurIPS 2023
- Bai, Y. et al. 2022 · Constitutional AI: Harmlessness from AI Feedback · arXiv:2212.08073
ML Academy · an interactive machine learning course that runs in your browser ·
All lessons