Pretraining LLM
InstructGPT Finetuning
Reinforcement Learning with Human Feedback (RLHF)
Last updated 3 years ago
Was this helpful?