Week 8 · LLM Post-Training & Building AI AssistantsRL based Fine Tuning← Previous8.3 Supervised Fine tuningNext →8.5 Pitfalls and Advanced RL