LLM Learning Hub

workspace/llm-course/home

MDSection 25 of 38 ยท 6 subtopics

Alignment

Making models helpful and safe. RLHF, reward models, PPO, and DPO.

Learning path

Work through each subtopic in order. Click any file below to open its lesson. Track your progress with the checkbox at the bottom.

Subtopics

Each subtopic includes a concise explanation, code examples, and references.