Large Language Models (LLMs) - Stanford CME295 Transformers & LLMs | Autumn 2025 | Lecture 6 - LLM Reasoning
4.0(2)
25 learners
What you'll learn
Analyze case studies to identify core principles
Apply theoretical frameworks to real-world scenarios
Evaluate evidence to support arguments
Synthesize information from multiple sources
For more information about Stanford’s graduate programs, visit: https://online.stanford.edu/graduate-education
November 7, 2025
This lecture covers:
• Reasoning models
• RL for reasoning
• GRPO
• Scaling
To follow along with the course schedule and syllabus, visit: https://cme295.stanford.edu/syllabus/
Chapters:
00:00:00 Introduction
00:12:43 Reasoning models
00:27:49 Benchmarks
00:32:04 Pass@k metric
00:48:07 Scaling with RL
00:57:44 GRPO
01:06:03 Comparison between GRPO and PPO
01:16:14 Length bias
01:25:00 DAPO, Dr. GRPO
01:29:38 DeepSeek R1 recipe
Afshine Amidi is an Adjunct Lecturer at Stanford University.
Shervine Amidi is an Adjunct Lecturer at Stanford University.
View the course playlist: https://www.youtube.com/playlist?list=PLoROMvodv4rOCXd21gf0CF4xr35yINeOy
Continue this lesson in the app
Install CourseHive on Android or iOS to keep learning while you move.