Course Hive
Search

Welcome

Sign in or create your account

Continue with Google
or
Stanford CME295 Transformers & LLMs | Autumn 2025 | Lecture 6 - LLM Reasoning
Play lesson

Large Language Models (LLMs) - Stanford CME295 Transformers & LLMs | Autumn 2025 | Lecture 6 - LLM Reasoning

4.0 (2)
25 learners

What you'll learn

Analyze case studies to identify core principles
Apply theoretical frameworks to real-world scenarios
Evaluate evidence to support arguments
Synthesize information from multiple sources

This course includes

  • 34.5 hours of video
  • Certificate of completion
  • Access on mobile and TV

Summary

Keywords

Full Transcript

For more information about Stanford’s graduate programs, visit: https://online.stanford.edu/graduate-education November 7, 2025 This lecture covers: • Reasoning models • RL for reasoning • GRPO • Scaling To follow along with the course schedule and syllabus, visit: https://cme295.stanford.edu/syllabus/ Chapters: 00:00:00 Introduction 00:12:43 Reasoning models 00:27:49 Benchmarks 00:32:04 Pass@k metric 00:48:07 Scaling with RL 00:57:44 GRPO 01:06:03 Comparison between GRPO and PPO 01:16:14 Length bias 01:25:00 DAPO, Dr. GRPO 01:29:38 DeepSeek R1 recipe Afshine Amidi is an Adjunct Lecturer at Stanford University. Shervine Amidi is an Adjunct Lecturer at Stanford University. View the course playlist: https://www.youtube.com/playlist?list=PLoROMvodv4rOCXd21gf0CF4xr35yINeOy

Course Hive

Continue this lesson in the app

Install CourseHive on Android or iOS to keep learning while you move.

Related Courses

FAQs

Course Hive
Download CourseHive
Keep learning anywhere