Stage 14 · Advanced Robotics / Reinforcement Learning · Lesson 7 · 12–18 min
Function Approximation Intuition
Instead of a table cell per state, approximate Q(s,a) with features or a small network — nearby states share parameters and thus share learning.