Stage 14 · Advanced Robotics / Reinforcement Learning · Lesson 8 · 14–20 min
DQN Core Ideas
DQN-style learning keeps a Q-network, stores transitions in a replay buffer, and bootstraps with a lagged target network — three ideas that stabilize value-based deep RL (intuition level).