Linear Function Approximation optimal control
Q-Learning with LFA may diverge. [B,G]
Sarsa with LFA converges [NR,S].
Monte Carlo with LFA converges.
השקופית הקודמת
השקופית הבאה
חזור אל השקופית הראשונה
הצג גירסה גרפית