Temporal-Difference Learning - TU Chemnitz
TD methods do not require a model of the environment, only experience! ? TD, but not MC, methods can be fully incremental!
Chapter 6: Temporal Difference LearningCompare efficiency of TD learning with MC learning. Then extend to control ... Figure 6.12: Q-learning: An off-policy TD control algorithm. Its simplest ... Gradient Temporal-Difference Learning Algorithms - Rich SuttonThree new algorithms. ? GTD, the original gradient TD algorithm. (Sutton, Szepevari & Maei, 2008). ? GTD-2, a second-generation GTD. ? TDC, TD with gradient ... Septembre 2006 N° 148 - Sites ENSFEA? Public : Etudiants Ingénieurs + Etudiants L1, L2, L3, M1 et M2. ? Niveau : BAC+3 à BAC+5. ? Cours (341 H équivalent TD) : ? 2007-2008 ...
Autres Cours: