Temporal Difference
الفارق الزمني الحسابي
منهجية تحديث تجمع بين مزايا مونت كارلو والبرمجة الديناميكية لتقدير القيم بناءً على التوقعات المتلاحقة بين الخطوات.
Temporal Difference
Also translated asخوارزميات التنبؤ المعتمدة على الخطوات البينية، معيار التعلم عبر الفجوات الزمنية المتتالية
First appears in this corpus in: Some Studies in Machine Learning Using the Game of Checkers (1959)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Mastering the Game of Go with Deep Neural Networks and Tree Search2016in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures2018in the sky ✦
- IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures2018in the sky ✦
- Between MDPs and Semi-MDPs — A Framework for Temporal Abstraction in Reinforcement Learning1999in the sky ✦
- Policy Gradient Methods for Reinforcement Learning with Function Approximation1999in the sky ✦
- Q-Learning1992in the sky ✦
- Q-Learning1992in the sky ✦
- Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning1992in the sky ✦
- Some Studies in Machine Learning Using the Game of Checkers1959in the sky ✦
- Some Studies in Machine Learning Using the Game of Checkers1959in the sky ✦
- Temporal Difference Learning and TD-Gammon1995in the sky ✦
- Temporal Difference Learning and TD-Gammon1995in the sky ✦
- Learning to Predict by the Methods of Temporal Differences1988in the sky ✦
- Learning to Predict by the Methods of Temporal Differences1988in the sky ✦