State
الحالة
اللقطة البيانية الحالية التي تلخص وضع النظام أو العميل في لحظة زمنية محددة.
State
Also translated asالوضع البنيوي، المتغير الحالي للنظام
First appears in this corpus in: Extension of the Law of Large Numbers to Dependent Quantities (1906)
Appears in these papers
- Mastering the Game of Go Without Human Knowledge2017in the sky ✦
- Learning Long-Term Dependencies with Gradient Descent is Difficult1994in the sky ✦
- Continuous Control with Deep Reinforcement Learning2015in the sky ✦
- Deep Reinforcement Learning with Double Q-Learning2016in the sky ✦
- Dynamic Programming1957in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- Hindsight Experience Replay2017in the sky ✦
- A Tutorial on Hidden Markov Models and Selected Applications in Speech Recognition1989in the sky ✦
- Neural Networks and Physical Systems with Emergent Collective Computational Abilities1982in the sky ✦
- Curiosity-Driven Exploration by Self-Supervised Prediction2017in the sky ✦
- A New Approach to Linear Filtering and Prediction Problems1960in the sky ✦
- Extension of the Law of Large Numbers to Dependent Quantities1906in the sky ✦
- Extension of the Law of Large Numbers to Dependent Quantities1906in the sky ✦
- The Monte Carlo Method1949in the sky ✦
- Between MDPs and Semi-MDPs — A Framework for Temporal Abstraction in Reinforcement Learning1999in the sky ✦
- Q-Learning1992in the sky ✦
- Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning1992in the sky ✦
- Optimization by Simulated Annealing1983in the sky ✦
- Temporal Difference Learning and TD-Gammon1995in the sky ✦