Exploitation
الاستغلال (اعتماد الأفعال الناجحة)
اعتماد العميل الذكي على الأفعال الناجحة والمعروفة مسبقاً لتعظيم المكاسب المضمونة والآنية.
Exploitation
Also translated asتعظيم العوائد بناءً على المعرفة الحالية، الاستثمار الفعلي للمسارات المجربة
First appears in this corpus in: Equation of State Calculations by Fast Computing Machines (1953)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Mastering the Game of Go Without Human Knowledge2017in the sky ✦
- Mastering the Game of Go with Deep Neural Networks and Tree Search2016in the sky ✦
- A General Reinforcement Learning Algorithm That Masters Chess, Shogi, and Go Through Self-Play2018in the sky ✦
- A General Reinforcement Learning Algorithm That Masters Chess, Shogi, and Go Through Self-Play2018in the sky ✦
- Practical Bayesian Optimization of Machine Learning Algorithms2012in the sky ✦
- Practical Bayesian Optimization of Machine Learning Algorithms2012in the sky ✦
- Conservative Q-Learning for Offline Reinforcement Learning2020in the sky ✦
- Continuous Control with Deep Reinforcement Learning2015in the sky ✦
- Human-Level Control Through Deep Reinforcement Learning2015in the sky ✦
- Hyperband: A Novel Bandit-Based Approach to Hyperparameter Optimization2018in the sky ✦
- Equation of State Calculations by Fast Computing Machines1953in the sky ✦
- MuZero: Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model2020in the sky ✦
- Between MDPs and Semi-MDPs — A Framework for Temporal Abstraction in Reinforcement Learning1999in the sky ✦
- Q-Learning1992in the sky ✦
- QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation2018in the sky ✦
- Rainbow: Combining Improvements in Deep Reinforcement Learning2018in the sky ✦
- Rainbow: Combining Improvements in Deep Reinforcement Learning2018in the sky ✦
- Simple Statistical Gradient-Following Algorithms for Connectionist Reinforcement Learning1992in the sky ✦
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor2018in the sky ✦
- Optimization by Simulated Annealing1983in the sky ✦
- Temporal Difference Learning and TD-Gammon1995in the sky ✦