State-Value Function
دالة قيمة الحالة الحالية
دالة تحسب الجدوى المطلقة للتواجد في وضعية راهنة تحت سياق السياسة التشغيلية المتبعة.
State-Value Function
Also translated asمعيار الجدوى المطلقة للوضعية، القيمة الرياضية للموقع الراهن
First appears in this corpus in: Policy Gradient Methods for Reinforcement Learning with Function Approximation (1999)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Dueling Network Architectures for Deep Reinforcement Learning2016in the sky ✦
- Dueling Network Architectures for Deep Reinforcement Learning2016in the sky ✦
- IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures2018in the sky ✦
- Policy Gradient Methods for Reinforcement Learning with Function Approximation1999in the sky ✦
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor2018in the sky ✦
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor2018in the sky ✦