Environment
البيئة التفاعلية
المحيط أو النطاق التشغيلي الافتراضي الذي يتفاعل معه العميل الذكي ويتلقى منه التغذية الراجعة.
Environment
Also translated asالنطاق التشغيلي، المحيط الافتراضي
First appears in this corpus in: A Proposal for the Dartmouth Summer Research Project on Artificial Intelligence (1956)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- Concrete Problems in AI Safety2016in the sky ✦
- Conservative Q-Learning for Offline Reinforcement Learning2020in the sky ✦
- Conservative Q-Learning for Offline Reinforcement Learning2020in the sky ✦
- A Proposal for the Dartmouth Summer Research Project on Artificial Intelligence1956in the sky ✦
- Deep Reinforcement Learning from Human Preferences2017in the sky ✦
- Human-Level Control Through Deep Reinforcement Learning2015in the sky ✦
- Dueling Network Architectures for Deep Reinforcement Learning2016in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- First Return, Then Explore2021in the sky ✦
- First Return, Then Explore2021in the sky ✦
- Curiosity-Driven Exploration by Self-Supervised Prediction2017in the sky ✦
- IMPALA: Scalable Distributed Deep-RL with Importance Weighted Actor-Learner Architectures2018in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- MuZero: Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model2020in the sky ✦
- MuZero: Mastering Atari, Go, Chess and Shogi by Planning with a Learned Model2020in the sky ✦
- Between MDPs and Semi-MDPs — A Framework for Temporal Abstraction in Reinforcement Learning1999in the sky ✦
- Policy Gradient Methods for Reinforcement Learning with Function Approximation1999in the sky ✦
- Proximal Policy Optimization Algorithms2017in the sky ✦
- Q-Learning1992in the sky ✦
- ReAct: Synergizing Reasoning and Acting in Language Models2023in the sky ✦
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor2018in the sky ✦
- Soft Actor-Critic: Off-Policy Maximum Entropy Deep Reinforcement Learning with a Stochastic Actor2018in the sky ✦
- WebArena: A Realistic Web Environment for Building Autonomous Agents2023in the sky ✦
- WebArena: A Realistic Web Environment for Building Autonomous Agents2023in the sky ✦
- World Models2018in the sky ✦
- World Models2018in the sky ✦