Reward Shaping
صياغة وهندسة دالة المكافأة
تقنية في التعلم بالتعزيز تضيف مكافآت بينية صغيرة وموجهة لتسهيل وتسريع تعلم العميل في البيئات المعقدة.
Reward Shaping
Also translated asتحسين قيم التحفيز لتسريع توجيه العميل الذكي، هيكلة المكافآت البينية لتسهيل التعلم
First appears in this corpus in: Deep Reinforcement Learning from Human Preferences (2017)
Appears in these papers
- Grandmaster Level in StarCraft II Using Multi-Agent Reinforcement Learning2019in the sky ✦
- Deep Reinforcement Learning from Human Preferences2017in the sky ✦
- Deep Reinforcement Learning from Human Preferences2017in the sky ✦
- Mastering Diverse Domains Through World Models2023in the sky ✦
- First Return, Then Explore2021in the sky ✦
- First Return, Then Explore2021in the sky ✦
- Hindsight Experience Replay2017in the sky ✦
- Hindsight Experience Replay2017in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- Learning Dexterous In-Hand Manipulation2019in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- QT-Opt: Scalable Deep Reinforcement Learning for Vision-Based Robotic Manipulation2018in the sky ✦