Glossary

Prioritized Experience Replay

إعادة التجربة بالأولوية

أسلوب في التعلّم المعزّز يُعيد تشغيل الانتقالات المهمة بوتيرة أعلى بدلاً من السحب المنتظم، مستخدماً مقدار خطأ الفرق الزمني كمقياس لأولوية كل انتقال، مع تصحيح الانحياز الناتج بأوزان أخذ العيّنات بالأهمية.

A reinforcement learning technique that replays important transitions more frequently instead of uniformly sampling, using the magnitude of the TD error as a priority measure for each transition, with importance sampling weights to correct the resulting bias.

Also translated asإعادة التشغيل بالأولوية، الإعادة ذات الأولوية

First appears in this corpus in: Rainbow: Combining Improvements in Deep Reinforcement Learning (2018)

Appears in these papers