Reinforcement Learning from Human Feedback
التعلم بالتعزيز القائم على التقييم البشري
تدريب السلوك اللغوي عبر دمج تقييمات وتفضيلات البشر كمكافآت لتوجيه مخرجات الشبكة العصبية.
Also translated asمواءمة السلوك بالتغذية الراجعة البشرية، التدريب التدعيمي بإشراف إنساني، RLHF، التعلم التعزيزي من الملاحظات البشرية، التعلم التعزيزي بالتغذية البشرية، التعلم التعزيزي من الملاحظات البشرية
First appears in this corpus in: On Information and Sufficiency (1951)
Appears in these papers
- Asynchronous Methods for Deep Reinforcement Learning2016in the sky ✦
- AI Safety via Debate2018in the sky ✦
- Alpaca: A Strong, Replicable Instruction-Following Model2023in the sky ✦
- Concrete Problems in AI Safety2016in the sky ✦
- A Proposal for the Dartmouth Summer Research Project on Artificial Intelligence1956in the sky ✦
- DeepSeek-R1: Incentivizing Reasoning Capability in LLMs via Reinforcement Learning2025in the sky ✦
- Direct Preference Optimization: Your Language Model Is Secretly a Reward Model2023in the sky ✦
- Finetuned Language Models Are Zero-Shot Learners2022in the sky ✦
- High-Dimensional Continuous Control Using Generalized Advantage Estimation2016in the sky ✦
- Gemini: A Family of Highly Capable Multimodal Models2023in the sky ✦
- Gemma: Open Models Based on Gemini Research and Technology2024in the sky ✦
- Google's Neural Machine Translation System: Bridging the Gap Between Human and Machine Translation2016in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- GPT-4 Technical Report2023in the sky ✦
- Training Language Models to Follow Instructions with Human Feedback2022in the sky ✦
- On Information and Sufficiency1951in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- Learning to Summarize from Human Feedback2020in the sky ✦
- LIMA: Less Is More for Alignment2023in the sky ✦
- LIMA: Less Is More for Alignment2023in the sky ✦
- Llama 2: Open Foundation and Fine-Tuned Chat Models2023in the sky ✦
- Llama 2: Open Foundation and Fine-Tuned Chat Models2023in the sky ✦
- Red Teaming Language Models to Reduce Harms: Methods, Scaling Behaviors, and Lessons Learned2022in the sky ✦
- WebGPT: Browser-Assisted Question-Answering with Human Feedback2021in the sky ✦
- WebGPT: Browser-Assisted Question-Answering with Human Feedback2021in the sky ✦