Constitutional AI
الذكاء الاصطناعي الدستوري
منهجية لتدريب أنظمة الذكاء الاصطناعي على عدم الإيذاء باستخدام مجموعة مبادئ مكتوبة بلغة طبيعية (دستور) بدلاً من تصنيفات بشرية، عبر النقد الذاتي والتنقيح وتوليد تصنيفات التفضيل آلياً.
A method for training AI systems to be harmless using a set of natural-language principles (a constitution) instead of human labels, through self-critique, revision, and AI-generated preference labels.
Also translated asAI الدستوري، الذكاء الدستوري، المحاذاة بالمبادئ، المواءمة الذاتية الموجهة بدستور نظامي، ضبط السلوك بناءً على قواعد ومبادئ مكتوبة، منهجية CAI
First appears in this corpus in: Concrete Problems in AI Safety (2016)
Appears in these papers
- Concrete Problems in AI Safety2016in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- Training Language Models to Follow Instructions with Human Feedback2022in the sky ✦
- RLAIF: Scaling Reinforcement Learning from Human Feedback with AI Feedback2023in the sky ✦
- RLAIF: Scaling Reinforcement Learning from Human Feedback with AI Feedback2023in the sky ✦