Scalable Oversight
الإشراف القابل للتوسُّع
تقنيات تُسخِّر أنظمة الذكاء الاصطناعي لمساعدة البشر في الإشراف على أنظمة ذكاء اصطناعي أخرى، مما يُخفِّف من عنق زجاجة التصنيف البشري مع ازدياد قدرة النماذج.
Techniques that leverage AI systems to help humans supervise other AIs, reducing the bottleneck of human annotation as models become more capable.
Also translated asالرقابة القابلة للتوسع، الإشراف المُتدرِّج
First appears in this corpus in: Concrete Problems in AI Safety (2016)
Appears in these papers
- AI Safety via Debate2018in the sky ✦
- Concrete Problems in AI Safety2016in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Constitutional AI: Harmlessness from AI Feedback2022in the sky ✦
- Weak-to-Strong Generalization: Eliciting Strong Capabilities With Weak Supervision2023in the sky ✦