Text Infilling
ملء النص
استراتيجية تشويش في التدريب المسبق تستبدل مقاطع نصية عشوائية — بأطوال مسحوبة من توزيع بواسون — بقناع واحد [MASK]، مما يُجبر النموذج على التنبؤ بمحتوى المقطع وعدد الرموز المفقودة معاً.
A pre-training corruption strategy that replaces random text spans — with lengths drawn from a Poisson distribution — with a single [MASK] token, forcing the model to predict both the span content and the number of missing tokens.
Also translated asملء الفراغات النصية، حشو النص
First appears in this corpus in: BART: Denoising Sequence-to-Sequence Pre-Training for Natural Language Generation, Translation, and Comprehension (2019)
Appears in these papers