Glossary

Block-Sparse Attention

الانتباه المتناثر الكُتَلي

امتداد لـ FlashAttention يتخطى كتلاً كاملة من مصفوفة الانتباه وفق قناع تناثر محدد مسبقاً، فيحقق تسريعاً إضافياً يتناسب مع نسبة الكتل المتخطّاة.

An extension of FlashAttention that skips entire blocks of the attention matrix according to a predefined sparsity mask, yielding additional speedup proportional to the fraction of blocks skipped.

Also translated asالانتباه المتفرق بالكتل، الانتباه الكتلي المبعثر

First appears in this corpus in: FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness (2022)

Appears in these papers