Materialization
التجسيد
كتابة مصفوفة وسيطة بالكامل في الذاكرة الرئيسية. FlashAttention يتجنب تجسيد مصفوفة الانتباه N×N وهو ما يوفّر الذاكرة بشكل جذري.
Writing out a complete intermediate matrix to main memory. FlashAttention avoids materializing the N×N attention matrix, which is what saves memory so dramatically.
Also translated asتجسيد المصفوفة، التخزين الوسيط الكامل
First appears in this corpus in: FlashAttention: Fast and Memory-Efficient Exact Attention with IO-Awareness (2022)
Appears in these papers