Patch Embedding
تضمين الرُّقع
عملية تقسيم الصورة إلى مربعات صغيرة ثابتة الحجم وبسط كل مربع إلى متجه يُسقط خطياً إلى فضاء التضمين.
The process of splitting an image into fixed-size square patches, flattening each into a vector, and linearly projecting it into embedding space.
Also translated asإسقاط الرُّقع، تشفير الأجزاء المُرقَّعة
First appears in this corpus in: An Image Is Worth 16×16 Words: Transformers for Image Recognition at Scale (2020)
Appears in these papers
- BEiT: BERT Pre-Training of Image Transformers2021in the sky ✦
- BEiT: BERT Pre-Training of Image Transformers2021in the sky ✦
- BLIP-2: Bootstrapping Language-Image Pre-Training with Frozen Image Encoders and Large Language Models2023in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- Emerging Properties in Self-Supervised Vision Transformers2021in the sky ✦
- Emerging Properties in Self-Supervised Vision Transformers2021in the sky ✦
- Gemini: A Family of Highly Capable Multimodal Models2023in the sky ✦
- A Time Series Is Worth 64 Words: Long-Term Forecasting with Transformers2023in the sky ✦
- SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers2021in the sky ✦
- SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers2021in the sky ✦
- Swin Transformer: Hierarchical Vision Transformer Using Shifted Windows2021in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- An Image Is Worth 16×16 Words: Transformers for Image Recognition at Scale2020in the sky ✦