Positional Encoding
الترميز الموضعي
قيم رياضية ثابتة أو متعلمة تدمج مع أوزان الرموز لتعريف النموذج بترتيب ومواقع الكلمات.
Positional Encoding
Also translated asالتشفير الدلالي للمواقع، ترميز رتبة الرمز
First appears in this corpus in: Attention Is All You Need (2017)
Appears in these papers
- DeBERTa: Decoding-Enhanced BERT with Disentangled Attention2020in the sky ✦
- Decision Transformer: Reinforcement Learning via Sequence Modeling2021in the sky ✦
- Decision Transformer: Reinforcement Learning via Sequence Modeling2021in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- Training Data-Efficient Image Transformers & Distillation Through Attention2021in the sky ✦
- End-to-End Object Detection with Transformers2020in the sky ✦
- Are Transformers Effective for Time Series Forecasting?2023in the sky ✦
- Graph Attention Networks2018in the sky ✦
- Scaling Language Models: Methods, Analysis & Insights from Training Gopher2022in the sky ✦
- Do Transformers Really Perform Bad for Graph Representation?2021in the sky ✦
- Do Transformers Really Perform Bad for Graph Representation?2021in the sky ✦
- Informer: Beyond Efficient Transformer for Long Sequence Time Series Forecasting2021in the sky ✦
- Informer: Beyond Efficient Transformer for Long Sequence Time Series Forecasting2021in the sky ✦
- LLaMA: Open and Efficient Foundation Language Models2023in the sky ✦
- NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis2020in the sky ✦
- NeRF: Representing Scenes as Neural Radiance Fields for View Synthesis2020in the sky ✦
- PaLM: Scaling Language Modeling with Pathways2022in the sky ✦
- RoFormer: Enhanced Transformer with Rotary Position Embedding2021in the sky ✦
- SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers2021in the sky ✦
- SegFormer: Simple and Efficient Design for Semantic Segmentation with Transformers2021in the sky ✦
- Segment Anything2023in the sky ✦
- Exploring the Limits of Transfer Learning with a Unified Text-to-Text Transformer2019in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- A Decoder-Only Foundation Model for Time Series Forecasting2024in the sky ✦
- Attention Is All You Need2017in the sky ✦
- Attention Is All You Need2017in the sky ✦
- An Image Is Worth 16×16 Words: Transformers for Image Recognition at Scale2020in the sky ✦
- Robust Speech Recognition via Large-Scale Weak Supervision2022in the sky ✦
- XLNet: Generalized Autoregressive Pretraining for Language Understanding2019in the sky ✦