Annotation Vector
متجه التعليق التوضيحي
متجه سمات يصف منطقة مكانية معينة من الصورة، يُستخرج من طبقة التفافية وسطى. في نموذج Show, Attend and Tell، تُنتج الشبكة الالتفافية 196 متجهاً (14×14)، كل منها بأبعاد 512.
A feature vector describing a specific spatial region of an image, extracted from an intermediate convolutional layer. In Show, Attend and Tell, the CNN produces 196 vectors (14×14), each of dimension 512.
Also translated asمتجه الشرح
First appears in this corpus in: Show, Attend and Tell: Neural Image Caption Generation with Visual Attention (2015)
Appears in these papers