Glossary

Annotation Vector

متجه التعليق التوضيحي

متجه سمات يصف منطقة مكانية معينة من الصورة، يُستخرج من طبقة التفافية وسطى. في نموذج Show, Attend and Tell، تُنتج الشبكة الالتفافية 196 متجهاً (14×14)، كل منها بأبعاد 512.

A feature vector describing a specific spatial region of an image, extracted from an intermediate convolutional layer. In Show, Attend and Tell, the CNN produces 196 vectors (14×14), each of dimension 512.

Also translated asمتجه الشرح

First appears in this corpus in: Show, Attend and Tell: Neural Image Caption Generation with Visual Attention (2015)

Appears in these papers