Dual Encoder
المُرمِّز الثنائي
بنية بمُرمِّزَين منفصلَين (مثلاً واحد للصور وآخر للنصوص) ينتجان تضمينات في فضاء مشترك، مما يتيح المقارنة بين الوسائط.
An architecture with two separate encoders (e.g. one for images, one for text) that produce embeddings in a shared space, enabling cross-modal comparison.
Also translated asالبنية ثنائية المُرمِّز، المُرمِّز المزدوج
First appears in this corpus in: Dense Passage Retrieval for Open-Domain Question Answering (2020)
Appears in these papers
- ALIGN: Scaling Up Visual and Vision-Language Representation Learning with Noisy Text Supervision2021in the sky ✦
- ALIGN: Scaling Up Visual and Vision-Language Representation Learning with Noisy Text Supervision2021in the sky ✦
- Atlas: Few-shot Learning with Retrieval Augmented Language Models2023in the sky ✦
- Learning Transferable Visual Models from Natural Language Supervision2021in the sky ✦
- Learning Transferable Visual Models from Natural Language Supervision2021in the sky ✦
- Dense Passage Retrieval for Open-Domain Question Answering2020in the sky ✦
- Dense Passage Retrieval for Open-Domain Question Answering2020in the sky ✦