Glossary

Multimodal

متعدد الوسائط

نموذج يستطيع معالجة أنواع متعددة من المدخلات — كالنص والصور والصوت — بدلاً من النص وحده، مما يمنحه فهماً أوسع للعالم.

A model that can process multiple types of input — such as text, images, and audio — rather than text alone, giving it a broader understanding of the world.

Also translated asشامل الوسائط، عابر الوسائط، عابر للوسائط، متعدد الأنماط، مُتعدِّد الأنماط

First appears in this corpus in: VQA: Visual Question Answering (2015)

Appears in these papers