Semantic Token
رمز دلالي
رمز منفصل يُستخلص من نموذج صوتي ذاتي الإشراف (مثل w2v-BERT) ويلتقط المعلومات البنيوية والدلالية للصوت كاللحن والإيقاع والنوع الموسيقي، بمعزل عن التفاصيل الصوتية الدقيقة.
A discrete token extracted from a self-supervised audio model (like w2v-BERT) that captures structural and semantic information of audio such as melody, rhythm, and genre, independent of fine acoustic details.
Also translated asوحدة دلالية
First appears in this corpus in: MusicLM: Generating Music From Text (2023)
Appears in these papers