Acoustic Token
رمز صوتي
رمز منفصل يُنتجه مرمِّز صوتي عصبي (مثل SoundStream) عبر التكميم المتجهي المتبقي، ويرمِّز التفاصيل الصوتية الدقيقة كالطابع والصدى والديناميكيات اللازمة لإعادة بناء صوت عالي الدقة.
A discrete token produced by a neural audio codec (like SoundStream) via residual vector quantization, encoding fine acoustic details such as timbre, reverb, and dynamics needed for high-fidelity audio reconstruction.
Also translated asوحدة صوتية
First appears in this corpus in: MusicLM: Generating Music From Text (2023)
Appears in these papers