Using Random Codebooks for Audio Neural AutoEncoders
EUROPEAN SIGNAL PROCESSING CONFERENCE 2024 [EUSIPCO], Aug 2024, Lyon, France Latent representation learning has been an active field of study for decades in numerous applications. Inspired among others by the tokenization from Natural Language Processing and motivated by the research of a simple dat...
Gespeichert in:
Hauptverfasser: | , , , |
---|---|
Format: | Artikel |
Sprache: | eng |
Online-Zugang: | Volltext bestellen |
Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Zusammenfassung: | EUROPEAN SIGNAL PROCESSING CONFERENCE 2024 [EUSIPCO], Aug 2024,
Lyon, France Latent representation learning has been an active field of study for decades
in numerous applications. Inspired among others by the tokenization from
Natural Language Processing and motivated by the research of a simple data
representation, recent works have introduced a quantization step into the
feature extraction. In this work, we propose a novel strategy to build the
neural discrete representation by means of random codebooks. These codebooks
are obtained by randomly sampling a large, predefined fixed codebook. We
experimentally show the merits and potential of our approach in a task of audio
compression and reconstruction. |
---|---|
DOI: | 10.48550/arxiv.2409.16677 |