Masked Autoregressive Speech Enhancement with Continuous Neural Audio Codec Representations
This work proposes masked autoregressive SE (MARSE), a method for SE based on iterative decoding of masked clean speech frames using continuous NAC representations of speech, which enables a flexible trade-off between SE performance and computational cost.
Yoto Fujita, Simon Leglaive, Laurent Girin
· 0 citations