MARSE uses continuous DAC codec representations for speech enhancement
September 4, 2026
MARSE performs speech enhancement by iteratively decoding masked clean speech frames using continuous latent representations from the DAC neural audio codec. Utilizing a Conformer-based architecture, the method allows for adjustable trade-offs between speech quality and computational overhead compared to discrete token approaches.
HOW THIS AFFECTS YOU
●
builderYou can tune the balance between audio fidelity and inference latency in speech enhancement pipelines.
●
researcherYou can explore the benefits of continuous versus discrete latent spaces for generative audio tasks.