Get Started
Home
Topics
Search
Library
Research questionHow can speech enhancement use continuous audio representations while balancing quality against decoding cost?Masked generative speech enhancement reconstructs clean speech from noisy inputs, while the representation and decoding process affect both enhancement quality and inference cost.
Audio & Speech
Audio & Speech Processing
Inference Optimization
Machine Learning
Latest papersRecent research connected to this question, newest first.Masked Autoregressive Speech Enhancement with Continuous Neural Audio Codec RepresentationsThe source studies iterative decoding of masked clean-speech frames in continuous neural audio codec representations using a Conformer and the DAC codec. It compares decoding policies under the same training setup and reports a trade-off between enhancement performance and computational cost.research paper · Sep 3, 2026
Related questions
How can streaming neural audio codecs preserve speech intelligibility under zero-lookahead, low-latency constraints?How can continuous diffusion language models reduce denoising steps without sacrificing text-generation quality?How can speech enhancement adapt to mismatched deployment acoustics without labeled target audio?How can masked diffusion language models retain generation quality with only a few denoising steps?