arXiv · 2609.25610
SRF-SVB: Style-Consistent Singing Voice Beautifying via Rectified Flow
Abstract
Singing voice beautifying (SVB) aims to correct pitch and rhythm of amateur singing while enhancing vocal quality, preserving lyrics and the singer's timbre. Existing methods, however, suffer from limited generation quality and efficiency, and tend to neglect the preservation of the singer's style. We propose SRF-SVB, a style-consistent model for SVB via rectified flow, which achieves high-fidelity and efficient beautification covering pitch and rhythm correction. Furthermore, we design a context-guided masked mel-spectrogram inpainting mechanism that effectively preserves the amateur singer's style, including unique timbre and expressive patterns. Experiments on both English and Chinese test sets show that SRF-SVB outperforms baseline models in most objective and subjective metrics.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Wenhui Li, Biao Dong, Liwei Hu, Jiqing Han, Yongjun He. 2026-09-22. SRF-SVB: Style-Consistent Singing Voice Beautifying via Rectified Flow. https://arxiv.org/abs/2609.25610
Cite the original work for its findings. Save a collection to share your selection of sources.