🤫 RE USE: Multilingual Universal Speech Enhancement Model Overview Description In universal speech enhancement, the goal is to restore the quality of diverse degraded speech while preserving fidelity , ensuring that all other factors remain unchanged, e.g., linguistic content, speaker identity, emotion, accent, and other paralinguistic attributes. Inspired by the distortion–perception trade off theory , our proposed single model achieves a good balance between these two objectives and has the following desirable properties: Robustness to diverse degradations , including additive noise, reverberation, clipping, bandwidth limitation, codec artifacts, packet loss and low quality mics . Support for multiple input sampling rates , including 8, 16, 22.05, 24, 32, 44.1, and 48 kHz. Strong language agnostic capability, enabling effective performance across different languages. This model is for research and development only. Usage Directly try our Gradio Interactive Demo by uploading your noisy audio/video !! Environment Setup 1. (For Mamba setup)Pre built Docker environments can be downloaded here to simplify Mamba setup. 2. If you need bandwidth extension: 3. Download and navigate to th…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy