UniverSR General Audio (Flagship) Vocoder free audio super resolution model that upsamples 8/12/16/24 kHz → 48 kHz audio using flow matching in the complex STFT domain. Trained on speech, music, and sound effects. This is the recommended model for general use. For speech only evaluation (e.g. VCTK benchmark), see universr speech. Paper : arXiv:2510.00771 Demo : woongzip1.github.io/universr demo Code : github.com/woongzip1/UniverSR Usage Citation
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy