GitHub F5 TTS Spanish Language Model Overview The F5 TTS model is finetuned specifically for Spanish language speech synthesis. This project aims to deliver high quality, regionally diverse speech synthesis capabilities for Spanish speakers. License This model is released under the CC0 1.0 license, which allows for free usage, modification, and distribution. Datasets The following datasets were used for training: Voxpopuli Dataset, with mainly Peninsular Spain accents Crowdsourced high quality Spanish speech data: Argentinian Spanish Chilean Spanish Colombian Spanish Peruvian Spanish Puerto Rican Spanish Venezuelan Spanish TEDx Spanish Corpus Additional sources: Crowdsourced high quality Argentinian Spanish speech data set Crowdsourced high quality Chilean Spanish speech data set Crowdsourced high quality Colombian Spanish speech data set Crowdsourced high quality Peruvian Spanish speech data set Crowdsourced high quality Puerto Rico Spanish speech data set Crowdsourced high quality Venezuelan Spanish speech data set TEDx Spanish Corpus Model Information Base Model: SWivid/F5 TTS Total Training Duration: 218 hours of audio Training Configuration: Batch Size: 3200 Max Samples: 64 Tr…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy