1. Introduction Introducing a novel accent database "IndicAccentDB" which satisfies the below requirements: Gender balance: The speech database should be a collection of a wide range of speakers balancing both the male and female speakers to display the characteristics of the speakers speech. Phonetically balanced uniform content: To make the classification task simpler and models to distinguish the speakers, we considered building the IndicAccentDB with uniform content, a collection of speech recordings for the Harvard sentences. These sentences gather intrinsic information by combining different phonemes and grammatically focused vocabulary. These sentences are appropriately expressing accents in sentence level discourse. You can access the Harvard sentences (sample shown below) dataset here: Harvard Sentences recited by the speakers in the recordings. The juice of lemons makes fine punch. The fish twisted and turned on the bent hook. IndicAccentDB contains speech recordings in six non native English accents of Gujarati, Hindi, Kannada, Malayalam, Tamil, and Telugu. We collected six non native accents from volunteers who had strong non native English accents and were well versed…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy