Devanagari Characters Image Dataset Dataset Summary The Devanagari Characters Image Dataset is a high resolution dataset designed to support research and experimentation in generative modeling, specifically for the Hindi script. It includes images for: Vowels (स्वर) Consonants (व्यंजन) Matra combinations (e.g., का, कि, की, कु) Hindi numerals (० ९) The dataset was created to address the limitations of existing Devanagari datasets, which often suffer from low resolution (typically 32x32 pixels) and limited font variety. This dataset provides a more robust and scalable resource for training advanced generative models such as diffusion models. But it can be used for some other problem statement as well. Use Cases Hindi character generation using diffusion models Full sentence level generation and rendering OCR pretraining and benchmarking for Devanagari Font style transfer and augmentation High resolution classification tasks Dataset Details Characters Covered : 80+ (including vowels, consonants, matra combinations, and numerals) Font Variations : 305 unique Devanagari Unicode fonts Resolution : 128x128 pixels (grayscale) Images per Character : ~305 Total Size : ~24,000 images Each cha…
We use cookies for essential functionality and analytics. You can accept or reject analytics cookies.Cookie policy