遇见数据集

test

收藏
Mendeley Data2026-04-18 收录
官方服务:

资源简介:

Gujarati is the formalized communication language in the state of Gujarat and united territory of Dadra and Nagar Haveli and Div and Daman in India predominantly spoken by the Gujarati. Gujarati language contained a wealthy set of characters that includes vowels, consonants, digits, various signs. This dataset contains 75,000 grayscale isolated handwritten character images with the size of 28 X 28 pixels. This dataset contains sample images for 34 consonants, 12 vowels, 12 vowel signs and 5 various signs. This dataset could be used for the Gujarati Handwritten Character recognition in the field of Natural Language Processing (NLP) and Deep Learning .

古吉拉特语(Gujarati)是印度古吉拉特邦以及达德拉-纳加尔哈维利和达曼-第乌联邦属地的官方交流语言,主要使用者为古吉拉特族民众。古吉拉特语拥有一套丰富的字符集,涵盖元音、辅音、数字及各类符号。本数据集包含75000张尺寸为28×28像素的孤立手写灰度字符图像,涵盖34个辅音、12个元音、12个元音符号以及5类通用符号的样本图像。该数据集可应用于自然语言处理(Natural Language Processing,NLP)与深度学习领域的古吉拉特语手写字符识别任务。

创建时间:
2022-10-18
二维码
社区交流群
二维码
科研交流群
商业服务