遇见数据集

Dzongkha Handwritten Digit Dataset

收藏
Zenodo2022-02-25 更新2026-04-07 收录
数据链接:
官方服务:

资源简介:

Dzongkha, the national language of Bhutan, has limited resources available for Natural Language Processing (NLP) tasks because the language is relatively understudied. However, there is no publicly available benchmark dataset for handwritten character identification in the Dzongkha digit script. The dataset contains 1000 images of handwritten Dzongkha digits that are captured using Google Jamboard in JPG format. The image data is assembled from a total of 100 indigenous and non-indigenous people of Bhutan irrespective of age, gender, educational background, etc. In the designed dataset, there are 10 different classes of Dzongkha digits which range from 0 to 9. The labels of these classes are: 0 (༠), 1 (༡), 2 (༢), 3 (༣), 4 (༤), 5 (༥), 6 (༦), 7 (༧), 8 (༨), 9 (༩).

创建时间:
2022-02-25
二维码
社区交流群
二维码
科研交流群
商业服务