遇见数据集

Twitter2015-Urdu

收藏
IEEE2026-04-17 收录
官方服务:

资源简介:

The Twitter2015-Urdu Dataset is a multimodal resource designed to advance Multimodal Named Entity Recognition (MNER) research in Urdu, a low-resource language. It adapts the widely used Twitter2015 English dataset with culturally grounded annotations tailored to Urdu's unique linguistic complexities. Featuring a balanced split across training, validation, and test sets, this dataset addresses the limitations of existing resources like CRULP, UCREL, and WikiDiverse, making it a critical tool for advancing Urdu-specific MNER research and multimodal applications.

提供机构:
Ahmad, Hussain
二维码
社区交流群
二维码
科研交流群
商业服务