Uddessho
收藏资源简介:
Uddessho数据集由艾哈迈德·沙希德·苏赫拉瓦尔迪大学创建,专门用于低资源孟加拉语中的多模态作者意图分类。该数据集包含3048条从社交媒体平台(如Facebook、X和Instagram)收集的帖子,涵盖六个类别:信息性、倡导性、推广性、展示性、表达性和争议性。数据集的创建过程包括手动收集和标注,确保了数据的质量和一致性。该数据集主要用于解决在低资源语言环境中,通过结合文本和图像信息来准确分类作者意图的问题。
The Uddessho dataset was developed by Ahmed Shahid Suhrawardy University, specifically tailored for multimodal author intent classification in low-resource Bengali. It contains 3048 posts collected from social media platforms including Facebook, X, and Instagram, covering six categories: informative, advocacy, promotional, demonstrative, expressive, and controversial. The dataset construction process includes manual collection and annotation, which ensures the quality and consistency of the data. This dataset is primarily used to address the challenge of accurate author intent classification by combining textual and visual information in low-resource language environments.




