A Seqential learning on Bengali News Title Semantic Learning with Explainable AI
收藏资源简介:
This dataset is a Bangla dataset. There are total 6 columns. The data is collected from different newspaper and more importantly the dataset is raw. Total number of data is 3192. There are 3 classes in Title Category. They are: National, International and Sports. Number of data from different category, National-1227, Sports-994, International-971. We aim to find the title category form title using ML and deep learning models. As we said the data is in Bangla, after downloading the data might look corrupted. If the data is uploaded on drive the data will look okay or the microsoft excel is up to date, the data will look okay.
本数据集为孟加拉语(Bangla)数据集,共包含6列数据。数据采集自多家不同报刊,尤为关键的是,本数据集为原始未处理数据集,总样本量达3192条。标题类别共设有3个分类,分别为:国内(National)、国际(International)与体育(Sports)。各分类的样本量具体如下:国内类1227条,体育类994条,国际类971条。本研究旨在通过机器学习(Machine Learning)与深度学习模型,依据标题文本识别其所属类别。如前所述,本数据集为孟加拉语格式,下载后可能出现显示异常或损坏的情况。若将数据集上传至云端硬盘,或使用版本更新至最新的Microsoft Excel打开该数据集,即可正常显示。




