遇见数据集

DaliaBarua/En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset

收藏
Hugging Face2025-11-13 更新2025-10-25 收录
官方服务:

资源简介:

En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset是一个面向情感分析的代码混合二元分类数据集,包含了均衡分布的正负情感类别。数据集文本涵盖了英文、孟加拉语以及不同程度的代码混合文本,适合研究多语言干扰和领域适应性,并可用于多种自然语言处理研究。

The En-Bn-Code-Mixed-Two-Class-Sentiment-Dataset is a code-mixed binary sentiment analysis dataset with a balanced distribution of positive and negative sentiment classes. It includes English, Bengali, and various degrees of code-mixed texts, suitable for researching multilingual interference and domain adaptation, and can be used in a variety of natural language processing research areas.

提供机构:
DaliaBarua
二维码
社区交流群
二维码
科研交流群
商业服务