filesystem_huggingface_5053_xgvl6k4b
收藏资源简介:
该数据集是一个推文情感语料库,包含25,000条经过标注的推文文本,用于情感极性分析研究。每条推文被标注为正面(positive)、负面(negative)或中性(neutral)三种情感类别之一。标注工作由专业标注团队完成,确保标签质量。每条记录包括推文文本、情感标签以及置信度分数,可用于训练和评估情感分类模型。数据集采用CC-BY-NC-4.0许可协议,仅限非商业用途。
This dataset is a tweet sentiment corpus containing 25,000 annotated tweet texts for sentiment polarity analysis research. Each tweet is labeled as one of three sentiment categories: positive, negative, or neutral. The annotation work was carried out by a professional annotation team to ensure label quality. Each record includes the tweet text, sentiment label, and confidence score, which can be used to train and evaluate sentiment classification models. The dataset is licensed under CC-BY-NC-4.0, for non-commercial use only.
Twitter情感语料库
数据集概览
- 许可证:CC BY-NC 4.0(非商业使用)
- 语言:英语(en)
- 用途:科学研究
数据集描述
该语料库包含25,000条推文,每条推文均带有情感极性标签,用于研究目的。情感标注由签约标注团队完成。
情感类别
数据集包含三种情感类别:
- 积极(Positive)
- 消极(Negative)
- 中立(Neutral)
数据结构
每条记录包含以下字段:
- 推文文本(tweet text):推文的原始内容
- 情感标签(sentiment label):对应的情感类别
- 置信度分数(confidence score):标注的置信度评估值




