遇见数据集

ID-SMSA: Indonesian Stock Market Dataset for Sentiment Analysis

收藏
Mendeley Data2026-04-18 收录
官方服务:

资源简介:

The ID-SMSA Dataset is a collection of stock market-related Indonesian tweets that were collected via X (formerly known as Twitter). The dataset contains tweets in the Indonesian language, each labeled with sentiment categories: positive, negative, or neutral. A team of annotators completes the annotations using annotation guidelines that a clinical psychology specialist has reviewed. To facilitate future studies in sentiment analysis and financial market studies, other variables are also incorporated, such as the tweet's date and user engagement metrics (Quote Count, Reply Count, Retweet Count, and Favorite Count).

ID-SMSA数据集是一批与股票市场相关的印尼语推文集合,数据通过X(前身为Twitter)平台采集。该数据集收录的推文均为印尼语,每条均标注了情感类别:正面、负面或中性。标注工作由专业标注团队依据经临床心理学专家审定的标注指南完成。为便于未来开展情感分析与金融市场相关研究,数据集还纳入了其他相关变量,包括推文发布日期以及用户互动指标(引用数(Quote Count)、回复数(Reply Count)、转推数(Retweet Count)、点赞数(Favorite Count))。

创建时间:
2025-01-20
二维码
社区交流群
二维码
科研交流群
商业服务