SPAMID-PAIR
收藏DataCite Commons2025-04-01 更新2025-04-16 收录
下载链接:
https://data.mendeley.com/datasets/fj5pbdf95t
下载链接
链接失效反馈官方服务:
资源简介:
Data post-comment pairs were collected from 13 selected Indonesian public figures (artists) / public accounts with more than 15 million followers and categorized as famous artists. It was collected from Instagram using an online tool and Selenium. Two persons labeled all pair data as an expert in a total of 72874 data. The data contains Unicode text (UTF-8) and emojis scrapped in posts and comments without account profile information.
It contains several fields:
-igid: Account ID,
-comment: Comment of a post,
-post: Post from an ID,
-emoji: Whether the data contains emojis or not (1 or 0),
-spam: Whether the data is spam or not (1 or 0),
-lengthcomment: The character length of the comment,
-lengthpost: The character length of the post,
-countemojicomment: Number of emoji symbol characters in comments,
-countemojicommentuniq: Number of emoji symbol characters in comments (unique),
-countemojipost: Number of emoji symbol characters in posts,
-countemojipostuniq: Number of emoji symbol characters in the post (unique)
提供机构:
Mendeley
创建时间:
2022-09-23



