SPAMID-PAIR
收藏NIAID Data Ecosystem2026-03-14 收录
下载链接:
https://data.mendeley.com/datasets/fj5pbdf95t
下载链接
链接失效反馈官方服务:
资源简介:
Data post-comment pairs were collected from 13 selected Indonesian public figures (artists) / public accounts with more than 15 million followers and categorized as famous artists. It was collected from Instagram using an online tool and Selenium. Two persons labeled all pair data as an expert in a total of 72874 data. The data contains Unicode text (UTF-8) and emojis scrapped in posts and comments without account profile information.
It contains several fields:
-igid: Account ID,
-comment: Comment of a post,
-post: Post from an ID,
-emoji: Whether the data contains emojis or not (1 or 0),
-spam: Whether the data is spam or not (1 or 0),
-lengthcomment: The character length of the comment,
-lengthpost: The character length of the post,
-countemojicomment: Number of emoji symbol characters in comments,
-countemojicommentuniq: Number of emoji symbol characters in comments (unique),
-countemojipost: Number of emoji symbol characters in posts,
-countemojipostuniq: Number of emoji symbol characters in the post (unique)
创建时间:
2022-09-23



