FanChuan
收藏资源简介:
FanChuan是一个包含中英两种语言的多语种、图结构化数据集,由南洋理工大学信息科学与系统中心创建。该数据集涵盖了多个主题,共包含21,210条注释和14,755个用户的注释。数据集通过构建用户互动关系异质图来提供丰富的上下文信息。数据集的构建包括数据收集、注释和预处理三个步骤,确保了数据的高多样性、精确注释和丰富的上下文。该数据集用于模仿检测、评论情绪分类和用户情绪分类等三个关键任务,旨在解决社交媒体上模仿内容识别和分析的问题。
FanChuan is a multilingual, graph-structured dataset containing both Chinese and English content, created by the Center for Information Science and Systems at Nanyang Technological University. It covers a wide range of topics, with a total of 21,210 annotated records and annotations from 14,755 distinct users. The dataset provides rich contextual information by constructing a heterogeneous graph of user interaction relationships. The construction of this dataset includes three sequential steps: data collection, annotation, and preprocessing, which ensures high data diversity, accurate annotations, and abundant contextual information. This dataset is applied to three core tasks, namely imitation detection, comment sentiment classification, and user sentiment classification, aiming to address the issues of imitation content recognition and analysis on social media.

- 1FanChuan: A Multilingual and Graph-Structured Benchmark For Parody Detection and Analysis南洋理工大学, 信息科学与系统中心 · 2025年



