遇见数据集

Snapshots of a Culture War: Dataset of X/Twitter and Reddit Posts from Conservative Christian and LGBTQIA+ Youth Issues Discourse Across Four Time Periods, with Sentiment Scores

收藏
Mendeley Data2026-08-04 收录
官方服务:

资源简介:

This dataset consists of 1161 entries scraped from Reddit and X/Twitter between February 10 and March 20, 2026, including at least 100 posts from each platform representing each of four different political/cultural/historical/technological moments in the online cultural struggle between the interests of LGBTQIA+ youth and those of conservative Christian ideologies and institutions. Posts were collected from before the Trump era, during the Trump Era but before the COVID-19 lockdowns, during the COVID Era, and after the COVID Era. Each post has been assigned three sentiment codes: One by a machine learning approach created for general language, one by a machine learning approach created specifically for X/Twitter, and one by a social scientist who researches LGBTQIA+ youth issues with conservative Christianity, who has previously published research that involved coding short-format text data. To keep work on this dataset within the realm of activity that is not human subjects research, the dataset does not include posts marked as deleted or removed in their respective databases. For the sake of technical completeness, the small percentage of posts found by our algorithms that were not really part of the discourse (e.g., product advertisements, “AI slop”) were left in. This dataset may be useful to investigators of LGBTQIA+ youths’ real-lived experiences with online discourse involving conservative Christian ideologies and institutions, and to computer scientists interested in training sentiment analysis models to address politically charged issues.

创建时间:
2026-07-29
二维码
社区交流群
二维码
科研交流群
商业服务