遇见数据集

Subreddit AskScience - 1% sample from 2016 - Learning in the Wild Coding Schema

收藏
The Canadian Dataverse Repository2020-01-01 更新2026-04-17 收录
官方服务:

资源简介:

Data Collection: Data was collected using a custom web application (Communalytic, available at: https://communalytic.com/) that used Reddit’s public API (https://www.reddit.com/dev/api/). We sampled one percent of public Reddit comments posted in 2016 from AskScience. Since the dataset was collected retroactively, it does not include comments deleted by authors or moderators. Manual Coding: The sample comments were then manually coded using the 'Leaning in the Wild' Coding Schema by three independent coders, each of whom had first completed a schema tutorial training-module. Each coder (1) reviewed a submission that started a thread and was often framed as a question (see the "submissions_title" column) and then (2) assigned up to three applicable codes to the reply message (see the "text" column). The values stored under columns C1-C8 represent the number of coders who agreed on a given code.

创建时间:
2020-01-01
二维码
社区交流群
二维码
科研交流群
商业服务