遇见数据集

Comparison of objective WhatsApp data and subjective self-reports before and after data-driven personalized feedback

收藏
Zenodo2026-04-01 更新2026-05-26 收录
官方服务:

资源简介:

General information Note that this dataset partically overlaps with another dataset published earlier. However, since the data are associated with different analyses and publications, we chose to release two targeted datasets corresponding to each. The dataset contains de-identified messaging meta-data from 68 WhatsApp data donations. The data was collected from August 2022 to June 2024 in an online study using the data donation platform Dona. The participants first answered questions to their sociodemographic information, current mood and aspects of their texting behavior, such as whether they send a higher number of words per month than they receive. After the survey, the participants donated their WhatsApp data on Dona and received visualizations of their messaging behavior, such as how much they write in different chats, then are they more active, et cetera. The goal was to compare whether data-driven visualizations change self-assessments of messaging behavior toward more objective values. For more information on Dona, the associated publications and updates, please visit https://mbp-lab.github.io/dona-blog/. File description donation_table_CHB.csv - contains general information about donations including donation_id: donation identifier donor_id: the ID of the donor to distinguish the messages sent by them from those sent by contacts source: the messaging platform from which the data is donated (WhatsApp) external_id: ID used to connect messaging data with the survey data donation_table_CHB_filtered.csv - same as donation_table_CHB excluding 3 participants who did not provide the required number of chats messages_table_CHB.csv - contains the donated messages including conversation_id: chat identifier sender_id: sender identifier datetime: time of the message, UNIX time for Facebook and device time for WhatsApp word_count: word count of the messages achieved by splitting the text based on whitespace donation_id: donation identifier (also listed in donation_table_CHB.csv) messages_table_CHB_filtered.csv - same structuve as messages_table_CHB.csv excluding the three donors who did not provide the required number of chats pre_survey_CHB.xlsx -> survey responses before data donation and visual feedback post_survey_CHB.xlsx -> survey responses after data donation and visual feedback survey_coding_CHB.xlsx → contains the mapping between the column names in surveysand their meaning, including the original survey questions and response options. Request access: If you would like to request access to these files, please reach out to Dr. Olya Hakobyan at olya.hakobyan@uni-bielefeld.de. You need to satisfy these conditions in order for this request to be accepted: Individuals wishing to use the data set must hold an academic affiliation. Further to this, they have to download and fill out the End User License Agreement (EULA) and submit it to us. This dataset is intended for research purposes only.

基本信息 请注意,本数据集与此前发布的另一数据集存在部分重叠。但由于本数据集关联了不同的分析与发表成果,我们选择分别发布两套针对性数据集。 本数据集包含68份WhatsApp数据捐赠所对应的去标识化即时通讯元数据。该数据集采集于2022年8月至2024年6月期间开展的一项线上研究,研究使用了数据捐赠平台Dona。参与者首先完成了关于社会人口学信息、当前情绪以及通讯行为特征的问卷,例如每月发送的文字量是否多于接收量。问卷完成后,参与者通过Dona平台捐赠其WhatsApp数据,并可获得自身通讯行为的可视化展示,例如在不同聊天中的发言量、自身活跃度高低等。本研究的目标为验证数据驱动的可视化工具是否能将参与者对自身通讯行为的自我评估调整为更客观的数值。 如需了解关于Dona、相关发表成果及最新动态的更多信息,请访问:https://mbp-lab.github.io/dona-blog/。 文件说明 donation_table_CHB.csv:包含捐赠项目的通用信息,具体字段如下: - donation_id:捐赠项目唯一标识符 - donor_id:捐赠者ID,用于区分捐赠者本人发送的消息与联系人发送的消息 - source:捐赠数据所来源的即时通讯平台(WhatsApp) - external_id:用于关联即时通讯数据与问卷数据的ID donation_table_CHB_filtered.csv:与donation_table_CHB.csv内容一致,但剔除了未提供指定数量聊天记录的3名参与者的数据。 messages_table_CHB.csv:包含捐赠的消息数据,具体字段如下: - conversation_id:聊天会话唯一标识符 - sender_id:发送者唯一标识符 - datetime:消息发送时间,Facebook数据采用UNIX时间戳,WhatsApp数据采用设备本地时间 - word_count:消息的单词计数,通过按空白符分割文本计算得到 - donation_id:捐赠项目唯一标识符(亦存在于donation_table_CHB.csv中) messages_table_CHB_filtered.csv:与messages_table_CHB.csv结构一致,但剔除了未提供指定数量聊天记录的3名捐赠者的数据。 pre_survey_CHB.xlsx:数据捐赠与可视化反馈前的问卷应答数据 post_survey_CHB.xlsx:数据捐赠与可视化反馈后的问卷应答数据 survey_coding_CHB.xlsx:包含问卷字段名与其含义的映射表,涵盖原始问卷题目与应答选项。 数据申请方式 如需申请获取本数据集文件,请联系Olya Hakobyan博士,邮箱地址:olya.hakobyan@uni-bielefeld.de。 申请需满足以下条件方可通过审核: 1. 申请者需具备学术机构附属身份; 2. 需下载并填写最终用户许可协议(End User License Agreement,EULA)后提交至我方。 本数据集仅可用于学术研究用途。

提供机构:
Zenodo
创建时间:
2025-11-11
二维码
社区交流群
二维码
科研交流群
商业服务