Hurricanes Harvey and Irma Tweet ids
收藏资源简介:
This dataset contains the tweet ids of 35,596,281 tweets related to Hurricanes Irma and Harvey. They were collected during these events from the Twitter API using Social Feed Manager. These tweet ids are broken up into 2 collections. Each collection was collected using the POST statuses/filter method of the Twitter Stream API. The collections are: Hurricane Irma: irma_filter_tweet_ids.txt Hurricane Harvey: harvey_filter_tweet_ids.txt There is a README.txt file for each collection containing additional documentation on how it was collected. The GET statuses/lookup method supports retrieving the complete tweet for a tweet id (known as hydrating). Tools such as Twarc or Hydrator can be used to hydrate tweets. Per Twitter’s Developer Policy, tweet ids may be publicly shared for academic purposes; tweets may not. Questions about this dataset can be sent to sfm@gwu.edu. George Washington University researchers should contact us for access to the tweets.
本数据集包含35,596,281条与飓风艾尔玛(Hurricane Irma)和哈维(Hurricane Harvey)相关推文的推文ID(tweet ID)。上述推文ID是在两场飓风事件发生期间,通过社交馈送管理器(Social Feed Manager)从推特应用程序编程接口(Twitter API)采集获取的。本数据集的推文ID被划分为两个数据集子集,每个子集均通过推特流式API(Twitter Stream API)的POST statuses/filter方法采集完成。两个数据集子集分别为:飓风艾尔玛子集:irma_filter_tweet_ids.txt;飓风哈维子集:harvey_filter_tweet_ids.txt。每个数据集子集均附带一个README.txt文件,其中包含该子集采集流程的补充说明文档。GET statuses/lookup方法支持通过推文ID获取完整推文内容(该操作被称为"推文水化"(hydrating)),可使用Twarc或Hydrator等工具完成推文水化流程。根据推特开发者政策(Twitter’s Developer Policy),推文ID可出于学术目的公开共享,但完整推文内容不得公开。若对本数据集存在疑问,可发送邮件至sfm@gwu.edu。乔治华盛顿大学的研究人员若需获取完整推文内容,请联系我方。




