Replication Data for: Predicting and Interpolating State-level Polls using Twitter Textual Data

NIAID Data Ecosystem2026-03-10 收录

下载链接：

https://doi.org/10.7910/DVN/RJAUNW

下载链接

链接失效反馈

官方服务：

资源简介：

Spatially or temporally dense polling remains both difficult and expensive using existing survey methods. In response, there have been increasing efforts to approximate various survey measures using social media, but most of these approaches remain methodologically flawed. To remedy these flaws, this paper combines 1200 state-level polls during the 2012 presidential campaign with over 100 million state-located political Tweets; models the polls as a function of the Twitter text using a new linear regularization feature-selection method; and shows via out-of-sample testing that when properly modeled, the Twitter-based measures track and to some degree predict opinion polls, and can be extended to unpolled states and potentially sub-state regions and sub-day timescales. An examination of the most predictive textual features reveals the topics and events associated with opinion shifts, sheds light on more general theories of partisan difference in attention and information processing, and may be of use for real-time campaign strategy.

创建时间：

2017-07-01

5,000+

优质数据集

54 个

任务类型

进入经典数据集