Code and data for research on characterization published as "Fleshing Out Models of Gender in English-Language Novels," Jonathan Cheng, Cultural Analytics 2019. The compressed character folder is a fr
Data, code, and figures accompanying the article 'Measuring the Rhetor's Craft: Computational Stylistics and the Orationes Panegyricae of Michael Psellos.' Contains: (1) a fully executed Jupyter noteb
This dataset includes derived data on a collection of ca. 2,700 books in English published between 2001-2021 and spanning twelve different genres. The data was manually collected to capture popular wr
The repository provides full data and processing / analysis pipeline for the paper 'From stage to page: language independent bootstrap measures of distinctiveness in fictional speech '