MOS Dataset - Template-based Abstractive Microblog Opinion Summarisation
收藏资源简介:
This dataset was used in the paper 'Template-based Abstractive Microblog Opinion Summarisation' (to be published at TACL, 2022). The data is structured as follows: each file represents a cluster of tweets which contains the tweet IDs and a summary of the tweets written by journalists. The gold standard summary follows a template structure and depending on its opinion content, it contains a main story, majority opinion (if any) and/or minority opinions (if any). Additionally, we will include the abstractive model baselines we have used in the paper. For ease of use, we distinguish between opinionated/non-opinionated and training/testing/agreement sets. Due to the recent changes in the availability of the Twitter / X academic API, please reach out to iman.bilal@warwick.ac.uk if you consider using the dataset. License: The annotations are provided under a CC-BY license, while Twitter retains the ownership and rights of the content of the tweets.



