MOS Dataset - Template-based Abstractive Microblog Opinion Summarisation

Iman Bilal, Bo Wang, Adam Tsakalidis, Dong Nguyen, Rob Procter, Maria Liakata · Figshare · 2022

This dataset was used in the paper 'Template-based Abstractive Microblog Opinion Summarisation' (to be published at TACL, 2022). The data is structured as follows: each file represents a cluster of tweets which contains the tweet IDs and a summary of the tweets written by journalists. The gold standard summary follows a template structure and depending on its opinion content, it contains a main story, majority opinion (if any) and/or minority opinions (if any). Additionally, we will include the abstractive model baselines we have used in the paper. For ease of use, we distinguish between opinionated/non-opinionated and training/testing/agreement sets. License: The annotations are provided under a CC-BY license, while Twitter retains the ownership and rights of the content of the tweets.

Read the paper · More papers on PaperTik