TweetC19SR-Spa - Manually annotated dataset of Spanish language COVID-19 tweets containing self-reports of symptoms

Ramya Tekumalla, Luis Alberto Robles Hernandez, Juan M. Banda · Zenodo (CERN European Organization for Nuclear Research) · 2023

In this work, we release two expert curated, manually annotated datasets of COVID-19 self-reported symptoms. The first dataset contains tweets in English and the second contains tweets in Spanish, both containing around 36,500 tweets in total. These datasets were used for the Sixth and Seventh Workshop on Social Media Mining For Health (2021 and 2022)

Read the paper · More papers on PaperTik