The Edinburgh Twitter Corpus
Saša Petrović, Miles Osborne, Victor P. Lavrenko · North American Chapter of the Association for Computational Linguistics · 2010
We describe the first release of our corpus of 97 million Twitter posts. We believe that this data will prove valuable to researchers working in social media, natural language processing, large-scale data processing, and similar areas.