Predicting a Scientific Community’s Response to an Article

Dani Yogatama, Michael Heilman, Brendan T. O’Connor, Chris Dyer, Bryan Routledge, Noah A. Smith · KiltHub Repository · 2011

We consider the problem of predicting measurable responses to scientific articles based primarily on their text content. Specifically, we consider papers in two fields (economics and computational linguistics) and make predictions about downloads and within-community citations. Our approach is based on generalized linear models, allowing interpretability; a novel extension that captures first-order temporal effects is also presented. We demonstrate that text features significantly improve accuracy of predictions over metadata features like authors, topical categories, and publication venues.

Read the paper · More papers on PaperTik