When Hyperparameters Help: Beneficial Parameter Combinations in Distributional Semantic Models
Alicia Krebs, Denis Paperno · 2016
Distributional semantic models can predict many linguistic phenomena, including word similarity, lexical ambiguity, and semantic priming, or even to pass TOEFL synonymy and analogy tests (Landauer and Dumais, 1997;Griffiths et al., 2007;Turney and Pantel, 2010).But what does it take to create a competitive distributional model?Levy et al. (2015) argue that the key to success lies in hyperparameter tuning rather than in the model's architecture.More hyperparameters trivially lead to potential performance gains, but what do they actually do to improve the models?Are individual hyperparameters' contributions independent of each other?Or are only specific parameter combinations beneficial?To answer these questions, we perform a quantitative and qualitative evaluation of major hyperparameters as identified in previous research.