Dictionary files used in the paper 'Combating the Curse of Multilinguality in Cross-Lingual WSD by Aligning Sparse Contextualized Word Representations'

Gabor Gabor · Zenodo (CERN European Organization for Nuclear Research) · 2022

The uploaded numpy ndarrays were used to obtain the results in the NAACL 2022 publication entitled 'Combating the Curse of Multilinguality in Cross-Lingual WSD by Aligning Sparse Contextualized Word Representations'. The four different files correspond to the 1024-by-3000 dictionary matrices that were obtained using sparse coding for the last four layers of the bert-large-cased model over the hidden representations derived from the SemCor dataset (as described in the paper).

Read the paper · More papers on PaperTik