StFX NLP at SemEval-2023 Task 1: Multimodal Encoding-based Methods for Visual Word Sense Disambiguation

Yu-Chen Wei, Milton King · 2023

SemEval-2023's Task 1, Visual Word Sense Disambiguation, a task about text semantics and visual semantics, is about selecting the best-matched image to represent a target word in a limited context.We explored several methods, including image captioning methods and CLIP-based methods, and submitted our predictions in the competition for this task.This paper will focus on the methods we used and their performance, and provide an analysis and discussion of their performance.

Read the paper · More papers on PaperTik