StFX NLP at SemEval-2023 Task 1: Multimodal Encoding-based Methods for Visual Word Sense Disambiguation
Yu-Chen Wei, Milton King · 2023
SemEval-2023's Task 1, Visual Word Sense Disambiguation, a task about text semantics and visual semantics, is about selecting the best-matched image to represent a target word in a limited context.We explored several methods, including image captioning methods and CLIP-based methods, and submitted our predictions in the competition for this task.This paper will focus on the methods we used and their performance, and provide an analysis and discussion of their performance.