Mention detection with LLMs in pair-programming dialogue
Cecilia Domingo, Paul Piwek, Svetlana Stoyanchev, Michel Wermelinger · 2025
We tackle the task of mention detection for pairprogramming dialogue, a setting which adds several challenges to the task due to the characteristics of natural dialogue, the dynamic environment of the dialogue task, and the domainspecific vocabulary and structures.We compare recent variants of the Llama and GPT families and explore different prompt and context engineering approaches.While aspects like hesitations and references to read-out code and variable names made the task challenging, GPT 4.1 approximated human performance when we provided few-shot examples similar to the inference text and corrected formatting errors.