AnthroSet: a Challenge Dataset for Anthropomorphic Language Detection
Dorielle Lonke, Jelke Bloem, Pia Sommerauer · 2025
This paper addresses the challenge of detecting anthropomorphic language in AI research.We introduce AnthroSet, a novel dataset of 600 manually annotated utterances covering various linguistic structures.Through the evaluation of two current approaches for anthropomorphism and atypical animacy detection, we highlight the limitations of a masked language model approach, arising from masking constraints as well as increasingly anthropomorphizing AIrelated terminology.Our findings underscore the need for more targeted methods and a robust definition of anthropomorphism.