Putting the pieces together

Ramanathan Subramanian, Jacopo Staiano, Kyriaki Kalimeri, Nicu Sebe, Fabio Pianesi · 2010

This paper presents a multimodal framework employing eye-gaze, head-pose and speech cues to explain observed social attention patterns in meeting scenes. We first investigate a few hypotheses concerning social attention and characterize meetings and individuals based on ground-truth data. This is followed by replication of ground-truth results through automated estimation of eye-gaze, head-pose and speech activity for each participant. Experimental results show that combining eye-gaze and head-pose estimates decreases error in social attention estimation by over 26%.

Read the paper · More papers on PaperTik