Deep gaze pooling: Inferring and visually decoding search intents from human gaze fixations

Sattara, Hosnieh and Fritz, Mario and Bulling, Andreas

(2020) Deep gaze pooling: Inferring and visually decoding search intents from human gaze fixations.

Neurocomputing, 387. pp. 369-382. ISSN 0925-2312

Full text not available from this repository.

Official URL: https://doi.org/10.1016/j.neucom.2020.01.028

Abstract

Predicting the target of visual search from human eye fixations (gaze) is a difficult problem with many applications, e.g. in human-computer interaction. While previous work has focused on predicting specific search target instances, we propose the first approach to predict categories and attributes of search intents from gaze data and to visually reconstruct plausible targets. However, state-of-the-art models for categorical recognition, in general, require large amounts of training data, which is prohibitive for gaze data. To address this challenge, we further propose a novel Gaze Pooling Layer that combines gaze information with visual representations from Deep Learning approaches. Our scheme incorporates both spatial and temporal aspects of human gaze behavior as well as the appearance of the fixated locations. We propose an experimental setup and novel dataset and demonstrate the effectiveness of our method for gaze-based search target prediction and reconstruction. We highlight several practical advantages of our approach, such as compatibility with existing architectures, no need for gaze training data, and robustness to noise from common gaze sources.

Item Type:	Article
Divisions:	Mario Fritz (MF)
Depositing User:	Anne Monzel-Busch
Date Deposited:	15 Oct 2020 15:58
Last Modified:	12 May 2021 12:57
Primary Research Area:	NRA1: Trustworthy Information Processing
URI:	https://publications.cispa.saarland/id/eprint/3249

Actions

Actions (login required)

View Item