Predicting Goal-directed Human Attention Using Inverse Reinforcement Learning

被引:0
|
作者
Yang, Zhibo [1 ]
Huang, Lihan [1 ]
Chen, Yupei [1 ]
Wei, Zijun [2 ]
Ahn, Seoyoung [1 ]
Zelinsky, Gregory [1 ]
Samaras, Dimitris [1 ]
Hoai, Minh [1 ]
机构
[1] SUNY Stony Brook, Stony Brook, NY 11794 USA
[2] Adobe Inc, San Jose, CA USA
来源
2020 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR) | 2020年
基金
美国国家科学基金会;
关键词
EYE-MOVEMENTS; SEARCH; MODEL; GUIDANCE; SCENES;
D O I
10.1109/CVPR42600.2020.00027
中图分类号
TP18 [人工智能理论];
学科分类号
081104 ; 0812 ; 0835 ; 1405 ;
摘要
Human gaze behavior prediction is important for behavioral vision and for computer vision applications. Most models mainly focus on predicting free-viewing behavior using saliency maps, but do not generalize to goal-directed behavior, such as when a person searches for a visual target object. We propose the first inverse reinforcement learning (IRL) model to learn the internal reward function and policy used by humans during visual search. We modeled the viewer's internal belief states as dynamic contextual belief maps of object locations. These maps were learned and then used to predict behavioral scanpaths for multiple target categories. To train and evaluate our IRL model we created COCO-Search18, which is now the largest dataset of high-quality search fixations in existence. COCO-Search18 has 10 participants searching for each of 18 target-object categories in 6202 images, making about 300,000 goal-directed fixations. When trained and evaluated on COCO-Search18, the IRL model outperformed baseline models in predicting search fixation scanpaths, both in terms of similarity to human search behavior and search efficiency. Finally, reward maps recovered by the IRL model reveal distinctive target-dependent patterns of object prioritization, which we interpret as a learned object context.
引用
收藏
页码:190 / 199
页数:10
相关论文
共 50 条
  • [31] Longitudinal Assessment of Impairments in Goal-Directed Reinforcement Learning for Money and Food in Anorexia Nervosa
    Foerde, Karin
    Daw, Nathaniel
    Shohamy, Daphna
    Rufin, Teresa
    Walsh, B. Timothy
    Steinglass, Joanna
    BIOLOGICAL PSYCHIATRY, 2019, 85 (10) : S324 - S325
  • [32] Reward function and initial values: Better choices for accelerated goal-directed reinforcement learning
    Matignon, Laetitia
    Laurent, Guillaume J.
    Le Fort-Piat, Nadine
    ARTIFICIAL NEURAL NETWORKS - ICANN 2006, PT 1, 2006, 4131 : 840 - 849
  • [33] Modeling goal-directed attention in tone sequences using a weighted Kalman filter
    Chakrabarty, Debmalya
    Elhilali, Mounya
    2015 49TH ANNUAL CONFERENCE ON INFORMATION SCIENCES AND SYSTEMS (CISS), 2015,
  • [34] Neurocognitive Development of Goal-directed and Habitual Learning
    Hartley, Catherine
    Decker, Johannes H.
    Otto, A. Ross
    Daw, Nathaniel D.
    Casey, B. J.
    BIOLOGICAL PSYCHIATRY, 2015, 77 (09) : 294S - 294S
  • [35] Goal-directed behavior and learning of living organisms
    Umryukhin, EA
    JOURNAL OF COMPUTER AND SYSTEMS SCIENCES INTERNATIONAL, 2003, 42 (03) : 425 - 434
  • [36] Goal-directed learning and obsessive - compulsive disorder
    Gillan, Claire M.
    Robbins, Trevor W.
    PHILOSOPHICAL TRANSACTIONS OF THE ROYAL SOCIETY B-BIOLOGICAL SCIENCES, 2014, 369 (1655)
  • [37] Goal-directed learning of features and forward models
    Saeb, Sohrab
    Weber, Cornelius
    Triesch, Jochen
    NEURAL NETWORKS, 2009, 22 (5-6) : 586 - 592
  • [38] Neural bases of goal-directed implicit learning
    Rostami, Maryam
    Hosseini, S. M. Hadi
    Takahashi, Makoto
    Sugiura, Motoaki
    Kawashima, Ryuta
    NEUROIMAGE, 2009, 48 (01) : 303 - 310
  • [39] Value learning modulates goal-directed actions
    Painter, David R.
    Kritikos, Ada
    Raymond, Jane E.
    QUARTERLY JOURNAL OF EXPERIMENTAL PSYCHOLOGY, 2014, 67 (06): : 1166 - 1175
  • [40] Reward Reinforcement Creates Enduring Facilitation of Goal-directed Behavior
    Ballard, Ian C.
    Waskom, Michael
    Nix, Kerry C.
    D'Esposito, Mark
    JOURNAL OF COGNITIVE NEUROSCIENCE, 2024, 36 (12) : 2847 - 2862