Human gaze behavior prediction is important for behavioral vision and for computer vision applications. Most models mainly focus on predicting free-viewing behavior using saliency maps, but do not generalize to goal-directed behavior, such as when a person searches for a visual target object. We propose the first inverse reinforcement learning (IRL) model to learn the internal reward function and policy used by humans during visual search. We modeled the viewer's internal belief states as dynamic contextual belief maps of object locations. These maps were learned and then used to predict behavioral scanpaths for multiple target categories. To train and evaluate our IRL model we created COCO-Search18, which is now the largest dataset of high-quality search fixations in existence. COCO-Search18 has 10 participants searching for each of 18 target-object categories in 6202 images, making about 300,000 goal-directed fixations. When trained and evaluated on COCO-Search18, the IRL model outperformed baseline models in predicting search fixation scanpaths, both in terms of similarity to human search behavior and search efficiency. Finally, reward maps recovered by the IRL model reveal distinctive target-dependent patterns of object prioritization, which we interpret as a learned object context.
机构:
Natl Inst Adv Ind Sci & Technol, Ctr Serv Res, Chiyoda Ku, Tokyo 1010021, JapanNatl Inst Adv Ind Sci & Technol, Ctr Serv Res, Chiyoda Ku, Tokyo 1010021, Japan
Xiang, Jianwen
Tian, Jing
论文数: 0引用数: 0
h-index: 0
机构:
JAIST, Grad Sch Knowledge Sci, Nomi, Ishikawa 9231292, JapanNatl Inst Adv Ind Sci & Technol, Ctr Serv Res, Chiyoda Ku, Tokyo 1010021, Japan
Tian, Jing
Mori, Akira
论文数: 0引用数: 0
h-index: 0
机构:
Natl Inst Adv Ind Sci & Technol, Ctr Serv Res, Chiyoda Ku, Tokyo 1010021, JapanNatl Inst Adv Ind Sci & Technol, Ctr Serv Res, Chiyoda Ku, Tokyo 1010021, Japan
机构:
New York Univ Shanghai, Div Arts & Sci, Shanghai, Peoples R ChinaZhejiang Univ, Coll Biomed Engn & Instrument Sci, Key Lab Biomed Engn Minist Educ, Hangzhou, Peoples R China
Li, Jialu
Tian, Xing
论文数: 0引用数: 0
h-index: 0
机构:
New York Univ Shanghai, Div Arts & Sci, Shanghai, Peoples R ChinaZhejiang Univ, Coll Biomed Engn & Instrument Sci, Key Lab Biomed Engn Minist Educ, Hangzhou, Peoples R China
Tian, Xing
Ding, Nai
论文数: 0引用数: 0
h-index: 0
机构:
Zhejiang Univ, Coll Biomed Engn & Instrument Sci, Key Lab Biomed Engn Minist Educ, Hangzhou, Peoples R China
Nanhu Brain Comp Interface Inst, Hangzhou, Peoples R ChinaZhejiang Univ, Coll Biomed Engn & Instrument Sci, Key Lab Biomed Engn Minist Educ, Hangzhou, Peoples R China
机构:
Brown Univ, Cognit, Linguist & Psychol Sci, Providence, RI 02912 USA
Williams Coll, Dept Psychol, Williamstown, MA 01267 USABrown Univ, Cognit, Linguist & Psychol Sci, Providence, RI 02912 USA
机构:
Signant Hlth, San Diego, CA USAUniv Maryland, Sch Med, Dept Psychiat, MPRC, Baltimore, MD 21201 USA
Schwartz, E. K.
Frank, M. J.
论文数: 0引用数: 0
h-index: 0
机构:
Brown Univ, Dept Cognit Linguist & Psychol Sci, Providence, RI 02912 USA
Brown Univ, Dept Psychiat, Providence, RI 02912 USA
Brown Univ, Brown Inst Brain Sci, Providence, RI 02912 USAUniv Maryland, Sch Med, Dept Psychiat, MPRC, Baltimore, MD 21201 USA
Frank, M. J.
Brown, E. C.
论文数: 0引用数: 0
h-index: 0
机构:
Arden Univ, Sch Hlth & Care Management, Berlin, GermanyUniv Maryland, Sch Med, Dept Psychiat, MPRC, Baltimore, MD 21201 USA