跳到主要导航 跳到搜索 跳到主要内容

Gaze Target Prediction with the Understanding of 3D Scenes

  • Leru Gao*
  • , Fengxi Sun
  • , Yue Liu
  • *此作品的通讯作者
  • Beijing Institute of Technology

科研成果: 书/报告/会议事项章节会议稿件同行评审

摘要

The goal of the gaze target prediction is to determine the location where the person is focusing and the probability of the gaze falls outside the image. Although prior works have addressed this task by regressing heatmaps centered on the gaze location, they typically fail to incorporate the scene's semantic information. In this work, we first generate 3D point cloud of the given image based on depth estimation and camera intrinsics. Then we combine the point cloud and estimated 3D gaze vector to generate the 3D field of view (FoV) heatmap. Scene contextual cues are finally merged to get the output heatmap. Our method achieves competitive results on the ChildPlay, GazeFollow, and VideoAttentionTarget datasets.

源语言英语
主期刊名Image and Graphics Technologies and Applications - 20th Chinese Conference, IGTA 2025, Revised Selected Papers
编辑Yongtian Wang, Yi Chen
出版商Springer Science and Business Media Deutschland GmbH
129-143
页数15
ISBN(印刷版)9789819549658
DOI
出版状态已出版 - 2026
已对外发布
活动20th Chinese Conference on Image and Graphics Technologies and Applications, IGTA 2025 - Beijing, 中国
期限: 9 8月 202510 8月 2025

丛书

姓名Communications in Computer and Information Science
2800 CCIS
ISSN(印刷版)1865-0929
ISSN(电子版)1865-0937

会议

会议20th Chinese Conference on Image and Graphics Technologies and Applications, IGTA 2025
国家/地区中国
Beijing
时期9/08/2510/08/25

学术指纹

探究 'Gaze Target Prediction with the Understanding of 3D Scenes' 的科研主题。它们共同构成独一无二的学术指纹。

引用此