期刊论文详细信息
Visual Computing for Industry, Biomedicine, and Art
EM-Gaze: eye context correlation and metric learning for gaze estimation
Original Article
Pengfei Wan1  Xiaoyan Guo1  Feng Shi1  Jinchao Zhou2  Miao Wang2  Guoan Li2 
[1] Kuaishou Technology, 100085, Beijing, China;State Key Laboratory of Virtual Reality Technology and Systems, Beihang University, 100191, Beijing, China;
关键词: Computer vision;    Gaze estimation;    Metric learning;    Attention;    Multi-task learning;   
DOI  :  10.1186/s42492-023-00135-6
 received in 2022-11-18, accepted in 2023-04-15,  发布年份 2023
来源: Springer
PDF
【 摘 要 】

In recent years, deep learning techniques have been used to estimate gaze—a significant task in computer vision and human-computer interaction. Previous studies have made significant achievements in predicting 2D or 3D gazes from monocular face images. This study presents a deep neural network for 2D gaze estimation on mobile devices. It achieves state-of-the-art 2D gaze point regression error, while significantly improving gaze classification error on quadrant divisions of the display. To this end, an efficient attention-based module that correlates and fuses the left and right eye contextual features is first proposed to improve gaze point regression performance. Subsequently, through a unified perspective for gaze estimation, metric learning for gaze classification on quadrant divisions is incorporated as additional supervision. Consequently, both gaze point regression and quadrant classification performances are improved. The experiments demonstrate that the proposed method outperforms existing gaze-estimation methods on the GazeCapture and MPIIFaceGaze datasets.

【 授权许可】

CC BY   
© The Author(s) 2023

【 预 览 】
附件列表
Files Size Format View
RO202308158541854ZK.pdf 1594KB PDF download
Fig. 4 110KB Image download
Fig. 1 723KB Image download
Fig. 1 664KB Image download
41116_2023_36_Article_IEq10.gif 1KB Image download
41116_2023_36_Article_IEq16.gif 1KB Image download
41116_2023_36_Article_IEq29.gif 1KB Image download
41116_2023_36_Article_IEq35.gif 1KB Image download
41116_2023_36_Article_IEq39.gif 1KB Image download
41116_2023_36_Article_IEq40.gif 1KB Image download
MediaObjects/12888_2023_4791_MOESM2_ESM.docx 14KB Other download
41116_2023_36_Article_IEq42.gif 1KB Image download
【 图 表 】

41116_2023_36_Article_IEq42.gif

41116_2023_36_Article_IEq40.gif

41116_2023_36_Article_IEq39.gif

41116_2023_36_Article_IEq35.gif

41116_2023_36_Article_IEq29.gif

41116_2023_36_Article_IEq16.gif

41116_2023_36_Article_IEq10.gif

Fig. 1

Fig. 1

Fig. 4

【 参考文献 】
  • [1]
  • [2]
  • [3]
  • [4]
  • [5]
  • [6]
  • [7]
  • [8]
  • [9]
  • [10]
  • [11]
  • [12]
  • [13]
  • [14]
  • [15]
  • [16]
  • [17]
  • [18]
  • [19]
  • [20]
  • [21]
  • [22]
  • [23]
  • [24]
  • [25]
  • [26]
  • [27]
  • [28]
  • [29]
  • [30]
  • [31]
  • [32]
  • [33]
  • [34]
  • [35]
  • [36]
  • [37]
  • [38]
  • [39]
  • [40]
  • [41]
  • [42]
  • [43]
  • [44]
  • [45]
  • [46]
  • [47]
  • [48]
  • [49]
  • [50]
  • [51]
  • [52]
  • [53]
  文献评价指标  
  下载次数:1次 浏览次数:3次