精确定位眼睛图像中的角膜反射使用基于合成数据训练的深度学习模型 (Precise localization of corneal reflections in eye images using deep learning trained on synthetic data) - 专知论文

会员服务 ·

0

角膜 · 精确定位 · 深度学习模型 · 学习模型 · 精度 ·

2023 年 4 月 12 日

Precise localization of corneal reflections in eye images using deep learning trained on synthetic data

翻译：精确定位眼睛图像中的角膜反射使用基于合成数据训练的深度学习模型

Sean Anthony Byrne,Marcus Nyström,Virmarie Maquiling,Enkelejda Kasneci,Diederick C. Niehorster

We present a deep learning method for accurately localizing the center of a single corneal reflection (CR) in an eye image. Unlike previous approaches, we use a convolutional neural network (CNN) that was trained solely using simulated data. Using only simulated data has the benefit of completely sidestepping the time-consuming process of manual annotation that is required for supervised training on real eye images. To systematically evaluate the accuracy of our method, we first tested it on images with simulated CRs placed on different backgrounds and embedded in varying levels of noise. Second, we tested the method on high-quality videos captured from real eyes. Our method outperformed state-of-the-art algorithmic methods on real eye images with a 35% reduction in terms of spatial precision, and performed on par with state-of-the-art on simulated images in terms of spatial accuracy.We conclude that our method provides a precise method for CR center localization and provides a solution to the data availability problem which is one of the important common roadblocks in the development of deep learning models for gaze estimation. Due to the superior CR center localization and ease of application, our method has the potential to improve the accuracy and precision of CR-based eye trackers

翻译：我们提出了一种深度学习方法，用于在眼睛图像中精确定位单个角膜反射（CR）的中心。与之前的方法不同，我们使用了仅使用模拟数据训练的卷积神经网络（CNN）。仅使用模拟数据的好处是完全避开了手工注释的耗时过程，这对于在真实的眼睛图像上进行监督训练是必须的。为了系统地评估我们方法的准确性，我们首先对放置在不同背景和嵌入不同噪声水平的图像上的模拟CR进行了测试。其次，我们在从真实眼睛捕获的高品质视频上测试了该方法。我们的方法在真实眼睛图像上表现比最先进的算法方法提高了35％的精度，而在模拟图像上则在空间精度方面与最先进的技术保持相同。我们得出结论，我们的方法提供了角膜反射中心定位的精确方法，并提供了解决数据可用性问题的解决方案，这是发展基于眼动估计的深度学习模型的重要共同障碍之一。由于具有更好的CR中心定位和应用便捷性，我们的方法有潜力提高基于CR的眼动跟踪器的准确性和精度。

0

相关内容

【CVPR2022】多视图聚合的大规模三维语义分割

【CVPR2022】多视图聚合的大规模三维语义分割

专知会员服务

21+阅读 · 2022年4月20日

【AI+军事】洛马AI中心paper速读：基于深度学习的多目标跟踪、轨迹预测，Multi-Object Tracking with Deep Learning Ensemble for Unmanned Aerial System Applications

【AI+军事】洛马AI中心paper速读：基于深度学习的多目标跟踪、轨迹预测，Multi-Object Tracking with Deep Learning Ensemble for Unmanned Aerial System Applications

专知会员服务

65+阅读 · 2022年3月22日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

[ICCV 2021] 从二到一：一种带有视觉语言建模网络的新场景文本识别器

专知会员服务

17+阅读 · 2021年10月17日

最新《3D医疗图像处理》综述论文，23页pdf，3D Deep Learning on Medical Images: A Review

最新《3D医疗图像处理》综述论文，23页pdf，3D Deep Learning on Medical Images: A Review

专知会员服务

60+阅读 · 2020年7月14日

【微软研究院】IMAGEBERT: CROSS-MODAL PRE-TRAINING WITH LARGE-SCALE WEAK-SUPERVISED IMAGE-TEXT DATA

【微软研究院】IMAGEBERT: CROSS-MODAL PRE-TRAINING WITH LARGE-SCALE WEAK-SUPERVISED IMAGE-TEXT DATA

专知会员服务

43+阅读 · 2020年1月28日

【新书】使用OpenCV学习计算机视觉，深度学习CNNs和RNNs，Learn Computer Vision Using OpenCV With Deep Learning CNNs and RNNs，附163页pdf，

【新书】使用OpenCV学习计算机视觉，深度学习CNNs和RNNs，Learn Computer Vision Using OpenCV With Deep Learning CNNs and RNNs，附163页pdf，

专知会员服务

111+阅读 · 2020年1月22日

【CVPR 2019 | tutorial】自主汽车的感知、预测和大规模数据采集：Perception, Prediction, and Large Scale Data Collection for Autonomous Cars

【CVPR 2019 | tutorial】自主汽车的感知、预测和大规模数据采集：Perception, Prediction, and Large Scale Data Collection for Autonomous Cars

专知会员服务

33+阅读 · 2019年11月28日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

163+阅读 · 2019年10月12日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

【泡泡一分钟】FarSight：从户外图像中实现远距离深度估计

【泡泡一分钟】FarSight：从户外图像中实现远距离深度估计

泡泡机器人SLAM

11+阅读 · 2019年5月22日

【泡泡一分钟】三维卷积神经网络实现实时非模态三维目标检测

【泡泡一分钟】三维卷积神经网络实现实时非模态三维目标检测

泡泡机器人SLAM

12+阅读 · 2019年5月20日

【泡泡一分钟】基于运动估计的激光雷达和相机标定方法

【泡泡一分钟】基于运动估计的激光雷达和相机标定方法

泡泡机器人SLAM

25+阅读 · 2019年1月17日

【泡泡一分钟】扫描环境：用于3D点云地图中场景识别的自我中心空间描述符

【泡泡一分钟】扫描环境：用于3D点云地图中场景识别的自我中心空间描述符

泡泡机器人SLAM

22+阅读 · 2019年1月17日

【泡泡一分钟】用于评估视觉惯性里程计的TUM VI数据集

【泡泡一分钟】用于评估视觉惯性里程计的TUM VI数据集

泡泡机器人SLAM

11+阅读 · 2019年1月4日

【泡泡一分钟】基于机器人的视觉惯性里程计（IROS2018-10）

【泡泡一分钟】基于机器人的视觉惯性里程计（IROS2018-10）

泡泡机器人SLAM

13+阅读 · 2019年1月3日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【泡泡一分钟】使用深度神经网络提取局部特征的大规模图像检索算法(ICCV-2)

【泡泡一分钟】使用深度神经网络提取局部特征的大规模图像检索算法(ICCV-2)

泡泡机器人SLAM

16+阅读 · 2018年2月10日

深度学习医学图像分析文献集

深度学习医学图像分析文献集

机器学习研究会

19+阅读 · 2017年10月13日

基于微镜器件和复合传感器的高反射回转面缺陷检测新方法

国家自然科学基金

0+阅读 · 2015年12月31日

基于光学扫描全息的多图像加密原理及方法研究

国家自然科学基金

0+阅读 · 2014年12月31日

图上的偏微分方程理论及其在图像处理中的应用

国家自然科学基金

2+阅读 · 2014年12月31日

六维非自治非线性动力学系统的全局分叉和多脉冲混沌动力学的研究及应用

国家自然科学基金

0+阅读 · 2013年12月31日

复杂地形下耦合多基元的低空倾斜立体影像匹配研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于四元数的彩色视频去噪方法

国家自然科学基金

0+阅读 · 2012年12月31日

公路隧道照明察觉对比设计方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于观测图像的发音器官运动合成研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于人类视觉感知的高分辨率卫星遥感图像智能分类方法研究

国家自然科学基金

1+阅读 · 2009年12月31日

高分遥感影像中地震倒塌房屋应急提取新方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

Learning from Children: Improving Image-Caption Pretraining via Curriculum

Arxiv

1+阅读 · 2023年5月30日

Language-Conditioned Imitation Learning with Base Skill Priors under Unstructured Data

Arxiv

0+阅读 · 2023年5月30日

Towards Weakly-Supervised Hate Speech Classification Across Datasets

Arxiv

0+阅读 · 2023年5月30日

Synfeal: A Data-Driven Simulator for End-to-End Camera Localization

Arxiv

0+阅读 · 2023年5月29日

GlyphControl: Glyph Conditional Control for Visual Text Generation

Arxiv

0+阅读 · 2023年5月29日

HGT: A Hierarchical GCN-Based Transformer for Multimodal Periprosthetic Joint Infection Diagnosis Using CT Images and Text

Arxiv

0+阅读 · 2023年5月29日

Visually-augmented pretrained language models for NLP tasks without images

Arxiv

0+阅读 · 2023年5月26日

Learning with Limited Annotations: A Survey on Deep Semi-Supervised Learning for Medical Image Segmentation

Learning with Limited Annotations: A Survey on Deep Semi-Supervised Learning for Medical Image Segmentation

Arxiv

13+阅读 · 2022年7月28日

Image Segmentation Using Deep Learning: A Survey

Image Segmentation Using Deep Learning: A Survey

Arxiv

47+阅读 · 2020年1月15日

An application of cascaded 3D fully convolutional networks for medical image segmentation

Arxiv

10+阅读 · 2018年3月20日

VIP会员

文章信息

相关主题

深度学习模型

相关VIP内容

【CVPR2022】多视图聚合的大规模三维语义分割

【CVPR2022】多视图聚合的大规模三维语义分割

专知会员服务

21+阅读 · 2022年4月20日

【AI+军事】洛马AI中心paper速读：基于深度学习的多目标跟踪、轨迹预测，Multi-Object Tracking with Deep Learning Ensemble for Unmanned Aerial System Applications

【AI+军事】洛马AI中心paper速读：基于深度学习的多目标跟踪、轨迹预测，Multi-Object Tracking with Deep Learning Ensemble for Unmanned Aerial System Applications

专知会员服务

65+阅读 · 2022年3月22日

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

【干货书】深度学习合成数据，354页pdf，Synthetic Data for Deep Learning

专知会员服务

104+阅读 · 2022年2月10日

[ICCV 2021] 从二到一：一种带有视觉语言建模网络的新场景文本识别器

专知会员服务

17+阅读 · 2021年10月17日

最新《3D医疗图像处理》综述论文，23页pdf，3D Deep Learning on Medical Images: A Review

最新《3D医疗图像处理》综述论文，23页pdf，3D Deep Learning on Medical Images: A Review

专知会员服务

60+阅读 · 2020年7月14日

【微软研究院】IMAGEBERT: CROSS-MODAL PRE-TRAINING WITH LARGE-SCALE WEAK-SUPERVISED IMAGE-TEXT DATA

【微软研究院】IMAGEBERT: CROSS-MODAL PRE-TRAINING WITH LARGE-SCALE WEAK-SUPERVISED IMAGE-TEXT DATA

专知会员服务

43+阅读 · 2020年1月28日

【新书】使用OpenCV学习计算机视觉，深度学习CNNs和RNNs，Learn Computer Vision Using OpenCV With Deep Learning CNNs and RNNs，附163页pdf，

【新书】使用OpenCV学习计算机视觉，深度学习CNNs和RNNs，Learn Computer Vision Using OpenCV With Deep Learning CNNs and RNNs，附163页pdf，

专知会员服务

111+阅读 · 2020年1月22日

【CVPR 2019 | tutorial】自主汽车的感知、预测和大规模数据采集：Perception, Prediction, and Large Scale Data Collection for Autonomous Cars

【CVPR 2019 | tutorial】自主汽车的感知、预测和大规模数据采集：Perception, Prediction, and Large Scale Data Collection for Autonomous Cars

专知会员服务

33+阅读 · 2019年11月28日

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

Keras François Chollet 《Deep Learning with Python 》, 386页pdf

专知会员服务

163+阅读 · 2019年10月12日

[综述]深度学习下的场景文本检测与识别

[综述]深度学习下的场景文本检测与识别

专知会员服务

78+阅读 · 2019年10月10日

热门VIP内容

开通专知VIP会员享更多权益服务

《俄乌战争中的无人系统：新的战争方式与新兴趋势——来自前线的印象》报告

《海上自主水面船舶远程操作中心：安全可持续运行的多维度分析》

多模态大语言模型下游调优中“保持自我”的重要性

隐身自主无人水下航行器技术如何变革水下作战并重塑海军竞争

相关资讯

VCIP 2022 Call for Demos

VCIP 2022 Call for Demos

CCF多媒体专委会

1+阅读 · 2022年6月6日

【泡泡一分钟】FarSight：从户外图像中实现远距离深度估计

【泡泡一分钟】FarSight：从户外图像中实现远距离深度估计

泡泡机器人SLAM

11+阅读 · 2019年5月22日

【泡泡一分钟】三维卷积神经网络实现实时非模态三维目标检测

【泡泡一分钟】三维卷积神经网络实现实时非模态三维目标检测

泡泡机器人SLAM

12+阅读 · 2019年5月20日

【泡泡一分钟】基于运动估计的激光雷达和相机标定方法

【泡泡一分钟】基于运动估计的激光雷达和相机标定方法

泡泡机器人SLAM

25+阅读 · 2019年1月17日

【泡泡一分钟】扫描环境：用于3D点云地图中场景识别的自我中心空间描述符

【泡泡一分钟】扫描环境：用于3D点云地图中场景识别的自我中心空间描述符

泡泡机器人SLAM

22+阅读 · 2019年1月17日

【泡泡一分钟】用于评估视觉惯性里程计的TUM VI数据集

【泡泡一分钟】用于评估视觉惯性里程计的TUM VI数据集

泡泡机器人SLAM

11+阅读 · 2019年1月4日

【泡泡一分钟】基于机器人的视觉惯性里程计（IROS2018-10）

【泡泡一分钟】基于机器人的视觉惯性里程计（IROS2018-10）

泡泡机器人SLAM

13+阅读 · 2019年1月3日

Unsupervised Learning via Meta-Learning

Unsupervised Learning via Meta-Learning

CreateAMind

43+阅读 · 2019年1月3日

【泡泡一分钟】使用深度神经网络提取局部特征的大规模图像检索算法(ICCV-2)

【泡泡一分钟】使用深度神经网络提取局部特征的大规模图像检索算法(ICCV-2)

泡泡机器人SLAM

16+阅读 · 2018年2月10日

深度学习医学图像分析文献集

深度学习医学图像分析文献集

机器学习研究会

19+阅读 · 2017年10月13日

相关论文

Learning from Children: Improving Image-Caption Pretraining via Curriculum

Arxiv

1+阅读 · 2023年5月30日

Language-Conditioned Imitation Learning with Base Skill Priors under Unstructured Data

Arxiv

0+阅读 · 2023年5月30日

Towards Weakly-Supervised Hate Speech Classification Across Datasets

Arxiv

0+阅读 · 2023年5月30日

Synfeal: A Data-Driven Simulator for End-to-End Camera Localization

Arxiv

0+阅读 · 2023年5月29日

GlyphControl: Glyph Conditional Control for Visual Text Generation

Arxiv

0+阅读 · 2023年5月29日

HGT: A Hierarchical GCN-Based Transformer for Multimodal Periprosthetic Joint Infection Diagnosis Using CT Images and Text

Arxiv

0+阅读 · 2023年5月29日

Visually-augmented pretrained language models for NLP tasks without images

Arxiv

0+阅读 · 2023年5月26日

Learning with Limited Annotations: A Survey on Deep Semi-Supervised Learning for Medical Image Segmentation

Learning with Limited Annotations: A Survey on Deep Semi-Supervised Learning for Medical Image Segmentation

Arxiv

13+阅读 · 2022年7月28日

Image Segmentation Using Deep Learning: A Survey

Image Segmentation Using Deep Learning: A Survey

Arxiv

47+阅读 · 2020年1月15日

An application of cascaded 3D fully convolutional networks for medical image segmentation

Arxiv

10+阅读 · 2018年3月20日

相关基金

基于微镜器件和复合传感器的高反射回转面缺陷检测新方法

国家自然科学基金

0+阅读 · 2015年12月31日

基于光学扫描全息的多图像加密原理及方法研究

国家自然科学基金

0+阅读 · 2014年12月31日

图上的偏微分方程理论及其在图像处理中的应用

国家自然科学基金

2+阅读 · 2014年12月31日

六维非自治非线性动力学系统的全局分叉和多脉冲混沌动力学的研究及应用

国家自然科学基金

0+阅读 · 2013年12月31日

复杂地形下耦合多基元的低空倾斜立体影像匹配研究

国家自然科学基金

0+阅读 · 2013年12月31日

基于四元数的彩色视频去噪方法

国家自然科学基金

0+阅读 · 2012年12月31日

公路隧道照明察觉对比设计方法研究

国家自然科学基金

0+阅读 · 2012年12月31日

基于观测图像的发音器官运动合成研究

国家自然科学基金

0+阅读 · 2011年12月31日

基于人类视觉感知的高分辨率卫星遥感图像智能分类方法研究

国家自然科学基金

1+阅读 · 2009年12月31日

高分遥感影像中地震倒塌房屋应急提取新方法研究

国家自然科学基金

0+阅读 · 2009年12月31日

微信扫码咨询专知VIP会员