Revisiting Blind Photography in the Context of Teachable Object Recognizers

Revisiting Blind Photography in the Context of Teachable Object Recognizers
复制标题

在可教学物体识别器的背景下重新审视盲人摄影

DOI:
10.1145/3308561.3353799
复制
发表时间:
2019
期刊:
The 21st International ACM SIGACCESS Conference on Computers and Accessibility
影响因子:
--
通讯作者:
Kacorri, Hernisa
Kacorri, Hernisa
中科院分区:
--
文献类型:
--
作者:
Lee, Kyungjun;Hong, Jonggi;Pimento, Simone;Jarjue, Ebrima;Kacorri, Hernisa

文献摘要

参考文献

被引文献

相似文献

对于有视力障碍的人来说,摄影对于通过远视帮助和图像识别应用程序识别物体至关重要。对于可教导的对象识别器来说尤其如此,其中识别模型是根据用户的照片进行训练的。在这里,我们提出实时反馈来传达相机帧中感兴趣对象的位置。我们的音频触觉反馈由深度学习模型提供支持,该模型根据对象中心与用户手部的距离来估计对象中心位置。为了评估我们的方法,我们在实验室进行了一项用户研究,其中有视觉障碍的参与者 (N=9) 使用我们的反馈在普通和杂乱的环境中训练和测试他们的对象识别器。我们发现很少有照片不包含该物体(2% 是普通照片,8% 是杂乱照片),即使对于没有相机经验的参与者来说,识别性能也很有希望。参与者倾向于相信反馈,即使他们知道反馈可能是错误的。我们的聚类分析表明,更好的反馈与包含整个对象的照片相关。我们的结果提供了对可能降低可教学界面中的反馈和识别性能的因素的见解。
For people with visual impairments, photography is essential in identifying objects through remote sighted help and image recognition apps. This is especially the case for teachable object recognizers, where recognition models are trained on user's photos. Here, we propose real-time feedback for communicating the location of an object of interest in the camera frame. Our audio-haptic feedback is powered by a deep learning model that estimates the object center location based on its proximity to the user's hand. To evaluate our approach, we conducted a user study in the lab, where participants with visual impairments (N=9) used our feedback to train and test their object recognizer in vanilla and cluttered environments. We found that very few photos did not include the object (2% in the vanilla and 8% in the cluttered) and the recognition performance was promising even for participants with no prior camera experience. Participants tended to trust the feedback even though they know it can be wrong. Our cluster analysis indicates that better feedback is associated with photos that include the entire object. Our results provide insights into factors that can degrade feedback and recognition performance in teachable interfaces.
使用手机和基于云的视觉搜索引擎进行实时物体扫描
DOI: 10.1145/2513383.2513443
发表时间: 2013
期刊: Proceedings of the 15th International ACM SIGACCESS Conference on Computers and Accessibility
影响因子: --
作者:
Yu Zhong;Pierre Garrigues;Jeffrey P. Bigham
通讯作者: Jeffrey P. Bigham
DOI: 10.1109/cvpr.2018.00380
发表时间: 2018-02
期刊: 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition
影响因子: --
作者:
D. Gurari;Qing Li;Abigale Stangl;Anhong Guo;Chi Lin;K. Grauman;Jiebo Luo;Jeffrey P. Bigham
通讯作者: D. Gurari;Qing Li;Abigale Stangl;Anhong Guo;Chi Lin;K. Grauman;Jiebo Luo;Jeffrey P. Bigham
DOI: 10.1145/2049536.2049573
发表时间: 2011
期刊: The proceedings of the 13th international ACM SIGACCESS conference on Computers and accessibility
影响因子: --
作者:
C. Jayant;H. Ji;Samuel White;Jeffrey P. Bigham
通讯作者: Jeffrey P. Bigham
EasySnap:盲人摄影的实时音频反馈
DOI: 10.1145/1866218.1866244
发表时间: 2010
期刊: Proceedings of the SIGCHI Conference on Human Factors in Computing Systems
影响因子: --
作者:
Samuel White;H. Ji;Jeffrey P. Bigham
通讯作者: Jeffrey P. Bigham
BlindCamera:盲人摄影师的中央和黄金比例构图
DOI: 10.1145/2814464.2814472
发表时间: 2015
期刊: Proceedings of the SIGCHI Conference on Human Factors in Computing Systems
影响因子: --
作者:
Jan Balata;Z. Míkovec;Lukas Neoproud
通讯作者: Lukas Neoproud