RegionSpeak: Quick Comprehensive Spatial Descriptions of Complex Images for Blind Users
RegionSpeak: Quick Comprehensive Spatial Descriptions of Complex Images for Blind Users
复制标题
RegionSpeak:为盲人用户提供复杂图像的快速综合空间描述
DOI:
10.1145/2702123.2702437
复制
发表时间:
2015
期刊:
影响因子:
--
通讯作者:
Jeffrey P. Bigham
中科院分区:
文献类型:
--
作者:
Yu Zhong;Walter S. Lasecki;Erin L. Brady;Jeffrey P. Bigham
Blind people often seek answers to their visual questions from remote sources, however, the commonly adopted single-image, single-response model does not always guarantee enough bandwidth between users and sources. This is especially true when questions concern large sets of information, or spatial layout, e.g., where is there to sit in this area, what tools are on this work bench, or what do the buttons on this machine do? Our RegionSpeak system addresses this problem by providing an accessible way for blind users to (i) combine visual information across multiple photographs via image stitching, em (ii) quickly collect labels from the crowd for all relevant objects contained within the resulting large visual area in parallel, and (iii) then interactively explore the spatial layout of the objects that were labeled. The regions and descriptions are displayed on an accessible touchscreen interface, which allow blind users to interactively explore their spatial layout. We demonstrate that workers from Amazon Mechanical Turk are able to quickly and accurately identify relevant regions, and that asking them to describe only one region at a time results in more comprehensive descriptions of complex images. RegionSpeak can be used to explore the spatial layout of the regions identified. It also demonstrates broad potential for helping blind users to answer difficult spatial layout questions.