Eye movements while viewing narrated, captioned, and silent videos.

Eye movements while viewing narrated, captioned, and silent videos.
复制标题

观看带旁白、字幕和无声视频时的眼球运动。

DOI:
10.1167/13.4.1
复制
发表时间:
2013
期刊:
影响因子:
1.8
通讯作者:
Kowler,Eileen
Kowler,Eileen
中科院分区:
医学4区
文献类型:
--
作者:
Ross,NicholasM;Kowler,Eileen

文献摘要

被引文献

相似文献

摘要:视频通常伴随着音频流或字幕的叙述,但在观看叙述视频显示时,人们对跳跃性模式知之甚少。在观看带有(a)音频解说、(b)字幕、(c)无字幕或(d)字幕和音频同时播放的视频片段时,记录眼球运动。即使在存在冗余音频流的情况下,也有相当大比例的时间(bbbb40 %)花在阅读字幕上。冗余的音频并不影响扫视阅读模式,但确实会导致跳过字幕的某些部分,并导致进入字幕区域的扫视延迟。在没有字幕的情况下,注意力被吸引到具有高密度信息的区域,如显示器的中心区域,以及具有高水平时间变化(动作和事件)的区域,而不管是否有旁白。无论是否有冗余音频,字幕都具有很强的吸引力,这就提出了一个问题,即是什么决定了如何在字幕和视频区域之间分配时间,以最大限度地减少信息损失。分配时间的策略可能基于几个因素,包括视线对任何可用文本的内在吸引力,对标题和视频中信息相对重要性的时刻印象,以及将伴随音频的视觉文本整合到单一叙事流中的驱动力。
Abstract:Abstract Videos are often accompanied by narration delivered either by an audio stream or by captions, yet little is known about saccadic patterns while viewing narrated video displays. Eye movements were recorded while viewing video clips with (a) audio narration,(b) captions,(c) no narration, or (d) concurrent captions and audio. A surprisingly large proportion of time (> 40%) was spent reading captions even in the presence of a redundant audio stream. Redundant audio did not affect the saccadic reading patterns but did lead to skipping of some portions of the captions and to delays of saccades made into the caption region. In the absence of captions, fixations were drawn to regions with a high density of information, such as the central region of the display, and to regions with high levels of temporal change (actions and events), regardless of the presence of narration. The strong attraction to captions, with or without redundant audio, raises the question of what determines how time is apportioned between captions and video regions so as to minimize information loss. The strategies of apportioning time may be based on several factors, including the inherent attraction of the line of sight to any available text, the moment by moment impressions of the relative importance of the information in the caption and the video, and the drive to integrate visual text accompanied by audio into a single narrative stream.