Associating video frames with text
Associating video frames with text
复制标题
将视频帧与文本关联
DOI:
--
复制
发表时间:
2003
期刊:
影响因子:
--
通讯作者:
H. Wactlar
中科院分区:
文献类型:
--
作者:
Pinar Duygulu;H. Wactlar
In this study, integration of visual and textual data is proposed to solve the correspondence problem between video frames and associated text in order to annotate video frames with more reliable labels and descriptions. Visual features extracted from video frames are linked to text that is obtained from the audio transcripts using joint statistics. The results show that using this approach it is possible to have better annotations for video frames that can be later used to improve the performance of text based queries. The proposed approach will be integrated into the Informedia Digital Video Library Project at Carnegie Mellon University.