Joint optimisation convex-negative matrix factorisation for multi-modal image collection summarisation based on images and tags
Joint optimisation convex-negative matrix factorisation for multi-modal image collection summarisation based on images and tags
复制标题
基于图像和标签的多模态图像采集摘要联合优化凸负矩阵分解
DOI:
10.1049/iet-cvi.2017.0568
复制
发表时间:
2019
期刊:
影响因子:
--
通讯作者:
Hongqi Wang
中科院分区:
文献类型:
--
作者:
Wenkai Zhang;Kun Fu;Xian Sun;Yuhang Zhang;Hao Sun;Hongqi Wang
Image collection summarisation aims to represent a large-scale multi-modal collection with a small subset of images and tags, helping navigate a large image dataset. Most extant methods leverage the contributions of text-to-visual summaries, ignoring the visual contribution to the textual topic. When the tags are weakly labelled, the textual topic cannot accurately reflect the visual summary. To solve this, the authors propose a novel model, joint optimisation of convex non-negative matrix factorisation, which incorporates images and tags in a beneficial way. The objective function contains visual and textual error functions, sharing the same indicator matrix, connecting different modal relations. Then, they propose an iterative algorithm to optimise the proposed model. Finally, they explore the effects of different visual feature representations (e.g. bag-of-words and deep learning) on multi-modal collection summary. Our proposed method is then compared with state-of-the-art algorithms using two multi-modal datasets (i.e. MIRFlickr and NUS-WIDE-SCENE). Experimental results demonstrate the effectiveness of their proposed approach.