Sparse Relational Topical Coding on multi-modal data
Sparse Relational Topical Coding on multi-modal data
复制标题
DOI:
10.1016/j.patcog.2017.08.005
复制
发表时间:
2017-12
期刊:
影响因子:
--
通讯作者:
Lingyun Song;Jun Liu;Minnan Luo;B. Qian;Kuan Yang
中科院分区:
文献类型:
--
作者:
Lingyun Song;Jun Liu;Minnan Luo;B. Qian;Kuan Yang
Multi-modal data modeling lately has been an active research area in pattern recognition community. Existing studies mainly focus on modeling the content of multi-modal documents, whilst the links amongst documents are commonly ignored. However, link information has shown being of key importance in many applications, such as document navigation, classification, and clustering. In this paper, we present a non-probabilistic formulation of Relational Topic Model (RTM), i.e., Sparse Relational Multi-Modal Topical Coding (SRMMTC), to model both multi-modal documents and the corresponding link information. SRMMTC has the following three appealing properties: i) It can effectively produce sparse latent representations via directly imposing sparsity-inducing regularizers. ii) It handles the imbalance issues on multi-modal data collections by introducing regularization parameters for positive and negative links, respectively; iii) It can be solved by an efficient coordinate descent algorithm. We also explore a generalized version of SRMMTC to find pairwise interactions amongst topics. Our methods are also capable of performing link prediction for documents, as well as the prediction of annotation words for attendant images in documents. Empirical studies on a set of benchmark datasets show that our proposed models significantly outperform many state-of-the-art methods.