Robust Subjective Visual Property Prediction from Crowdsourced Pairwise Labels
Robust Subjective Visual Property Prediction from Crowdsourced Pairwise Labels
复制标题
根据众包成对标签进行稳健的主观视觉属性预测
DOI:
10.1109/tpami.2015.2456887
复制
发表时间:
2016-03-01
影响因子:
23.6
通讯作者:
Yao, Yuan
中科院分区:
文献类型:
--
作者:
Fu, Yanwei;Hospedales, Timothy M.;Yao, Yuan
The problem of estimating subjective visual properties from image and video has attracted increasing interest. A subjective visual property is useful either on its own (e.g. image and video interestingness) or as an intermediate representation for visual recognition (e.g. a relative attribute). Due to its ambiguous nature, annotating the value of a subjective visual property for learning a prediction model is challenging. To make the annotation more reliable, recent studies employ crowdsourcing tools to collect pairwise comparison labels. However, using crowdsourced data also introduces outliers. Existing methods rely on majority voting to prune the annotation outliers/errors. They thus require a large amount of pairwise labels to be collected. More importantly as a local outlier detection method, majority voting is ineffective in identifying outliers that can cause global ranking inconsistencies. In this paper, we propose a more principled way to identify annotation outliers by formulating the subjective visual property prediction task as a unified robust learning to rank problem, tackling both the outlier detection and learning to rank jointly. This differs from existing methods in that (1) the proposed method integrates local pairwise comparison labels together to minimise a cost that corresponds to global inconsistency of ranking order, and (2) the outlier detection and learning to rank problems are solved jointly. This not only leads to better detection of annotation outliers but also enables learning with extremely sparse annotations.