Cross-modal Image Retrieval Considering Semantic Relationships with Object Information
Cross-modal Image Retrieval Considering Semantic Relationships with Object Information
复制标题
DOI:
10.1109/gcce56475.2022.10014358
复制
发表时间:
2022-10
期刊:
影响因子:
--
通讯作者:
Huaying Zhang;Rintaro Yanagi;Ren Togo;Takahiro Ogawa;M. Haseyama
中科院分区:
文献类型:
--
作者:
Huaying Zhang;Rintaro Yanagi;Ren Togo;Takahiro Ogawa;M. Haseyama
Cross-modal image retrieval methods enable users to find desired images from a text query via the embedding space. However, most existing methods do not consider the semantically similar texts and images in the embedding space. In this paper, we propose a novel cross-modal image retrieval method that can consider the relationships between semantically similar texts and images. Our method constructs an embedding space consistent with the semantic similarity by using the object information in images. Experimental results verify that our method is effective for keeping the semantically similar texts and images close in the embedding space compared to the existing methods.