Semi-Supervised Visual Representation Learning for Fashion Compatibility

被引：13

作者：

Revanur, Ambareesh ^{[1
]}

Kumar, Vijay ^{[2
]}

Sharma, Deepthi ^{[2
]}

机构：

[1] Carnegie Mellon Univ, Pittsburgh, PA 15213 USA

[2] Walmart Global Tech Bangalore, Bangalore, Karnataka, India

来源：

15TH ACM CONFERENCE ON RECOMMENDER SYSTEMS (RECSYS 2021) | 2021年

关键词：

fashion compatibility; semi-supervised learning; self-supervision; product recommendation;

D O I：

10.1145/3460231.3474233

中图分类号：

TP18 [人工智能理论];

学科分类号：

081104 ; 0812 ; 0835 ; 1405 ;

摘要：

We consider the problem of complementary fashion prediction. Existing approaches focus on learning an embedding space where fashion items from different categories that are visually compatible are closer to each other. However, creating such labeled outfits is intensive and also not feasible to generate all possible outfit combinations, especially with large fashion catalogs. In this work, we propose a semi-supervised learning approach where we leverage large unlabeled fashion corpus to create pseudo positive and negative outfits on the fly during training. For each labeled outfit in a training batch, we obtain a pseudo-outfit by matching each item in the labeled outfit with unlabeled items. Additionally, we introduce consistency regularization to ensure that representation of the original images and their transformations are consistent to implicitly incorporate colour and other important attributes through self-supervision. We conduct extensive experiments on Polyvore, Polyvore-D and our newly created large-scale Fashion Outfits datasets, and show that our approach with only a fraction of labeled examples performs on-par with completely supervised methods.

引用

页码：463 / 472

页数：10

共 46 条

[1]

Agrawal R., 1994, VLDB 94, P487

[2]

Bengio Y., 2005, Advances in Neural Information Processing Systems (NeurIPS)

[3]

Berthelot D, 2019, ADV NEUR IN, V32

[4] Deep Clustering for Unsupervised Learning of Visual Features [J].

Caron, Mathilde ;

Bojanowski, Piotr ;

Joulin, Armand ;

Douze, Matthijs .

COMPUTER VISION - ECCV 2018, PT XIV, 2018, 11218 :139-156

[5]

Chen T, 2020, PR MACH LEARN RES, V119

[6] Knowledge-guided Deep Reinforcement Learning for Interactive Recommendation [J].

Chen, Xiaocong ;

Huang, Chaoran ;

Yao, Lina ;

Wang, Xianzhi ;

Liu, Wei ;

Zhang, Wenjie .

2020 INTERNATIONAL JOINT CONFERENCE ON NEURAL NETWORKS (IJCNN), 2020,

[7] AutoAugment: Learning Augmentation Strategies from Data [J].

Cubuk, Ekin D. ;

Zoph, Barret ;

Mane, Dandelion ;

Vasudevan, Vijay ;

Le, Quoc V. .

2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, :113-123

[8] Context-Aware Visual Compatibility Prediction [J].

Cucurull, Guillem ;

Taslakian, Perouz ;

Vazquez, David .

2019 IEEE/CVF CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION (CVPR 2019), 2019, :12609-12618

[9] Learning Large-Scale Automatic Image Colorization [J].

Deshpande, Aditya ;

Rock, Jason ;

Forsyth, David .

2015 IEEE INTERNATIONAL CONFERENCE ON COMPUTER VISION (ICCV), 2015, :567-575

[10]

Duan Jiali, 2019, NEURAL INFORM PROCES

← 1 2 3 4 5 →