Semi-Supervised Visual Representation Learning for Fashion Compatibility

About

We consider the problem of complementary fashion prediction. Existing approaches focus on learning an embedding space where fashion items from different categories that are visually compatible are closer to each other. However, creating such labeled outfits is intensive and also not feasible to generate all possible outfit combinations, especially with large fashion catalogs. In this work, we propose a semi-supervised learning approach where we leverage large unlabeled fashion corpus to create pseudo-positive and pseudo-negative outfits on the fly during training. For each labeled outfit in a training batch, we obtain a pseudo-outfit by matching each item in the labeled outfit with unlabeled items. Additionally, we introduce consistency regularization to ensure that representation of the original images and their transformations are consistent to implicitly incorporate colour and other important attributes through self-supervision. We conduct extensive experiments on Polyvore, Polyvore-D and our newly created large-scale Fashion Outfits datasets, and show that our approach with only a fraction of labeled examples performs on-par with completely supervised methods.

Ambareesh Revanur, Vijay Kumar, Deepthi Sharma• 2021

Related benchmarks

Task	Dataset	Result
Fill-In-The-Blank	Polyvore Disjoint (test)	FITB Accuracy54.6	20
Compatibility prediction	Polyvore Standard (test)	Compatibility AUC0.89	12
Compatibility prediction	Polyvore Disjoint (test)	Comp. AUC0.84	12
Fill-In-The-Blank	Polyvore Standard (test)	Accuracy57.9	12

Showing 4 of 4 rows

Other info

Follow for update

@wizwand_team Discord