Our new X account is live! Follow @wizwand_team for updates
WorkDL logo mark

Semi-Supervised Visual Representation Learning for Fashion Compatibility

About

We consider the problem of complementary fashion prediction. Existing approaches focus on learning an embedding space where fashion items from different categories that are visually compatible are closer to each other. However, creating such labeled outfits is intensive and also not feasible to generate all possible outfit combinations, especially with large fashion catalogs. In this work, we propose a semi-supervised learning approach where we leverage large unlabeled fashion corpus to create pseudo-positive and pseudo-negative outfits on the fly during training. For each labeled outfit in a training batch, we obtain a pseudo-outfit by matching each item in the labeled outfit with unlabeled items. Additionally, we introduce consistency regularization to ensure that representation of the original images and their transformations are consistent to implicitly incorporate colour and other important attributes through self-supervision. We conduct extensive experiments on Polyvore, Polyvore-D and our newly created large-scale Fashion Outfits datasets, and show that our approach with only a fraction of labeled examples performs on-par with completely supervised methods.

Ambareesh Revanur, Vijay Kumar, Deepthi Sharma• 2021

Related benchmarks

TaskDatasetResultRank
Fill-In-The-BlankPolyvore Disjoint (test)
FITB Accuracy54.6
20
Compatibility predictionPolyvore Standard (test)
Compatibility AUC0.89
12
Compatibility predictionPolyvore Disjoint (test)
Comp. AUC0.84
12
Fill-In-The-BlankPolyvore Standard (test)
Accuracy57.9
12
Showing 4 of 4 rows

Other info

Follow for update