TVAE: Triplet-Based Variational Autoencoder using Metric Learning

About

Deep metric learning has been demonstrated to be highly effective in learning semantic representation and encoding information that can be used to measure data similarity, by relying on the embedding learned from metric learning. At the same time, variational autoencoder (VAE) has widely been used to approximate inference and proved to have a good performance for directed probabilistic models. However, for traditional VAE, the data label or feature information are intractable. Similarly, traditional representation learning approaches fail to represent many salient aspects of the data. In this project, we propose a novel integrated framework to learn latent embedding in VAE by incorporating deep metric learning. The features are learned by optimizing a triplet loss on the mean vectors of VAE in conjunction with standard evidence lower bound (ELBO) of VAE. This approach, which we call Triplet based Variational Autoencoder (TVAE), allows us to capture more fine-grained information in the latent embedding. Our model is tested on MNIST data set and achieves a high triplet accuracy of 95.60% while the traditional VAE (Kingma & Welling, 2013) achieves triplet accuracy of 75.08%.

Haque Ishfaq, Assaf Hoogi, Daniel Rubin• 2018

Related benchmarks

Task	Dataset	Result
Fine-Grained Sketch-Based Image Retrieval (FG-SBIR)	Chair V2 (test)	Top-1 Accuracy49.37	72
Classification	Credit	ROCAUC58	63
Fine-Grained Sketch-Based Image Retrieval (FG-SBIR)	Shoe V2 (test)	Recall@127.62	63
Regression	King (test)	R²0.44	21
Regression	News (test)	MSE0.88	17
Classification	Census	Accuracy93	13
Classification	Cabs	Accuracy66	13
Concept Composition	3DIdent 1.0 (test)	Compo25.6	9
Reconstruction	quad independent 1.0 (test)	Recon2.048	9
Reconstruction	quad dependent 1.0 (test)	Recon Score2.455	9

Showing 10 of 32 rows

Other info

Follow for update

@wizwand_team Discord