Supervised Learning of Universal Sentence Representations from Natural Language Inference Data

About

Many modern NLP systems rely on word embeddings, previously trained in an unsupervised manner on large corpora, as base features. Efforts to obtain embeddings for larger chunks of text, such as sentences, have however not been so successful. Several attempts at learning unsupervised representations of sentences have not reached satisfactory enough performance to be widely adopted. In this paper, we show how universal sentence representations trained using the supervised data of the Stanford Natural Language Inference datasets can consistently outperform unsupervised methods like SkipThought vectors on a wide range of transfer tasks. Much like how computer vision uses ImageNet to obtain features, which can then be transferred to other tasks, our work tends to indicate the suitability of natural language inference for transfer learning to other NLP tasks. Our encoder is publicly available.

Alexis Conneau, Douwe Kiela, Holger Schwenk, Loic Barrault, Antoine Bordes• 2017

Related benchmarks

Task	Dataset	Result
Natural Language Inference	SNLI (test)	Accuracy85.3	694
Semantic Textual Similarity	STS tasks (STS12, STS13, STS14, STS15, STS16, STS-B, SICK-R) various (test)	STS12 Score52.86	412
Semantic Textual Similarity	STS-B	Spearman's Rho (x100)68.03	156
Subjectivity Classification	Subj (test)	Accuracy93.7	152
Sentiment Classification	CR	Accuracy86.3	142
Sentiment Classification	MR (test)	Accuracy81.1	142
Text-to-Image Retrieval	MSCOCO (1K test)	R@133.9	118
Semantic Textual Similarity	STS-12	Spearman Correlation (rho)0.5286	91
Natural Language Inference	SciTail (test)	Accuracy85.1	86
Caption Retrieval	MS COCO Karpathy 1k (test)	R@142.6	62

Showing 10 of 49 rows

Other info

Code

Follow for update

@wizwand_team Discord