ReLoop: A Self-Correction Continual Learning Loop for Recommender Systems

About

Deep learning-based recommendation has become a widely adopted technique in various online applications. Typically, a deployed model undergoes frequent re-training to capture users' dynamic behaviors from newly collected interaction logs. However, the current model training process only acquires users' feedbacks as labels, but fail to take into account the errors made in previous recommendations. Inspired by the intuition that humans usually reflect and learn from mistakes, in this paper, we attempt to build a self-correction learning loop (dubbed ReLoop) for recommender systems. In particular, a new customized loss is employed to encourage every new model version to reduce prediction errors over the previous model version during training. Our ReLoop learning framework enables a continual self-correction process in the long run and thus is expected to obtain better performance over existing training strategies. Both offline experiments and an online A/B test have been conducted to validate the effectiveness of ReLoop.

Guohao Cai, Jieming Zhu, Quanyu Dai, Zhenhua Dong, Xiuqiang He, Ruiming Tang, Rui Zhang• 2022

Related benchmarks

Task	Dataset	Result
Recommendation	Gowalla	--	153
Recommendation	RetailRocket	NDCG @ 1024.6	44
Recommendation	Yelp	NDCG@100.111	32
Recommendation	MovieLens 20M	nDCG@1031.7	29
Top-K Recommendation	Amazon Electronics	NDCG@1013.4	23
Top-K Recommendation	Amazon Books	NDCG@100.116	23
Top-K Recommendation	Taobao	NDCG@100.219	23
Top-K Recommendation	H&M	NDCG@1026.5	23

Showing 8 of 8 rows

Other info

Follow for update

@wizwand_team Discord