Wukong: Towards a Scaling Law for Large-Scale Recommendation

About

Scaling laws play an instrumental role in the sustainable improvement in model quality. Unfortunately, recommendation models to date do not exhibit such laws similar to those observed in the domain of large language models, due to the inefficiencies of their upscaling mechanisms. This limitation poses significant challenges in adapting these models to increasingly more complex real-world datasets. In this paper, we propose an effective network architecture based purely on stacked factorization machines, and a synergistic upscaling strategy, collectively dubbed Wukong, to establish a scaling law in the domain of recommendation. Wukong's unique design makes it possible to capture diverse, any-order of interactions simply through taller and wider layers. We conducted extensive evaluations on six public datasets, and our results demonstrate that Wukong consistently outperforms state-of-the-art models quality-wise. Further, we assessed Wukong's scalability on an internal, large-scale dataset. The results show that Wukong retains its superiority in quality over state-of-the-art models, while holding the scaling law across two orders of magnitude in model complexity, extending beyond 100 GFLOP/example, where prior arts fall short.

Buyun Zhang, Liang Luo, Yuxin Chen, Jade Nie, Xi Liu, Daifeng Guo, Yanli Zhao, Shen Li, Yuchen Hao, Yantao Yao, Guna Lakshminarayanan, Ellie Dingqiao Wen, Jongsoo Park, Maxim Naumov, Wenlin Chen• 2024

Related benchmarks

Task	Dataset	Result
CTR Prediction	Criteo	AUC0.8073	309
Click-Through Rate Prediction	Avazu (test)	AUC0.7893	207
CTR Prediction	Avazu	AUC77.56	171
Click-Through Rate Prediction	Criteo (test)	AUC0.8013	57
Click prediction	KuaiVideos (test)	AUC0.8842	41
CTR Prediction	TaobaoAds	AUC0.6388	41
CTR Prediction	Industrial	AUC83.34	33
CTR Prediction	KuaiVideo	GAUC0.6629	27
CTR Prediction	AMAZON	AUC0.8663	26
Follow Prediction	KuaiVideo (test)	AUC79.47	23

Showing 10 of 35 rows

Other info

Follow for update

@wizwand_team Discord