Kraken: Memory-Efficient Continual Learning For Large-Scale Real-Time Recommendations

Minhui Xie,Kai Ren,Youyou Lu,Guangxu Yang,Qingxing Xu,Bihai Wu,Jiazhen Lin,Hongbo Ao,Wanhong Xu,Jiwu Shu

SC（2020）

引用 16|浏览36

暂无评分

摘要

Modern recommendation systems in industry often use deep learning (DL) models that achieve better model accuracy with more data and model parameters. However, current open-source DL frameworks, such as TensorFiow and PyTorch, show relatively low scalability on training recommendation models with terabytes of parameters. To efficiently learn large-scale recommendation models from data streams that generate hundreds of terabytes training data daily, we introduce a continual learning system called Kraken. Kraken contains a special parameter server implementation that dynamically adapts to the rapidly changing set of sparse features for the continual training and serving of recommendation models. Kraken provides a sparsity-aware training system that uses different learning optimizers for dense and sparse parameters to reduce memory overhead. Extensive experiments using real-world datasels confirm the effectiveness and scalability of Kraken. Kraken can benefit the accuracy of recommendation tasks with the same memory resources, or trisect the memory usage while keeping model performance.

查看译文

关键词

Systems for Machine Learning,Continual Learning,Recommendation System

AI 理解论文

溯源树

样例

生成溯源树，研究论文发展脉络

Chat Paper

正在生成论文摘要