Video Recommendation with Multi-gate Mixture of Experts Soft Actor Critic

Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval Pub Date : 2020-07-25 DOI:10.1145/3397271.3401238

Dingcheng Li, Xu Li, Jun Wang, P. Li

引用次数: 20

Abstract

In this paper, we propose a reinforcement learning based large scale multi-objective ranking system for optimizing short-video recommendation on an industrial video sharing platform. Multiple competing ranking objective and implicit selection bias in user feedback are the main challenges in real-world platform. In order to address those challenges, we integrate multi-gate mixture of experts and soft actor critic into the ranking system. We demonstrated that our proposed framework can greatly reduce the loss function compared with systems only based on single strategies.

查看原文本刊更多论文

视频推荐与多门混合专家软演员评论家

本文提出了一种基于强化学习的大规模多目标排名系统，用于优化工业视频分享平台上的短视频推荐。用户反馈中的多重竞争排名目标和隐式选择偏差是现实平台中的主要挑战。为了解决这些挑战，我们将专家和软演员评论家的多门混合集成到排名系统中。我们证明，与仅基于单一策略的系统相比，我们提出的框架可以大大减少损失函数。

本文章由计算机程序翻译，如有差异，请以英文原文为准。

求助全文

约1分钟内获得全文求助全文

来源期刊

Proceedings of the 43rd International ACM SIGIR Conference on Research and Development in Information Retrieval

自引率

0.00%

发文量