Get our free extension to see links to code for papers anywhere online!Free add-on: code for papers everywhere!Free add-on: See code for papers anywhere!

Add to Chrome

Add to Firefox

Add to Edge

Title:Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Jun 26, 2022

Mengyan Zhang, Thanh Nguyen-Tang, Fangzhao Wu, Zhenyu He, Xing Xie, Cheng Soon Ong

Figure 1 for Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Figure 2 for Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Figure 3 for Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Figure 4 for Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Share this with someone who'll enjoy it:

Abstract:We consider the problem of personalised news recommendation where each user consumes news in a sequential fashion. Existing personalised news recommendation methods focus on exploiting user interests and ignores exploration in recommendation, which leads to biased feedback loops and hurt recommendation quality in the long term. We build on contextual bandits recommendation strategies which naturally address the exploitation-exploration trade-off. The main challenges are the computational efficiency for exploring the large-scale item space and utilising the deep representations with uncertainty. We propose a two-stage hierarchical topic-news deep contextual bandits framework to efficiently learn user preferences when there are many news items. We use deep learning representations for users and news, and generalise the neural upper confidence bound (UCB) policies to generalised additive UCB and bilinear UCB. Empirical results on a large-scale news recommendation dataset show that our proposed policies are efficient and outperform the baseline bandit policies.

View paper on

Share this with someone who'll enjoy it:

Title:Two-Stage Neural Contextual Bandits for Personalised News Recommendation

Paper and Code