Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning

03/21/2023
by   Lingzi Zhang, et al.
0

Sequential recommendation systems utilize the sequential interactions of users with items as their main supervision signals in learning users' preferences. However, existing methods usually generate unsatisfactory results due to the sparsity of user behavior data. To address this issue, we propose a novel pre-training framework, named Multimodal Sequence Mixup for Sequential Recommendation (MSM4SR), which leverages both users' sequential behaviors and items' multimodal content (text and images) for effectively recommendation. Specifically, MSM4SR tokenizes each item image into multiple textual keywords and uses the pre-trained BERT model to obtain initial textual and visual features of items, for eliminating the discrepancy between the text and image modalities. A novel backbone network, Multimodal Mixup Sequence Encoder (M^2SE), is proposed to bridge the gap between the item multimodal content and the user behavior, using a complementary sequence mixup strategy. In addition, two contrastive learning tasks are developed to assist M^2SE in learning generalized multimodal representations of the user behavior sequence. Extensive experiments on real-world datasets demonstrate that MSM4SR outperforms state-of-the-art recommendation methods. Moreover, we further verify the effectiveness of MSM4SR on other challenging tasks including cold-start and cross-domain recommendation.

READ FULL TEXT

page 1

page 3

research
05/23/2023

Text Is All You Need: Learning Language Representations for Sequential Recommendation

Sequential recommendation aims to model dynamic user behavior from histo...
research
10/23/2020

A Pre-training Strategy for Recommendation

The side information of items has been shown to be effective in building...
research
06/06/2022

ID-Agnostic User Behavior Pre-training for Sequential Recommendation

Recently, sequential recommendation has emerged as a widely studied topi...
research
07/20/2023

Language-Enhanced Session-Based Recommendation with Decoupled Contrastive Learning

Session-based recommendation techniques aim to capture dynamic user beha...
research
08/21/2018

LRMM: Learning to Recommend with Missing Modalities

Multimodal learning has shown promising performance in content-based rec...
research
09/19/2023

RUEL: Retrieval-Augmented User Representation with Edge Browser Logs for Sequential Recommendation

Online recommender systems (RS) aim to match user needs with the vast am...
research
10/27/2020

Contrastive Pre-training for Sequential Recommendation

Sequential recommendation methods play a crucial role in modern recommen...

Please sign up or login with your details

Forgot password? Click here to reset