Skip to main navigation Skip to search Skip to main content

MMMLP: Multi-modal Multilayer Perceptron for Sequential Recommendations

  • Jiahao Liang
  • , Xiangyu Zhao*
  • , Muyang Li
  • , Zijian Zhang
  • , Wanyu Wang
  • , Haochen Liu
  • , Zitao Liu
  • *Corresponding author for this work
  • City University of Hong Kong
  • University of Sydney
  • Jilin University
  • Michigan State University
  • University of Jinan

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Sequential recommendation aims to offer potentially interesting products to users by capturing their historical sequence of interacted items. Although it has facilitated extensive physical scenarios, sequential recommendation for multi-modal sequences has long been neglected. Multi-modal data that depicts a user's historical interactions exists ubiquitously, such as product pictures, textual descriptions, and interacted item sequences, providing semantic information from multiple perspectives that comprehensively describe a user's preferences. However, existing sequential recommendation methods either fail to directly handle multi-modality or suffer from high computational complexity. To address this, we propose a novel Multi-Modal Multi-Layer Perceptron (MMMLP) for maintaining multi-modal sequences for sequential recommendation. MMMLP is a purely MLP-based architecture that consists of three modules - the Feature Mixer Layer, Fusion Mixer Layer, and Prediction Layer - and has an edge on both efficacy and efficiency. Extensive experiments show that MMMLP achieves state-of-the-art performance with linear complexity. We also conduct ablating analysis to verify the contribution of each component. Furthermore, compatible experiments are devised, and the results show that the multi-modal representation learned by our proposed model generally benefits other recommendation models, emphasizing our model's ability to handle multi-modal information. We have made our code available online to ease reproducibility1.

Original languageEnglish
Title of host publicationACM Web Conference 2023 - Proceedings of the World Wide Web Conference, WWW 2023
PublisherAssociation for Computing Machinery, Inc
Pages1109-1117
Number of pages9
ISBN (Electronic)9781450394161
DOIs
StatePublished - 30 Apr 2023
Externally publishedYes
Event32nd ACM World Wide Web Conference, WWW 2023 - Austin, United States
Duration: 30 Apr 20234 May 2023

Publication series

NameACM Web Conference 2023 - Proceedings of the World Wide Web Conference, WWW 2023

Conference

Conference32nd ACM World Wide Web Conference, WWW 2023
Country/TerritoryUnited States
CityAustin
Period30/04/234/05/23

Keywords

  • Multi-modal Data
  • Multimedia
  • Sequential Recommendation

Fingerprint

Dive into the research topics of 'MMMLP: Multi-modal Multilayer Perceptron for Sequential Recommendations'. Together they form a unique fingerprint.

Cite this