Skip to main navigation Skip to search Skip to main content

Dual-Enhanced Item Representation for Bundle Construction via Category-Wise and Cross-Modality Learning

  • Long Hai Nguyen
  • , Huy Son Nguyen
  • , Cam Van Thi Nguyen
  • , Duc Trong Le
  • , Atsuhiro Takasu
  • , Hoang Quynh Le*
  • *Corresponding author for this work

Research output: Chapter in Book/Conference proceedings/Edited volumeConference contributionScientificpeer-review

1 Downloads (Pure)

Abstract

Bundle recommender systems merely learn from existing bundles, but obtaining large-scale, high-quality bundle datasets remains a challenge, especially for platforms newly adopting bundle services. Bundle construction is the task of automatically selecting a set of compatible items to form a coherent bundle, a vital step before making recommendations on bundle-aware platforms. Groundbreaking work on bundle construction, like CLHE, has been designed solely on user-item interaction and self-attention modules to learn item/bundle representations. These techniques fall short of the standards for coherent bundles in real-world applications, where the relation among the semantic information of items should be considered more thoroughly. To address these challenges, we explicitly leverage category-wise information and employ cross-modal fusion to enhance item representations. By doing so, we propose Caro: Dual-Enhanced Item Representation for Bundle Construction via Category-Wise and Cross-Modality Learning. Caro captures the inherent relationships between items within analogous categories, improving bundle coherence. It comprises three main components: (1) cross-modality enhanced item representation, (2) category-enhanced item representation, and (3) bundle contrastive learning. Extensive experiments and detailed analyzes using multiple real-world datasets demonstrate that our method outperforms existing state-of-the-art techniques and provides valuable insight into the bundle construction problem. Notably, Caro achieves a 5-8% higher Recall@20 than the strongest baseline, underscoring its performance gains through dual category-wise and cross-modal enhancements. [...]
Original languageEnglish
Title of host publicationSIGIR-AP 2025 - Proceedings of the 2025 Annual International ACM SIGIR Conference on Research and Development in Information Retrieval in the Asia Pacific Region
Place of PublicationNew York, NY
PublisherAssociation for Computing Machinery (ACM)
Pages272-280
Number of pages9
ISBN (Electronic)9798400722189
DOIs
Publication statusPublished - 2025
Event3rd International ACM SIGIR Conference on Research and Development in Information Retrieval in the Asia Pacific Region, SIGIR-AP 2025 - Xi'an, China
Duration: 7 Dec 202510 Dec 2025
https://www.sigir-ap.org/sigir-ap-2025/

Conference

Conference3rd International ACM SIGIR Conference on Research and Development in Information Retrieval in the Asia Pacific Region, SIGIR-AP 2025
Country/TerritoryChina
CityXi'an
Period7/12/2510/12/25
Internet address

Bibliographical note

Green Open Access added to TU Delft Institutional Repository as part of the Taverne amendment. More information about this copyright law amendment can be found at https://www.openaccess.nl. Otherwise as indicated in the copyright section: the publisher is the copyright holder of this work and the author uses the Dutch legislation to make this work public.

Keywords

  • bundle construction
  • category-wise
  • cross-attention
  • multimodal recommendation

Fingerprint

Dive into the research topics of 'Dual-Enhanced Item Representation for Bundle Construction via Category-Wise and Cross-Modality Learning'. Together they form a unique fingerprint.

Cite this