← 返回论文检索
ACM Multimedia 2025Engagement: Multimedia Search and Recommendation

DeCoRec: Decoupled Collaborative Refinement for Multi-Modal Sequential Recommendations

Zhaoqi Chen, Wanni Xu, Yunfeng Zhang, Yawei Hou, Zhenyu Wen, Cong Wang 0006

PDF 由论文原始站点提供,PaperCompass 不保存论文文件。DOI 10.1145/3746027.3755475 ↗

摘要

While multi-modal features offer rich semantic signals to enhance sequential recommendation systems, their integration with ID-based embeddings remains challenging. Conventional fusion strategies often degrade performance despite the semantic potential of multimodal data. Through empirical analysis, we identify asymmetric convergence dynamics between rapidly adapting ID embeddings and slowly evolving modality representations as the fundamental barrier. To address this, we propose DeCoRec, a novel framework to decouple ID and modality optimization trajectories to prevent gradient interference. To further reconcile ID and multi-modal data, we introduce modality-aware interest clustering and cross-modal contrastive learning to align semantic neighborhoods with behavioral patterns. Extensive experiments demonstrate 5-7% improvements in NDCG/HiT metrics against the existing schemes and particular robustness in cold-start scenarios. The code is available: https://github.com/KIKIENAO/decorec