Cross-Modally Aligned and Temporally Gated Mixture of Experts for Multimodal Sequential Recommendation
Multimodal Sequential recommendation alleviates the semantic insufficiency and data sparsity of item-ID-based models by incorporating side information such as text and images. However, multimodal systems face the dual challenges of feature-space heterogeneity and modality-specific noise, in addition to the dynamic evol...