Pith. sign in

REVIEW 2 cited by

MSCRS: Multi-modal Semantic Graph Prompt Learning Framework for Conversational Recommender Systems

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2504.10921 v2 pith:3OSMRNQI submitted 2025-04-15 cs.IR

classification cs.IR
keywords semanticconversationalcontextsmulti-modalusergraphlearningprompt
verification ladder T0 review T1 audit T2 compute T3 formal

Signed reviews

No signed human review yet.

0 comments
read the original abstract

Conversational Recommender Systems (CRSs) aim to provide personalized recommendations by interacting with users through conversations. Most existing studies of CRS focus on extracting user preferences from conversational contexts. However, due to the short and sparse nature of conversational contexts, it is difficult to fully capture user preferences by conversational contexts only. We argue that multi-modal semantic information can enrich user preference expressions from diverse dimensions (e.g., a user preference for a certain movie may stem from its magnificent visual effects and compelling storyline). In this paper, we propose a multi-modal semantic graph prompt learning framework for CRS, named MSCRS. First, we extract textual and image features of items mentioned in the conversational contexts. Second, we capture higher-order semantic associations within different semantic modalities (collaborative, textual, and image) by constructing modality-specific graph structures. Finally, we propose an innovative integration of multi-modal semantic graphs with prompt learning, harnessing the power of large language models to comprehensively explore high-dimensional semantic relationships. Experimental results demonstrate that our proposed method significantly improves accuracy in item recommendation, as well as generates more natural and contextually relevant content in response generation.

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 2 Pith papers

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Beyond Whole Dialogue Modeling: Contextual Disentanglement for Conversational Recommendation

    cs.IR 2025-04 conditional novelty 5.0 of 10

    DisenCRS splits dialogue context into focus and background signals using contrastive and counterfactual losses, then adaptively selects prompts, improving movie recommendation and response generation on ReDial and INSPIRED.

  2. Multi-Type Context-Aware Conversational Recommender Systems via Mixture-of-Experts

    cs.CL 2025-04 conditional novelty 4.0 of 10

    A learned gating chair coordinates separate conversation, knowledge-graph, and review experts, and the paper reports improved movie recommendation accuracy and response diversity on ReDial and INSPIRED.

Pith tools