Instruction-tuned Mistral 7B achieves modestly higher F1 than CNN/LSTM on next merchant category prediction, but the evaluation lacks significance tests and the weighted F1 is dominated by an 'Other' class.
Fusing Similarity Models with Markov Chains for Sparse Sequential Recommendation
1 Pith paper cite this work. Polarity classification is still indexing.
abstract
Predicting personalized sequential behavior is a key task for recommender systems. In order to predict user actions such as the next product to purchase, movie to watch, or place to visit, it is essential to take into account both long-term user preferences and sequential patterns (i.e., short-term dynamics). Matrix Factorization and Markov Chain methods have emerged as two separate but powerful paradigms for modeling the two respectively. Combining these ideas has led to unified methods that accommodate long- and short-term dynamics simultaneously by modeling pairwise user-item and item-item interactions. In spite of the success of such methods for tackling dense data, they are challenged by sparsity issues, which are prevalent in real-world datasets. In recent years, similarity-based methods have been proposed for (sequentially-unaware) item recommendation with promising results on sparse datasets. In this paper, we propose to fuse such methods with Markov Chains to make personalized sequential recommendations. We evaluate our method, Fossil, on a variety of large, real-world datasets. We show quantitatively that Fossil outperforms alternative algorithms, especially on sparse datasets, and qualitatively that it captures personalized dynamics and is able to make meaningful recommendations.
citation-role summary
citation-polarity summary
fields
cs.IR 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Instruction-Based Fine-tuning of Open-Source LLMs for Predicting Customer Purchase Behaviors
Instruction-tuned Mistral 7B achieves modestly higher F1 than CNN/LSTM on next merchant category prediction, but the evaluation lacks significance tests and the weighted F1 is dominated by an 'Other' class.