Pith. sign in

REVIEW 1 cited by

Zero-Shot Recommendation as Language Modeling

Not yet reviewed by Pith; the record is open.

This paper has not been read by Pith yet. Machine review is queued; the pith claim, tier, and objections will appear here once it completes.

SPECIMEN: schema-true, not a live event

T0 review · schema-true

One-sentence machine reading of the paper's core claim.

pith:XXXXXXXX · record.json · timestamp

arxiv 2112.04184 v1 pith:BYJT3Y4N submitted 2021-12-08 cs.CL cs.IR

classification cs.CLcs.IR
keywords recommendationdatamatrixtextitinceptionlanguagemoviesprompt
verification ladder T0 review T1 audit T2 compute T3 formal
0 comments
abstract

Recommendation is the task of ranking items (e.g. movies or products) according to individual user needs. Current systems rely on collaborative filtering and content-based techniques, which both require structured training data. We propose a framework for recommendation with off-the-shelf pretrained language models (LM) that only used unstructured text corpora as training data. If a user $u$ liked \textit{Matrix} and \textit{Inception}, we construct a textual prompt, e.g. \textit{"Movies like Matrix, Inception, ${<}m{>}$"} to estimate the affinity between $u$ and $m$ with LM likelihood. We motivate our idea with a corpus analysis, evaluate several prompt structures, and we compare LM-based recommendation with standard matrix factorization trained on different data regimes. The code for our experiments is publicly available (https://colab.research.google.com/drive/1f1mlZ-FGaLGdo5rPzxf3vemKllbh2esT?usp=sharing).

Discussion (0). Continue with ORCID to comment.

Forward citations

Cited by 1 Pith paper

Reviewed papers in the Pith corpus that reference this work. Sorted by Pith novelty score. Full citation record

  1. Instruction-Based Fine-tuning of Open-Source LLMs for Predicting Customer Purchase Behaviors

    cs.IR 2025-01 conditional novelty 4.0 of 10

    Instruction-tuned Mistral 7B achieves modestly higher F1 than CNN/LSTM on next merchant category prediction, but the evaluation lacks significance tests and the weighted F1 is dominated by an 'Other' class.

Pith tools