Pith. sign in

Shop-r1: Rewarding llms to simulate human behavior in online shopping via reinforcement learning

5 Pith papers cite this work. Polarity classification is still indexing.

5 Pith papers citing it

fields

cs.AI 3 cs.CL 2

years

2026 5

representative citing papers

Large Behavior Model: A Promptable Digital Twin of the Retail Customer

cs.AI · 2026-07-08 · conditional · novelty 5.0

Grounding an LLM in verbalized transaction histories via Person–Environment prompting, continued pre-training, SFT, and GRPO yields stronger retail decision simulation than frontier models, with partial cross-domain transfer.

citing papers explorer

Showing 5 of 5 citing papers.