Pith. sign in

Language models are few-shot learners

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

citation-role summary

background 1

citation-polarity summary

fields

cs.AI 1

years

2025 1

verdicts

CONDITIONAL 1

roles

background 1

polarities

background 1

representative citing papers

SCAR: Shapley Credit Assignment for More Efficient RLHF

cs.AI · 2025-05-26 · conditional · novelty 5.0

SCAR redistributes the terminal RLHF reward to tokens and spans via Shapley values, preserving the total return while improving training efficiency and final reward across three LLM alignment tasks.

citing papers explorer

Showing 1 of 1 citing paper.

  • SCAR: Shapley Credit Assignment for More Efficient RLHF cs.AI · 2025-05-26 · conditional · none · ref 6

    SCAR redistributes the terminal RLHF reward to tokens and spans via Shapley values, preserving the total return while improving training efficiency and final reward across three LLM alignment tasks.