Pith. sign in

Robust preference optimization through reward model distillation.arXiv preprint arXiv:2405.19316

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

citation-role summary

background 1

citation-polarity summary

years

2026 2 2025 2

roles

background 1

polarities

support 1

representative citing papers

Generating Place-Based Compromises Between Two Points of View

cs.CL · 2026-04-27 · unverdicted · novelty 5.0

Empathic similarity feedback in prompts generates more acceptable compromises than chain-of-thought, and margin-based training on the resulting data lets smaller models produce them without ongoing empathy estimation.

LLM Harms: A Taxonomy and Discussion

cs.CY · 2025-12-05 · reject · novelty 3.0

Proposes a five-bucket taxonomy of LLM harms and calls for dynamic auditing, but the systematic review behind it is not reproducible and contains mismatched citations.

citing papers explorer

Showing 4 of 4 citing papers.