← back to paper
arxiv: 2506.16712 · 2 revisions
ReasonGRM: Enhancing Generative Reward Models through Large Reasoning Models