Pith. sign in

Gemini models, 2025

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.AI 1

years

2025 1

verdicts

UNVERDICTED 1

representative citing papers

Reinforcement Learning with Rubric Anchors

cs.AI · 2025-08-18 · unverdicted · novelty 6.0

Rubric-based rewards extend reinforcement learning to open-ended text generation, yielding a 30B model that outperforms a 671B model on humanities-style benchmarks.

citing papers explorer

Showing 1 of 1 citing paper.

  • Reinforcement Learning with Rubric Anchors cs.AI · 2025-08-18 · unverdicted · none · ref 5

    Rubric-based rewards extend reinforcement learning to open-ended text generation, yielding a 30B model that outperforms a 671B model on humanities-style benchmarks.