Pith. sign in

Med-rlvr: Emerging medical reasoning from a 3b base model via reinforcement learning

7 Pith papers cite this work. Polarity classification is still indexing.

7 Pith papers citing it

citation-role summary

background 1 method 1

citation-polarity summary

years

2026 6 2025 1

representative citing papers

Trading Human Curation for Synthetic Augmentation in RLVR

cs.LG · 2026-06-02 · conditional · novelty 5.0

Gated synthetic augmentations of a 10-task human base substitute for ~87 extra human RLVR tasks on aggregate held-out pass@1, with cost-adjusted trade rate ρ_cost in [1.4×, 11.6×].

citing papers explorer

Showing 7 of 7 citing papers.