Pith. sign in

https://novasky- ai.github.io/posts/sky-t1

1 Pith paper cite this work. Polarity classification is still indexing.

1 Pith paper citing it

fields

cs.CL 1

years

2025 1

verdicts

CONDITIONAL 1

representative citing papers

Can A Gamer Train A Mathematical Reasoning Model?

cs.CL · 2025-06-10 · conditional · novelty 4.0

Fine-tuning Qwen2.5-Math-1.5B with LoRA and GRPO on one RTX 3080 Ti improves GSM8K accuracy from 71.65 to 73.69, matching or beating several larger base models.

citing papers explorer

Showing 1 of 1 citing paper.

  • Can A Gamer Train A Mathematical Reasoning Model? cs.CL · 2025-06-10 · conditional · none · ref 12

    Fine-tuning Qwen2.5-Math-1.5B with LoRA and GRPO on one RTX 3080 Ti improves GSM8K accuracy from 71.65 to 73.69, matching or beating several larger base models.