Pith. sign in

Chinese tiny llm: Pretraining a chinese-centric large language model

4 Pith papers cite this work. Polarity classification is still indexing.

4 Pith papers citing it

citation-role summary

background 1

citation-polarity summary

years

2026 1 2025 3

roles

background 1

polarities

background 1

representative citing papers

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

cs.LG · 2026-05-29 · unverdicted · novelty 5.0

PVPO is a sample-efficient RL method that improves semantic, geometric, and physical quality in LLM LEGO assembly generation by mitigating the PhysHack failure mode where validity alone fails to ensure fidelity.

Scaling Latent Reasoning via Looped Language Models

cs.CL · 2025-10-29 · conditional · novelty 5.0

A 1.4B and a 2.6B looped (weight-tied, recurrent-depth) language model trained on 7.7T tokens match or exceed several 4B–8B transformer baselines on selected reasoning benchmarks.

citing papers explorer

Showing 4 of 4 citing papers.