Pith. sign in

Varshney, Mohit Bansal, Sanmi Koyejo, and Yang Liu

8 Pith papers cite this work. Polarity classification is still indexing.

8 Pith papers citing it

representative citing papers

Model Unlearning Objectives Vary for Distinct Language Functions

cs.CL · 2026-05-26 · unverdicted · novelty 6.0

Unlearning objectives should be tailored to distinct language functions, with a meta-learned RMU variant for dangerous knowledge and a multi-layer probe objective for toxicity, yielding strong results on four 7-8B models.

EEPO: Exploration-Enhanced Policy Optimization via Sample-Then-Forget

cs.CL · 2025-10-07 · unverdicted · novelty 5.0

EEPO uses sample-then-forget rollouts with adaptive unlearning to boost exploration in RLVR, delivering relative gains of 24.3% on Qwen2.5-3B, 33.0% on Llama3.2-3B-Instruct, and 10.4% on Qwen3-8B-Base over GRPO across five reasoning benchmarks.

citing papers explorer

Showing 8 of 8 citing papers.