Pith. sign in

Title resolution pending

6 Pith papers cite this work, alongside 1 external citations. Polarity classification is still indexing.

6 Pith papers citing it
1 external citations · external index

citation-role summary

background 1

citation-polarity summary

years

2026 5 2025 1

verdicts

UNVERDICTED 6

roles

background 1

polarities

background 1

representative citing papers

WriteSAE: Sparse Autoencoders for Recurrent State

cs.LG · 2026-05-12 · unverdicted · novelty 8.0 · 2 refs

WriteSAE introduces sparse autoencoders with rank-1 matrix atoms for recurrent state updates, allowing replacement tests that outperform deletion on 92.4% of positions and a formula predicting logit changes with R²=0.98.

A Systematic Study of Behavioral Cloning for Scientific Data Annotation

cs.HC · 2026-05-26 · unverdicted · novelty 6.0

Introduces 9 synthetic annotation tasks and benchmarks for behavioral cloning, finding hierarchical skill learning, scaling benefits, effective multi-task pretraining, and shared internal representations of task phases and mistakes.

Finding Belief Geometries with Sparse Autoencoders

cs.LG · 2026-04-03 · unverdicted · novelty 6.0

A new pipeline identifies candidate simplex geometries in Gemma-2-9B representations, with five clusters showing significant barycentric prediction advantages consistent with belief-state encoding.

Language models fail at extended rule following

cs.CL · 2026-05-03 · unverdicted · novelty 5.0 · 2 refs

LLMs fail at extended counting of repeated characters due to finite internal states, with abrupt errors persisting across model scales and inference methods.

Geometric Scaling of Bayesian Inference in LLMs

cs.LG · 2025-12-27 · unverdicted · novelty 5.0

Large language models preserve a geometric substrate in value representations that correlates with uncertainty and matches patterns from small models performing exact Bayesian inference.

citing papers explorer

Showing 6 of 6 citing papers.

  • WriteSAE: Sparse Autoencoders for Recurrent State cs.LG · 2026-05-12 · unverdicted · none · ref 58 · 2 links

    WriteSAE introduces sparse autoencoders with rank-1 matrix atoms for recurrent state updates, allowing replacement tests that outperform deletion on 92.4% of positions and a formula predicting logit changes with R²=0.98.

  • A Systematic Study of Behavioral Cloning for Scientific Data Annotation cs.HC · 2026-05-26 · unverdicted · none · ref 82

    Introduces 9 synthetic annotation tasks and benchmarks for behavioral cloning, finding hierarchical skill learning, scaling benefits, effective multi-task pretraining, and shared internal representations of task phases and mistakes.

  • Finding Belief Geometries with Sparse Autoencoders cs.LG · 2026-04-03 · unverdicted · none · ref 4

    A new pipeline identifies candidate simplex geometries in Gemma-2-9B representations, with five clusters showing significant barycentric prediction advantages consistent with belief-state encoding.

  • Language models fail at extended rule following cs.CL · 2026-05-03 · unverdicted · none · ref 24 · 2 links

    LLMs fail at extended counting of repeated characters due to finite internal states, with abrupt errors persisting across model scales and inference methods.

  • Geometric Scaling of Bayesian Inference in LLMs cs.LG · 2025-12-27 · unverdicted · none · ref 10

    Large language models preserve a geometric substrate in value representations that correlates with uncertainty and matches patterns from small models performing exact Bayesian inference.

  • Gradient Smoothing: Coupling Layer-wise Updates for Improved Optimization cs.LG · 2026-06-29 · unverdicted · none · ref 27

    Gradient Smoothing applies depth-wise smoothing to optimizer updates from base methods like Adam, yielding consistent gains in optimization and generalization on language, RL, diffusion, and vision tasks.