← back to paper
arxiv: 2607.14306 · 2 revisions
Tracing LLM Behavior to the Training Data with Empirical Next-Token Distributions