Pith. sign in

Paper Citation Record · LEDGER

Tuning Language Models by Proxy

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:2401.08565.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2401.08565 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:37:58.926261Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T12:33:45.021126Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 4a8f6943-03bf-4117-baa5-36fa9e476602 · inbound

A Roadmap to Pluralistic Alignment cites this paper.

A Roadmap to Pluralistic Alignment Tuning Language Models by Proxy

Reference 279

Resolution
verified exact
arxiv_id, observed 2026-05-16T14:37:53.483023Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-16T14:37:53.279275Z digest=sha256:c16a3865f3dcc56f56699f10228d3a2b477970ee041e3e6ca6c5d737585c42ac

Observation e32aa794-f5d7-4a2e-92c8-0823552377f8 · inbound

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models cites this paper.

Neural Incompatibility: The Unbridgeable Gap of Cross-Scale Parametric Knowledge Transfer in Large Language Models Tuning Language Models by Proxy

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T15:37:58.926261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T15:37:58.926261Z digest=sha256:a086701bfb3f33318fdd02015930ffb05825abcd3683f7beb8d4181eb37f35d8

Observation fcb289a6-a2f4-4b35-b02e-0cc2d0a45638 · inbound

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations cites this paper.

Scalable, Symbiotic, AI and Non-AI Agent Based Parallel Discrete Event Simulations Tuning Language Models by Proxy

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T13:07:49.678073Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:07:49.678073Z digest=sha256:2f74ed377d7b642a70aa7ebad62e6bff15ceb572b2cb414108a4ee9de4511678

Observation 48185574-370d-49dd-8c9f-6ea8dab5867c · inbound

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training cites this paper.

Logit Arithmetic Elicits Long Reasoning Capabilities Without Training Tuning Language Models by Proxy

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:58.459933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:45:58.459933Z digest=sha256:54db9d2d1466de4ba93f25e4b3ca947f1e9c669c8f6d3a762b6f59b57dda648b

Observation a3f7c747-c499-49ab-a392-045e562f2926 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models Tuning Language Models by Proxy

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:43.942431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:43.942431Z digest=sha256:8a7b27078f485b15316311530bd936f16632a95ed52f27da3bd9f81b5107f59f

Observation 0447a07e-40a2-4891-b1f9-d0430be7c590 · inbound

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment cites this paper.

The Master Key Hypothesis: Unlocking Cross-Model Capability Transfer via Linear Subspace Alignment Tuning Language Models by Proxy

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:15:56.453878Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T18:36:44.401045Z digest=sha256:382362107e6603a3cdc3da051def90c9fc597798592cee08419165751aa4e49c

Observation 6b425db2-3532-46b7-93d3-477c1b3e4efb · inbound

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration cites this paper.

Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration Tuning Language Models by Proxy

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T09:16:04.584157Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:07:20.180349Z digest=sha256:28d85ae52b559dcd6f830843552a1963d116c0cd5489b7d43d015486ecc5bfaf

Observation 9ad932fa-1dbe-4f2f-a761-b03deac132c5 · inbound

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models cites this paper.

Temporal Contrastive Decoding: A Training-Free Method for Large Audio-Language Models Tuning Language Models by Proxy

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T10:19:20.680811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T10:16:26.526986Z digest=sha256:95ada1377a3dd3e6a094aa3f9ccfc53587bb8602033a0788913a63531c35cf4a

Observation 72559b78-756f-4a17-ab44-4b65a2e8e7d7 · inbound

Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs cites this paper.

Complementing reinforcement learning with SFT through logit averaging in the post training of LLMs Tuning Language Models by Proxy

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-21T06:39:43.651320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-21T06:39:06.365240Z digest=sha256:84dd1e100644c07b762a73104a379cfafbe2f0fb61204b6dec35abea5f9847f8

Observation ab6a4246-7973-4358-8b72-a7543faf22f8 · inbound

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them) cites this paper.

Failed Reasoning Traces Tell You What Is Fixable (But Not by Reading Them) Tuning Language Models by Proxy

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-02T07:46:45.672823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T06:48:28.788806Z digest=sha256:4ab9875f29c91af17a4c3e74bdfbb79a2552795e77ee29eeb0d41a78bd69602d

Observation cde6781f-db19-42ab-8564-a536ac633c9c · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Tuning Language Models by Proxy

Reference 110

Resolution
verified exact
local_arxiv, observed 2026-07-07T12:33:45.022809Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-07-07T12:31:42.224094Z digest=sha256:0afae1b9f8181ec5792866a3cb4e08d5d9f9831d046e9aa6b05aeb9b5e1b7290

Observation 1957543d-eade-4ff2-8eb8-308fa5535ebc · inbound

Weak-to-Strong Generalization via Direct On-Policy Distillation cites this paper.

Weak-to-Strong Generalization via Direct On-Policy Distillation Tuning Language Models by Proxy

Reference 106

Resolution
unresolved
no resolver link, observed 2026-07-11T07:01:56.628017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T07:01:56.628017Z digest=sha256:50853bbc7441b343b07b1bcfe79d18461f3d3da63787959c55cafd7b523c128d

Observation 9780541d-452f-4e6e-86cf-3b72c5293382 · inbound

Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-Guided Update Signals cites this paper.

Proxy Exploration and Reusable Guidance: A Modular LLM Post-Training Paradigm via Proxy-Guided Update Signals Tuning Language Models by Proxy

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-14T05:09:00.865375Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T05:09:00.865375Z digest=sha256:245d9499cb064d1a766556db568b7ffb523508a17a63baa64a9f0982d132cadc

Observation 8d2eae3d-df45-4cb0-b8b3-41f5bfc8df15 · inbound

Weak-to-Strong On-Policy Distillation cites this paper.

Weak-to-Strong On-Policy Distillation Tuning Language Models by Proxy

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-01T00:26:23.349392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T00:26:23.349392Z digest=sha256:afc72e738fcfa630b11a28c45eecb08563fe9c3a7306448ae80af0a3d484f719