Pith. sign in

Paper Citation Record · LEDGER

Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 48 inbound Pith citation observations for arXiv:2002.06305.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2002.06305 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 48 of 48 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 48 of 48 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:25:05.166635Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

216
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 99b8bf49-1363-404e-aeac-d67604dc7577 · inbound

Learning to summarize from human feedback cites this paper.

Learning to summarize from human feedback Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-18T01:46:18.549473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T01:46:18.486086Z digest=sha256:acec5936ea68101e736acca56a0f95f6ba38d59a4d92a8d3d094b1f9f04ad789

Observation afa31a33-3065-4a80-b260-884b4fb6ebde · inbound

Editing Models with Task Arithmetic cites this paper.

Editing Models with Task Arithmetic Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T08:09:12.915111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T08:09:12.716163Z digest=sha256:646ffdd6dfc08f5a22d8d4456e68fb93ace8974978a870d484d0aa7efc8dd3aa

Observation 47d63d2b-5a44-4e1e-9950-9b31a0884199 · inbound

LIMO: Less is More for Reasoning cites this paper.

LIMO: Less is More for Reasoning Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 177

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T02:11:37.389452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-17T02:11:36.932541Z digest=sha256:c9b75b7e7272ad156aefaa9272c3ea8aec9cf077db44dfc1360c89321e30ef44

Observation 931f7b98-c0b7-4002-a349-0c72d5c17338 · inbound

Revisiting Bayesian Model Averaging in the Era of Foundation Models cites this paper.

Revisiting Bayesian Model Averaging in the Era of Foundation Models Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-07T13:25:05.166635Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:25:05.166635Z digest=sha256:ae62929a6ac8f4d5dbea3f0b83c516d95ebb7779b616b2f41353fc854b54814a

Observation 66db38e6-6110-49fd-a4ed-4db14092ef50 · inbound

Behavioral Augmentation of UML Class Diagrams: An Empirical Study of Large Language Models for Method Generation cites this paper.

Behavioral Augmentation of UML Class Diagrams: An Empirical Study of Large Language Models for Method Generation Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T12:02:52.397171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:02:52.397171Z digest=sha256:02ba2449b7015fb1b0c930daa050192fc063b19137a8af69ea495eb8de3e6513

Observation b2940480-ba75-4f17-9cc1-20080a795505 · inbound

Gradient-Based Model Fingerprinting for LLM Similarity Detection and Family Classification cites this paper.

Gradient-Based Model Fingerprinting for LLM Similarity Detection and Family Classification Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T11:45:37.662426Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:45:37.662426Z digest=sha256:25a8382acfb7414d778b920e934f15f378cda9cee73ab3cb7a763b905818c758

Observation 79f2caed-bde8-4a7a-88e4-8b78accb4f8e · inbound

RewardAnything: Generalizable Principle-Following Reward Models cites this paper.

RewardAnything: Generalizable Principle-Following Reward Models Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-07T11:04:06.940032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:04:06.940032Z digest=sha256:83c6339b1b23cc2a71ea56a7c0fe57d48f379bfb2465dd0a57f748edd389d1e4

Observation 67031d66-ee9b-4206-bcf1-c8a8d09c266d · inbound

GeistBERT: Breathing Life into German NLP cites this paper.

GeistBERT: Breathing Life into German NLP Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T01:08:58.370021Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T01:08:58.370021Z digest=sha256:f4a5ad12ae84d3128a836808cd281c463617a095354f2561a67b24823b0c3233

Observation 022713d0-2942-4f9d-bdb6-55ae00c8bcbc · inbound

Breaking a Logarithmic Barrier in the Stopping Time Convergence Rate of Stochastic First-order Methods cites this paper.

Breaking a Logarithmic Barrier in the Stopping Time Convergence Rate of Stochastic First-order Methods Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T21:58:12.293683Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:58:12.293683Z digest=sha256:326b68b51f40f37172e55361a868a007e4ac266899363f64b85939a31483a795

Observation e77693ec-4bdf-40a8-a9a4-c9f12bf8f1e5 · inbound

Should We Still Pretrain Encoders with Masked Language Modeling? cites this paper.

Should We Still Pretrain Encoders with Masked Language Modeling? Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T06:32:07.608630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-19T06:31:37.201344Z digest=sha256:a9bd91f9f7a92633068be071245cf6a0bb376b64b8e0cc8691a4e547dbd9a6e2

Observation 8ee894d9-55b1-49dd-870c-d18c85448a54 · inbound

Can Interpretation Predict Behavior on Unseen Data? cites this paper.

Can Interpretation Predict Behavior on Unseen Data? Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-06T19:10:28.427648Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:10:28.427648Z digest=sha256:7a1f63c04b77681a8b37ab27f715a1be87845e37bbb5278e39b0d739431a400d

Observation c0dbd269-8080-4cb3-937d-35fa7a6e832a · inbound

Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis cites this paper.

Improving Data and Parameter Efficiency of Neural Language Models Using Representation Analysis Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-06T17:03:41.362392Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:03:41.362392Z digest=sha256:a240f615c6fcb5bc4d2c3610c0c3ccc031629df2d15c7219f2a2cfb75a305353

Observation 71b7df51-d7ec-4042-8fb5-cb1fa778e286 · inbound

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges cites this paper.

Alignment and Safety in Large Language Models: Safety Mechanisms, Training Paradigms, and Emerging Challenges Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 143

Resolution
unresolved
no resolver link, observed 2026-08-06T14:13:06.458268Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T14:13:06.458268Z digest=sha256:9a0e1573300ab633e7cf1c0a27f9f3ec2366023e7451e13c4f70120b13d7db65

Observation 372fe7b1-d6a9-4fdd-90dd-37a98ba3ad85 · inbound

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning cites this paper.

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T22:21:53.341085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-18T22:17:41.758059Z digest=sha256:e75d72f90acaf4273cbaa54fb2b65763e89ee460b8db368e96bbb8dfc73852f9

Observation 17266caa-5e18-4b01-9029-9b298e9c4a77 · inbound

LobRA: Multi-tenant Fine-tuning over Heterogeneous Data cites this paper.

LobRA: Multi-tenant Fine-tuning over Heterogeneous Data Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T12:56:40.554825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:56:40.554825Z digest=sha256:9dc910382400f63ab3bd9033c4e158991a13b3901a403da930522b177506b9d6

Observation d233f646-f69c-43cf-811c-201b5b3eebc7 · inbound

SindBERT, the Sailor: Charting the Seas of Turkish NLP cites this paper.

SindBERT, the Sailor: Charting the Seas of Turkish NLP Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-04T08:22:13.081531Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T08:22:13.081531Z digest=sha256:bf98e03c3b41e797e530f4428428fdd2a6dfdb0662708ff54bf6b69665cc237c

Observation 21fea34f-8013-43e0-a42d-cb2833372c9d · inbound

Stay Unique, Stay Efficient: Preserving Model Personality in Multi-Task Merging cites this paper.

Stay Unique, Stay Efficient: Preserving Model Personality in Multi-Task Merging Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-03T19:15:52.462406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T19:15:52.462406Z digest=sha256:c8b5a057e59d1a2e9d3d63ee26b7d5a940b0fbb7c9ba2e74666220e404f2437b

Observation 6fa9eb8a-cea6-47fe-86f3-d622164bb94b · inbound

In-Context Probing for Membership Inference in Fine-Tuned Language Models cites this paper.

In-Context Probing for Membership Inference in Fine-Tuned Language Models Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-03T15:38:43.752637Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:38:43.752637Z digest=sha256:b1b86f3ed3c37cbfe46971d4ee17e3ded1ae69faf27c9eaabe510e9403bf1e77

Observation aa8e59b5-9381-4230-9af2-ba4fe17b6c80 · inbound

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation cites this paper.

Beyond Transfer Accuracy: Faithful Circuits for Controlled Low-Resource Adaptation Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-03T10:56:21.770183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T10:56:21.770183Z digest=sha256:dac8f71823a7a90c74aab2db6391ec9f059d0aeefbdabdf20c8a93454a88824b

Observation 7c8a1f8c-512f-4627-a737-1e2586556a52 · inbound

Robust Policy Optimization to Prevent Catastrophic Forgetting cites this paper.

Robust Policy Optimization to Prevent Catastrophic Forgetting Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 16

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T05:37:24.444353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T05:33:42.965249Z digest=sha256:5ee5e52bb60fa5e3ee3c629bbf9fc3ce023ed33525fde4749410c5954904f21f

Observation 363af6f4-f403-4ab6-98fa-81ecc4229353 · inbound

If It's Good Enough for You, It's Good Enough for Me: Transferability of Audio Sufficiencies across Models cites this paper.

If It's Good Enough for You, It's Good Enough for Me: Transferability of Audio Sufficiencies across Models Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T18:58:08.985758Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T18:53:48.881039Z digest=sha256:90d59d80d70fea1017d1db98a04b1edd218a76f51ee57decc3897d53fb2125cb

Observation 38a09036-9f9e-4b69-b036-4e492a8d5e0b · inbound

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution cites this paper.

Towards Adaptive Continual Model Merging via Manifold-Aware Expert Evolution Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 1

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T19:21:07.214859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T12:16:22.278245Z digest=sha256:99e6faeb92c3f5efcc1fc2e7b9861ba04f0654e45bf7dbf1f236a707a1216b59

Observation 047e66d7-5297-43eb-959f-2de1f40b206f · inbound

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages cites this paper.

Dependency Parsing Across the Resource Spectrum: Evaluating Architectures on High and Low-Resource Languages Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 38

Resolution
metadata mismatch
arxiv_id, observed 2026-05-09T05:45:21.657339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-08T19:36:09.771806Z digest=sha256:96c5d2b799d53c03edb1e6cae9e3161f8fbc0b0f8ae467b09deb37d134511039

Observation 5ca5e47a-f5e4-4711-a8cd-5d314f4cd172 · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 196

Resolution
verified exact
arxiv_id, observed 2026-05-13T03:12:09.663244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-13T03:09:02.902912Z digest=sha256:f72fafd9f14dc66d60f997920b48e35039815f349e7e016ddd39808c657b69de

Observation f6984efe-51e9-4b11-8b18-944aa6f334d4 · inbound

Instructions Shape Production of Language, not Processing cites this paper.

Instructions Shape Production of Language, not Processing Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 196

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:02:58.776483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-14T21:02:02.135970Z digest=sha256:d2456e52e4679d9f24c8241f60ddabfeb5a10d35a09d9a865b0d77886f36a61b

Observation 7bfd5b7e-726a-4d1c-8a0d-1c10414dc513 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:57:09.430005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T02:52:43.674969Z digest=sha256:273079ae477d6585219ffa4574a8aa34346d262fd16b6daa5fdb56678f4c2811

Observation e9480a13-e071-48a0-acd9-5af89844e208 · inbound

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone cites this paper.

20/20 Vision Language Models: A Prescription for Better VLMs through Data Curation Alone Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-14T21:29:28.696436Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-14T21:28:37.680681Z digest=sha256:10d220572439a90403e27c790213392f049cde10117ef02264a2bf6f89cabff1

Observation 2517ff75-3a82-4aac-8bd5-25be0473d21a · inbound

LoRA vs. Full Fine-Tuning: A Theoretical Perspective cites this paper.

LoRA vs. Full Fine-Tuning: A Theoretical Perspective Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:18:16.801931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T12:14:11.947836Z digest=sha256:e49065f873e703212bfdd8c9d8c3ac3bc501d82743bd3c94c7cea157b308b44f

Observation ca35d6f6-a5e9-45e3-9454-965f879c84cf · inbound

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling cites this paper.

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:43:05.863325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-20T05:41:17.997800Z digest=sha256:2d8b4e877150ee9baaadee07fc18c7716171183dc134e74f7fa74b728dfce045

Observation f49a51fc-ea9f-4be6-a8c9-42960eecb63d · inbound

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling cites this paper.

PromptRad: Knowledge-Enhanced Multi-Label Prompt-Tuning for Low-Resource Radiology Report Labeling Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:54:02.637557Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-05-21T07:53:46.273508Z digest=sha256:a8f599d3fd445e04a2ca90cdd00f4588edb8e11d76fb4dc3c4a07f3f7e0488f8

Observation 9dce5377-20ef-4699-b079-4fc60b94cadd · inbound

Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias cites this paper.

Do Deep Networks Forget Initialization? A Forgetting-Time View of Practical Inductive Bias Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-06-29T13:23:28.023366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-29T13:20:54.303605Z digest=sha256:604a0b6de70620e779598639d7af4a80953cfca8fd5a6d3ceacf2322b6da9c9d

Observation 9bb93867-19eb-442a-8a51-673fb8b042b9 · inbound

BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning cites this paper.

BadBone: Backdoor Attacks Against Backbone Models in Visual Prompt Learning Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T19:56:11.294800Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:52:23.150188Z digest=sha256:47f140c0821cdffc19fbe2dfc5a891af63c3d8d9f4280edec345065ec988d8db

Observation b9ff0799-230a-4d9c-9b78-a41fa84c65d2 · inbound

PortBERT: Navigating the Depths of Portuguese Language Models cites this paper.

PortBERT: Navigating the Depths of Portuguese Language Models Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T23:06:20.221903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-28T14:43:28.401687Z digest=sha256:f7cdb69b90a5660e7385cf61ea86b26596e52f0b6445ac7d20c5b000a9bd9c3e

Observation 801769d5-86c1-4dff-b51d-0461a5974d56 · inbound

On the Geometry of On-Policy Distillation cites this paper.

On the Geometry of On-Policy Distillation Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 3

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:27:08.862860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T22:47:04.213358Z digest=sha256:2ab1b73e74ca1357bb13bc5c072c29e0b6a5865de845d0276cde727fc047124f

Observation 5e248862-10b4-4d88-96c3-58754aed301f · inbound

Phantom Transitions in Language Model Fine-Tuning: A Density-Matrix Analysis cites this paper.

Phantom Transitions in Language Model Fine-Tuning: A Density-Matrix Analysis Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T22:04:00.576999Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-29T21:54:30.991573Z digest=sha256:68d55b0ef0c01b43e95ce8170ee6f12d0cde0d84ba7f04e434c722bc14b3c529

Observation ebd6ee2f-ea30-48b4-850b-6e8ad4d48ac0 · inbound

Phantom Transitions in Language Model Fine-Tuning: A Density-Matrix Analysis cites this paper.

Phantom Transitions in Language Model Fine-Tuning: A Density-Matrix Analysis Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-02T13:16:38.124570Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-02T13:16:38.124570Z digest=sha256:fa4e39a3a4fb2c66a6399edd87797cfaba07373c2f8de4a2d5bd228f2a259b08

Observation fd0521bb-ffb5-495e-9e51-a482d1e2dc0f · inbound

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents cites this paper.

Afrispeech Semantics: Evaluating Audio Semantic Reasoning in Spoken Language Models Across Domains and Accents Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 272

Resolution
verified exact
arxiv_id, observed 2026-06-30T22:15:05.450842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T22:11:44.891731Z digest=sha256:8e92d5a3bda4bd3d44eeebef70c6cf3f223c5b1e74b07c9b96a14e036f34b958

Observation 184ff01d-adbf-4fc1-a310-130609772a5f · inbound

Sparsity Curse: Understanding RLVR Model Parameter Space from Model Merging cites this paper.

Sparsity Curse: Understanding RLVR Model Parameter Space from Model Merging Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T21:08:57.821995Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-27T00:57:02.938997Z digest=sha256:244a96a43b93d2f92f1c0c89ec95faae1d88e3219d82ceac9b76181071d9daa3

Observation e927e833-4491-4ed6-961a-358ea600a34d · inbound

MiqraBERT: Regression-Based Sentence-BERT Finetuning for Biblical Hebrew Parallel Detection cites this paper.

MiqraBERT: Regression-Based Sentence-BERT Finetuning for Biblical Hebrew Parallel Detection Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 4

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T01:19:20.994944Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T20:30:05.043803Z digest=sha256:59929bf9893015d0f005cc44fca8a212ed89dc6e145d2e3808fb87921ced35bb

Observation 461299ca-31ab-42f5-9b20-daf5450a4ce1 · inbound

MiqraBERT: Regression-Based Sentence-BERT Finetuning for Biblical Hebrew Parallel Detection cites this paper.

MiqraBERT: Regression-Based Sentence-BERT Finetuning for Biblical Hebrew Parallel Detection Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-06-26T20:49:57.860299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T20:30:05.043803Z digest=sha256:acd3d43174699c33f67dea66dd640b3fe1468ad8c18542e2f48c2a9db796dcc4

Observation 014aceff-902a-4c32-b1a8-af77508d665f · inbound

Repository-Level Solidity Code Generation with Large Language Models: From Prompting to Fine-Tuning cites this paper.

Repository-Level Solidity Code Generation with Large Language Models: From Prompting to Fine-Tuning Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T04:49:34.445786Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T16:48:12.177090Z digest=sha256:4da3f7562fcfa326506838e6b93d706e4d00ba07b80cc50e7df89fee528dec07

Observation 4cef2d21-e320-4d4b-863c-71bb436ec705 · inbound

The FID Lottery: Quantifying Hidden Randomness in Generative-Model Evaluation cites this paper.

The FID Lottery: Quantifying Hidden Randomness in Generative-Model Evaluation Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T03:49:29.528251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-26T17:42:21.628047Z digest=sha256:6411cdb54fc3051f951b4f17252873dfca19b75f3a5db1bc3ea10110bce4747e

Observation 2be36366-81bd-4a57-acdb-4dbf56f663f0 · inbound

Sub-Billion, Super-Frontier: Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction cites this paper.

Sub-Billion, Super-Frontier: Small Language Models Rival Zero-Shot Frontier LLMs on General and Literary Relation Extraction Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 249

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:19:42.859218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T10:18:29.700444Z digest=sha256:895baae14e489a998b75f4720bf6d544ced83a065e9d2e3a6260b42018f0655f

Observation 75c89c94-3105-427c-ab92-09316fecc380 · inbound

GRAIN: Group Aggregation via Min-Norm Objective cites this paper.

GRAIN: Group Aggregation via Min-Norm Objective Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 72

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T09:59:45.269882Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-26T09:18:55.049767Z digest=sha256:b6f2814bd1a30178144e00b68280418a587e275aa0d592e76b31b95c3ed1123b

Observation 8039ae6c-1868-4493-9144-5659599b240b · inbound

Optimizer Memory Makes Shuffle Order a First-Order Source of Fine-Tuning Noise cites this paper.

Optimizer Memory Makes Shuffle Order a First-Order Source of Fine-Tuning Noise Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T07:24:21.865183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-30T07:15:56.974540Z digest=sha256:53eed5a6dc30b5bd32b771056fa9675c3b1cea96782fb696be1085dd50e1537d

Observation 3cfef4e6-f123-47ec-990d-3c71ffae3f54 · inbound

Training Large Language Models for Self-Explanation Faithfulness cites this paper.

Training Large Language Models for Self-Explanation Faithfulness Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-01T08:36:23.469133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T08:36:23.469133Z digest=sha256:6309c2d6ff6bd0c1740db0afa6adbb64acad3163638130c2f8788edc56937993

Observation f0d39616-0efc-41c9-9614-13a0b18ab9e9 · inbound

Latent-LoRA: Compact Latent-Space Adapters with Gradient-Free Routing for Continual Learning cites this paper.

Latent-LoRA: Compact Latent-Space Adapters with Gradient-Free Routing for Continual Learning Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-30T10:55:15.500484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-30T10:55:15.500484Z digest=sha256:3aa63b91cd6bd88824570cd7dc5c167aa2298cffabbb09c0a6e10ca63c6fb768

Observation e787555d-0ada-4d70-89b4-998b2c209f89 · inbound

What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend cites this paper.

What We Observe as LLM Behavior Can Be a Side-effect of Inference Backend Fine-Tuning Pretrained Language Models: Weight Initializations, Data Orders, and Early Stopping

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-06T18:22:28.042785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:22:28.042785Z digest=sha256:7610b2bd8d02efa80850ac02599d5535c9d8762a4fa649493bf6f2b4c3bad1d4