Pith. sign in

Paper Citation Record · LEDGER

Optimal Training-Time Scaling in Gradual Adaptation

As of 9 August 2026, this Paper Citation Record lists 100 of 107 outbound references and 0 inbound Pith citation observations for arXiv:2608.04927.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2608.04927 v2

Coverage vector

measured 100 of 107 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-08T17:23:55.291013Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 107 outbound references displayed

  • verified exact4
  • verified fuzzy55
  • unresolved40
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation e23d139f-1991-478d-b669-65b08eb3a25c · outbound

This paper cites Neural Networks , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Neural Networks , volume=

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.951480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.951480Z digest=sha256:e8a50e28a395e6cbd89b83aeef8f375d6514b876ee39b8d2f06553a240c17fe0

Observation 7a282466-fb66-4004-b1e4-f155f8454736 · outbound

This paper cites IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Transactions on Pattern Analysis and Machine Intelligence , volume=

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.957471Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.957471Z digest=sha256:acfc0b8e274d1e4bc1a08467ca8fcd09a3d7f3aff549690104756f0e3c953306

Observation 6d1d07d9-1105-4b79-ad3a-65696087eac6 · outbound

This paper cites Proceedings of the 37th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 37th International Conference on Machine Learning , series=

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.961685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.961685Z digest=sha256:a411cac3c748a0c253eeb996eb8e732e2dd6b3bb3dd557e3806902f07bd86733

Observation 3fcc8628-8b6b-4a8c-a121-7c461c655c90 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.965989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.965989Z digest=sha256:90fb3e71e78ce8e4dc5cd7e8017fe3e604fa8d87978149f156621860c0d9f66a

Observation 6125fc55-c75c-435d-a75a-cffe541fb413 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.970044Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.970044Z digest=sha256:4e82306c973ef540dd03ae86305303be75ef289acd8fc8d3ebbdfdabb0736f00

Observation 50435ea4-65a9-41b5-a061-14af19d3ba2b · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.973946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.973946Z digest=sha256:84a7bf157bb6e6425632783a3d4cbf78e080d556be63dc50239e3d036a96239b

Observation 82b4073b-100c-475f-8d38-686a70e53b46 · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.978059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.978059Z digest=sha256:cd911a403910faa298960de4940407aa271382ea347e2218e013145ede9d0b40

Observation 71bc4d43-8295-4b69-818e-e173224bb37b · outbound

This paper cites Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE/CVF Conference on Computer Vision and Pattern Recognition , pages=

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.981486Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.981486Z digest=sha256:0fdfc95c234893d33e5e7e4c6549a034c3ae4c237601cdc6edc5fcec8b131be7

Observation 80d2337f-8934-4f0c-81c9-2f90be31c8b3 · outbound

This paper cites IEEE Robotics and Automation Letters , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Robotics and Automation Letters , volume=

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.984888Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.984888Z digest=sha256:44cb51971550f3ed2ac954dc292672efcdbb68bc659d9dbf2ee80ea747348ae6

Observation e793334b-7eb1-4cb4-be75-763fa4db1626 · outbound

This paper cites 2025 , publisher=.

Optimal Training-Time Scaling in Gradual Adaptation 2025 , publisher=

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.988603Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.988603Z digest=sha256:0a3c53ca4fbcf6a09d3f62b9f6bbd7f67191000fc2b6d0885954689ab1bc2a09

Observation 79410c6d-b280-4e80-b3c5-428bba7960af · outbound

This paper cites The Eleventh International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation The Eleventh International Conference on Learning Representations , year=

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.992100Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.992100Z digest=sha256:e1ea3e0e9bc8b27d71de6d279af210937a106f33e6d6b06331d3761174ca5aa0

Observation 778a81ab-3bf9-49b0-b3a8-df6fdc3026a0 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 34th International Conference on Machine Learning , series=

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.995740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.995740Z digest=sha256:3b68bb54373a5d00dcf9874fd36860667b483a85e447e47b37a325b1150b6c4b

Observation 4a65cce3-68b0-40eb-b4ef-ad8fd3237142 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:54.999264Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:54.999264Z digest=sha256:8330d5954579992c4493bcabd2f7d94cca33461224072f25e06bc6da311686b4

Observation b33f205b-2587-45fe-b14f-f02654ca301d · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.002614Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.002614Z digest=sha256:4e901b619e0a50192bb9a79c78ba028c1a34833503840773795b49ebe5a515ed

Observation cc1cdec2-fe01-464d-9920-b033d85110a2 · outbound

This paper cites The Thirteenth International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation The Thirteenth International Conference on Learning Representations , year=

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.005983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.005983Z digest=sha256:5ff41c89ad25f40c6dd1f21ade5c04492d202f8001fafa77e40014b2761b0787

Observation b4176eaf-f84f-407f-b9a4-9200126447ac · outbound

This paper cites Constructive Approximation , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Constructive Approximation , volume=

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.009041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.009041Z digest=sha256:9b0143d27614fd2c3c29c6ed2ce27496da1ee864993623bbfa88c0f422c01e52

Observation 28723244-f315-412c-8e3d-d17372ff3acf · outbound

This paper cites Journal of Machine Learning Research , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Machine Learning Research , volume=

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.011834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.011834Z digest=sha256:553b0a3975b3191b1f4af41b07d36b537ec410029c51b1a6e9609f6160373232

Observation bf8d2157-ba62-4526-88fe-550b97a6beda · outbound

This paper cites Proceedings of the 22nd International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 22nd International Conference on Artificial Intelligence and Statistics , series=

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.015260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.015260Z digest=sha256:95f468ee77c0a96eb3f05e88010313cf039c6aaa85952a38ddf3dde208a75869

Observation 24150688-c0c6-4dda-a958-1f92ea3cc9d5 · outbound

This paper cites Proceedings of the 33rd International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 33rd International Conference on Machine Learning , series=

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.018585Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.018585Z digest=sha256:d1a4125d09b6db29dbb159ee8e3e2dc0a61ba7d50e93c5dcc8ca4125b74c50a7

Observation 8b943b9d-b7c7-4560-b07a-2f8dd14a64d3 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.021522Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.021522Z digest=sha256:a08f9f3ed2f24e545c9455a6b5a664de56b59e5eeb020e70ac8636e4d9bdaecb

Observation 1ae1ee5b-0bd6-4b00-b173-6f0f6f9196d5 · outbound

This paper cites Proceedings of the 41st International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 41st International Conference on Machine Learning , series=

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.024692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.024692Z digest=sha256:1c5411a8a0fa009dd7d3dcbc5dbd1aebc84ca9935ade190aa1039f4ffd4cbac0

Observation 8cf76e94-e1cb-4a1e-bd21-856ad541e518 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.027910Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.027910Z digest=sha256:5ba9f4288ba40a8d2624c847a10809084b3c2386c706a620e5df95751a469b38

Observation c11a9925-a61c-4d91-b5d6-575ac9889460 · outbound

This paper cites Proceedings of the 2nd Conference on Lifelong Learning Agents , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 2nd Conference on Lifelong Learning Agents , series=

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.030952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.030952Z digest=sha256:b6fde4cea8573bf2c6d48135f1f941a82352c2bd16370e81a746b2bd10a585ca

Observation 8e5d90de-de4b-4aaf-9776-b05eea297ec4 · outbound

This paper cites Proceedings of the 42nd International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 42nd International Conference on Machine Learning , series=

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.033985Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.033985Z digest=sha256:d50eeaa000fa0d7249f7ecb34b49c04abdb0ec11fefbbee9829335eb51b4ecb5

Observation 42fb60bd-eb1d-49df-88a4-e812f61aebba · outbound

This paper cites Proceedings of the 35th Conference on Learning Theory , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 35th Conference on Learning Theory , series=

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.037524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.037524Z digest=sha256:f3fa33aabeb20ae53a704ea8a8db833c931114c02591a905b5820fa1b316ef52

Observation 0bfc006c-c88d-4bd6-b4d3-9d91718a74f4 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.041086Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.041086Z digest=sha256:69c8dafc41c848eb6da299fa9a549c5cb0bee9e56ff91f034784d747870ef8bd

Observation fcc5815e-fcb0-40bc-a5de-1d755b6e7484 · outbound

This paper cites From Continual Learning to.

Optimal Training-Time Scaling in Gradual Adaptation From Continual Learning to

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.044205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.044205Z digest=sha256:7a0558dd8e34061c611c9c7f38a4e4dd29c242c058b2606b0b636a2009b6d8c8

Observation ee320a5c-54d4-42a2-9011-38725bb6a8ce · outbound

This paper cites Journal of the Physical Society of Japan , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of the Physical Society of Japan , volume=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.476394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.047395Z digest=sha256:aeb45b764ce4384ca96461b1c64b14197368af92d3d520310a3d2e4ac70497a5

Observation c7381799-d744-44a3-b69a-239176c81b92 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 29

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:23:56.467980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.050567Z digest=sha256:9ee6f2e588d10e4441cfce8b9f7d63f414c9d731182b157e485b2a220a202437

Observation d27339df-c071-47d2-b9ba-f1b2336d30cd · outbound

This paper cites Communications in Mathematical Physics , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Communications in Mathematical Physics , volume=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.458708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.053770Z digest=sha256:756887839c6259a15a63b466a5c85dcd853fe6fef5d7ef39fea4de2a71404e9f

Observation c5b818af-c52f-40a9-9308-9a8e5ba2c282 · outbound

This paper cites Journal of Dynamics and Games , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Dynamics and Games , volume=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.449281Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.061270Z digest=sha256:6a974cb47d6a637e5599ddae51f3ec13f4672692fa54feff15bedb65306c9a96

Observation e0536149-9791-4654-909e-8c130c9f8f53 · outbound

This paper cites Proceedings of the 26th International Conference on Machine Learning , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 26th International Conference on Machine Learning , pages=

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.440004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.064990Z digest=sha256:f2426843ec1adff565bef4775e4bf584dd7461b4e88259aed1dc3eeee776f3e3

Observation 5e2c05b4-f0ce-4ee3-8c72-c5efb30900e5 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.429766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.068951Z digest=sha256:1904174164ee9a415c4ae15109a567977c70fff338f5047540348cf8be8cdf9d

Observation a3d88299-68a2-4e5e-acba-c6d9040b0ebc · outbound

This paper cites Proceedings of the 35th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 35th International Conference on Machine Learning , series=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.419862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.072405Z digest=sha256:41a49624396ba81e929bbcb54b79af4b3fa84dab24ae38377f27245b473fecda

Observation b48ecbe0-5187-4281-b5bf-d5962d64dfcf · outbound

This paper cites Proceedings of the 36th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 36th International Conference on Machine Learning , series=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.409412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.075631Z digest=sha256:3da39bd9b71fa263ebe7456dee0815daf7ee81ed2813dd0188b6badde6adb211

Observation 0bce7133-5bdf-4270-a477-480d3037b8f5 · outbound

This paper cites Proceedings of the 37th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 37th International Conference on Machine Learning , series=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.399323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.078849Z digest=sha256:3d9541ee8e28cbf5913a3eb2055c2d30867e3cc52a05ddcc46aa0d1234bbbac0

Observation abf703a8-30b2-4a38-8f41-ddbc38c79f19 · outbound

This paper cites 2021 , url=.

Optimal Training-Time Scaling in Gradual Adaptation 2021 , url=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.388724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.081993Z digest=sha256:fbcd068bf56b5bef65aace5d933ebef1e2dd1d6accca4a166576a0f5c9418be5

Observation 22ed719c-ec9a-482e-8698-58d4def3ff01 · outbound

This paper cites Proceedings of the 39th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 39th International Conference on Machine Learning , series=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.378213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.085529Z digest=sha256:51f07faf04e5d3c7db17222712c062f4dbd4d4d14e2e97b4fc8a47198953871c

Observation 4a862d50-9737-467d-b0a8-d46b6e8e50c6 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-08T17:23:56.367407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.088866Z digest=sha256:9829f0c678688defa07ac4ba52e9e75a63a0e7afcf6325a9bea56394b854609b

Observation 12ba1344-cd11-4d56-bcb0-fa9718b68417 · outbound

This paper cites International Conference on Learning Representations , year=.

Optimal Training-Time Scaling in Gradual Adaptation International Conference on Learning Representations , year=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.357236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.092702Z digest=sha256:2bb6f76512d2de64b1f6374a8ab87ace92ebed1ac8ff2221ee934d8b4b4c6ae0

Observation e077afb1-fcca-4305-b605-444d669c55c0 · outbound

This paper cites Machine Learning , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Machine Learning , volume=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.346734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.096213Z digest=sha256:a857f02f51202ba92c8fcf927e067e7d9b9126ba246598b89880660bdb407efd

Observation 52125504-a4da-4673-a341-6e68c74ada28 · outbound

This paper cites Journal of Machine Learning Research , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Journal of Machine Learning Research , volume=

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.099569Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.099569Z digest=sha256:f45db87c6f78c2394599a757c3ad4d916128b49d7e84517cf14a53fc903530e7

Observation 579a2bfe-8a9b-4cb5-892d-b76d3bf39c1e · outbound

This paper cites and Darrell, Trevor , booktitle=.

Optimal Training-Time Scaling in Gradual Adaptation and Darrell, Trevor , booktitle=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.331512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.103037Z digest=sha256:aeb1276fc36ae0c91de7ce74faebef678bb1741cc3a6eacc1488b77555299dab

Observation 1697a3b2-e63e-4565-8b87-0c89df3b5285 · outbound

This paper cites Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the IEEE Conference on Computer Vision and Pattern Recognition , pages=

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.322673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.106482Z digest=sha256:d83030133060a51206abe2e11d83be5b69061143843a32ccc81e68e128016e27

Observation 7d3cb934-a885-410e-995c-dcdd03f2eb6c · outbound

This paper cites Proceedings of the National Academy of Sciences , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the National Academy of Sciences , volume=

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.313575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.110104Z digest=sha256:cf7061be4425f07c0382c48b529b1bd96834d3eebb93e4ee1ae9aff5d5f8f72c

Observation ec8791c9-b6df-4e94-a37d-107286cefb90 · outbound

This paper cites Proceedings of the 34th International Conference on Machine Learning , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 34th International Conference on Machine Learning , series=

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.303863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.113652Z digest=sha256:4ced37f491bfa9f7f4e185e966080a8159aa386817303ccb38d3953db31a7e33

Observation dbf5d8cf-6174-4948-8eaa-3029c5e32ae8 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.116812Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.116812Z digest=sha256:6165c9abfa5c7a4e529930e49cc4cbc499da37243ff46e91cca80a9f1a573d28

Observation f885f70f-dd49-4e99-b0ce-30488489f77e · outbound

This paper cites Efficient Lifelong Learning with.

Optimal Training-Time Scaling in Gradual Adaptation Efficient Lifelong Learning with

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.120123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.120123Z digest=sha256:865984bcd6afb1a059feeef036ce60c67213511518f60f0dd080994eb4aa73bc

Observation 0385501a-c298-4012-a4f4-a28d5958e14f · outbound

This paper cites Proceedings of the 23rd International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 23rd International Conference on Artificial Intelligence and Statistics , series=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.279286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.123655Z digest=sha256:0ef7551f501b9b19b3e9f551661310ac209db75b7d1ff84063c9f7b62a936031

Observation aceb1571-d2a2-4446-a9d3-ec04bb57153d · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems , volume=

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.268028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.127136Z digest=sha256:82f156a578517ffa29495780fff64afedc8bafc3a607ab1e78384baf7b4b83bc

Observation 1e56ca72-5632-47dd-a3c5-4f41127f93b4 · outbound

This paper cites Proceedings of the 25th International Conference on Artificial Intelligence and Statistics , series=.

Optimal Training-Time Scaling in Gradual Adaptation Proceedings of the 25th International Conference on Artificial Intelligence and Statistics , series=

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.256433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.130420Z digest=sha256:b76f92886150b3552f79ffaeba78fcf89c72d71ca1ee5a2e45d0212fb80eb8ec

Observation 21bae2fe-2e8a-45f4-9c63-f696dd4eb6e2 · outbound

This paper cites Nature Machine Intelligence , volume=.

Optimal Training-Time Scaling in Gradual Adaptation Nature Machine Intelligence , volume=

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.245235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.133891Z digest=sha256:2ba252e859947a784fff14d0355283a7b6fb4f96c074eca6f5b7cc5ca9287958

Observation dda83449-15c7-4440-b343-65eef2a4f8dc · outbound

This paper cites IEEE Transactions on Computational Imaging , volume=.

Optimal Training-Time Scaling in Gradual Adaptation IEEE Transactions on Computational Imaging , volume=

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.235261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.137163Z digest=sha256:040a5ce2006382b07464472338d6bd1f910d0519848c519fa7cb9b6c463e872b

Observation ea11b1be-039e-4c82-ba3a-a108e88eb7e1 · outbound

This paper cites Advances in Neural Information Processing Systems Datasets and Benchmarks Track , year=.

Optimal Training-Time Scaling in Gradual Adaptation Advances in Neural Information Processing Systems Datasets and Benchmarks Track , year=

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.225014Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.140355Z digest=sha256:7563b2b6de3ee3ed719f3ca98a2af8a39b9bbbf11254e7350a7d5727a76002ca

Observation 9713cd79-b8a4-4b70-a7d0-530f0c7f2efa · outbound

This paper cites Zico Kolter, and Ryan J.

Optimal Training-Time Scaling in Gradual Adaptation Zico Kolter, and Ryan J

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.214530Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.143750Z digest=sha256:f4acfb28723dec74cff7f7b10a72e3caf1fd80823bc7a9b6e6bc8489f15908a5

Observation d8e4909d-b907-4bf2-a8f4-866881226438 · outbound

This paper cites an unresolved cited work.

Optimal Training-Time Scaling in Gradual Adaptation Unresolved cited work

Reference 57

Resolution
verified exact
doi, observed 2026-08-08T17:23:55.389672Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.147041Z digest=sha256:baaf32d7f77681cfbc954ebfdd2495e9d54f5ebf1314a60b4ff0f00da8c3fab3

Observation 525b6b8a-5743-494d-8216-0e23aa91a675 · outbound

This paper cites A theory of learning from different domains.

Optimal Training-Time Scaling in Gradual Adaptation A theory of learning from different domains

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.150234Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.150234Z digest=sha256:954aa4e9ac1e65e9f45fc0ccf7c5eda3e7672acda24959e556cbfb97c2f1a512

Observation bad29b34-21bf-45b2-866f-d0af9ecd2e5a · outbound

This paper cites Curriculum learning.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum learning

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.153061Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.153061Z digest=sha256:24f88f51425a90e1fcebaf4519d55c5f39e2e431ef1fce3c2ef649ce10c4dd00

Observation 667f9263-4f93-4677-bd48-e071811041de · outbound

This paper cites Unifying importance based regularisation methods for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Unifying importance based regularisation methods for continual learning

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.204195Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.155697Z digest=sha256:b43be6808992a2be0c2b5898c7fcb74601e60145089b8a0baa2176019d355288

Observation 32af7795-7e32-44c2-bf63-6c51dddddf4d · outbound

This paper cites Efficient lifelong learning with A-GEM.

Optimal Training-Time Scaling in Gradual Adaptation Efficient lifelong learning with A-GEM

Reference 61

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.192599Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.158282Z digest=sha256:c53ea5c5228cd3c030d2e8dbc766e50320e632328374cb51fce85496be692aec

Observation ba1927b4-4334-4475-9e8d-e9b32a3d67d4 · outbound

This paper cites Gradual domain adaptation without indexed intermediate domains.

Optimal Training-Time Scaling in Gradual Adaptation Gradual domain adaptation without indexed intermediate domains

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.180838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.160809Z digest=sha256:22e99b503982ce6219140b5428e5540d763a8724748412bcfd35ba972cc45e75

Observation 9c8b1a8b-80bd-4395-88a3-5e0f785e7ee9 · outbound

This paper cites A continual learning survey: Defying forgetting in classification tasks.

Optimal Training-Time Scaling in Gradual Adaptation A continual learning survey: Defying forgetting in classification tasks

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.163358Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.163358Z digest=sha256:28c4afff9083a67536bf18cba9ec520cb154a9dbf5eb69487ac1c5bd723dba44

Observation 3f2ca09f-6bde-4cda-ad4f-6d84f81fd830 · outbound

This paper cites Marsden, and Bin Yang.

Optimal Training-Time Scaling in Gradual Adaptation Marsden, and Bin Yang

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.167856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.165952Z digest=sha256:f668d6dd163a16ba79a1d83fe55bc5b97df56cb935a943898baa56363c11bf7b

Observation 19350aa3-5693-4370-83fe-412333536589 · outbound

This paper cites Dynamic update-to-data ratio: Minimizing world model overfitting.

Optimal Training-Time Scaling in Gradual Adaptation Dynamic update-to-data ratio: Minimizing world model overfitting

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.156451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.168575Z digest=sha256:fd850f2c9ca3807caf8c395d0c74ae36b63665dc20a8f2c8fb7fb6707a7185d0

Observation b737c49a-0958-4c88-9422-5a91fe48e588 · outbound

This paper cites Ward, Nathan Srebro, and Daniel Soudry.

Optimal Training-Time Scaling in Gradual Adaptation Ward, Nathan Srebro, and Daniel Soudry

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.145832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.171526Z digest=sha256:672fed956d381d1bb390ae9c8367f99d3d49124c756f2427c4c21931ac6448a2

Observation cf04c0e0-676f-4026-8cd9-695e5550a704 · outbound

This paper cites From continual learning to SGD and back: Better rates for continual linear models.

Optimal Training-Time Scaling in Gradual Adaptation From continual learning to SGD and back: Better rates for continual linear models

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.133975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.174951Z digest=sha256:77e9f22abb2a8c83199880f8925f5738029a11b71b0342af085066067983c6ca

Observation a0fd1e7a-1392-403b-9b9a-ff44a0e293ac · outbound

This paper cites Orthogonal gradient descent for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Orthogonal gradient descent for continual learning

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.122217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.178276Z digest=sha256:39969414851cd6ee36bca729e90bcbcc79ad1928e928013b98a9aaca00bb54ed

Observation 81540749-9d34-45e6-944f-989b88512639 · outbound

This paper cites Domain-adversarial training of neural networks.

Optimal Training-Time Scaling in Gradual Adaptation Domain-adversarial training of neural networks

Reference 69

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.181785Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.181785Z digest=sha256:18308b84ab251eee1c15d4177c2afaf2348b218b7f0dc6100a7503df15c6caff

Observation b5025385-0bae-4ef9-ac6c-69f0d699f458 · outbound

This paper cites a henb \.

Optimal Training-Time Scaling in Gradual Adaptation a henb \

Reference 70

Resolution
metadata mismatch
raw_fallback, observed 2026-08-08T17:23:55.597503Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.184959Z digest=sha256:ff9cb8ed21c2264281bae016fe37dfab5590ab3d771733aead00b8ffeae39f51

Observation 3242efd6-3052-4981-8000-e54763bf9572 · outbound

This paper cites The importance of being lazy: Scaling limits of continual learning.

Optimal Training-Time Scaling in Gradual Adaptation The importance of being lazy: Scaling limits of continual learning

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.102028Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.188394Z digest=sha256:44e08ddd6680fab45075cf23b65ab20040ff626f99e31189037d8d8a3816cb94

Observation b9d84df5-0cf4-437c-9d25-2b2bbe59f688 · outbound

This paper cites Bellemare, Jacob Menick, R \'e mi Munos, and Koray Kavukcuoglu.

Optimal Training-Time Scaling in Gradual Adaptation Bellemare, Jacob Menick, R \'e mi Munos, and Koray Kavukcuoglu

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.089738Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.191925Z digest=sha256:53d76812d53ed59a4a9df259f40838c30f3c5ede98344a92cc298ace99331352

Observation a13ea8ad-30f6-451e-86ed-880454ec70ca · outbound

This paper cites On the power of curriculum learning in training deep networks.

Optimal Training-Time Scaling in Gradual Adaptation On the power of curriculum learning in training deep networks

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.077937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.195309Z digest=sha256:abe29061fe0d2542e10018e4795f9499ce8bf5b7ace9d8a58b5f1a4fe9699580

Observation 39c725d8-48fc-454a-9bc7-fd07d05d5fd2 · outbound

This paper cites Train faster, generalize better: Stability of stochastic gradient descent.

Optimal Training-Time Scaling in Gradual Adaptation Train faster, generalize better: Stability of stochastic gradient descent

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.066390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.199394Z digest=sha256:c9765c632d834912fc79d9fc38030a690d8480a014c33be8d7a8ebe8870afdb9

Observation ecfdf522-f9ea-4d47-b8f4-888b1ecfbe11 · outbound

This paper cites Efros, and Trevor Darrell.

Optimal Training-Time Scaling in Gradual Adaptation Efros, and Trevor Darrell

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.054085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.202784Z digest=sha256:61e5b4c92db893b01d934706f12d53956b72481b4e8834c758be4171d291a945

Observation 07f56e1d-15a6-497f-a5b2-1541be0264f3 · outbound

This paper cites Curriculum reinforcement learning using optimal transport via gradual domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum reinforcement learning using optimal transport via gradual domain adaptation

Reference 76

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.042421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.206339Z digest=sha256:a27e8dfe5bb38e57d8e8076f28323936f7b7fe372cd53d78e663b68f3fa0e383

Observation f980245a-77d2-4992-bc01-25e7c75483d8 · outbound

This paper cites On the adiabatic theorem of quantum mechanics.

Optimal Training-Time Scaling in Gradual Adaptation On the adiabatic theorem of quantum mechanics

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.209876Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.209876Z digest=sha256:83cae5901da65a8a08dea8b59ec86c50fc70576b703a099f1b8c562844dd00f3

Observation 915bb5f1-f909-4f5d-a57d-65d7d5e9dbb8 · outbound

This paper cites Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell.

Optimal Training-Time Scaling in Gradual Adaptation Rusu, Kieran Milan, John Quan, Tiago Ramalho, Agnieszka Grabska-Barwinska, Demis Hassabis, Claudia Clopath, Dharshan Kumaran, and Raia Hadsell

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.213423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.213423Z digest=sha256:fb132626fc3040b3cba7868cf2bc000245abf10ee75b55dbccc9936113c9ea33

Observation 8efb8743-af9f-47cd-b4ae-70bd546a27aa · outbound

This paper cites Curriculum reinforcement learning via constrained optimal transport.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum reinforcement learning via constrained optimal transport

Reference 79

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.029953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.216889Z digest=sha256:c140e8d1cd2571298c31e77cc905c4270e139356a264450358f0840d21ce2dfe

Observation fa6f26bf-c0aa-4ae0-b206-1b5a1c646498 · outbound

This paper cites Understanding self-training for gradual domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Understanding self-training for gradual domain adaptation

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.017534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.220564Z digest=sha256:092b73fc7f511e9fbfbb3df11e4404e9050b3b0747e89d81c8f5e817c89c45e0

Observation 4fd5523c-07d7-4eaf-a037-4108aad70d6a · outbound

This paper cites Pawan Kumar, Benjamin Packer, and Daphne Koller.

Optimal Training-Time Scaling in Gradual Adaptation Pawan Kumar, Benjamin Packer, and Daphne Koller

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:56.004767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.224025Z digest=sha256:e669ae200331bc83e9672e2f72a66dbef1023e83542511881808b12514f19bbe

Observation e768f1fe-6e5c-40ed-858a-d3baeb0be961 · outbound

This paper cites Challenging common assumptions about catastrophic forgetting and knowledge accumulation.

Optimal Training-Time Scaling in Gradual Adaptation Challenging common assumptions about catastrophic forgetting and knowledge accumulation

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.992349Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.227650Z digest=sha256:065c4b43580793b31687e979b96cf651ebf31bc16a08bebe68921623c013a769

Observation 17326707-639d-4958-899a-6990a7e19c3d · outbound

This paper cites Optimal rates in continual linear regression via increasing regularization.

Optimal Training-Time Scaling in Gradual Adaptation Optimal rates in continual linear regression via increasing regularization

Reference 83

Resolution
verified exact
raw_fallback, observed 2026-08-08T17:23:55.512645Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.231073Z digest=sha256:f4f79b9703df2a556c20d20a850479287b8e2dd1094e92c6b222e0923ea6b861

Observation bc245f15-1dea-4e33-9dc2-d726e8e996c3 · outbound

This paper cites Gradient episodic memory for continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Gradient episodic memory for continual learning

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.979861Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.234563Z digest=sha256:47480c8c839b5456cdc1ae56deb66cf0dcc5a7e041207458501d71218e176774

Observation 5758210b-7cf1-4e05-b5bf-1a0d2f393328 · outbound

This paper cites Understanding the role of training regimes in continual learning.

Optimal Training-Time Scaling in Gradual Adaptation Understanding the role of training regimes in continual learning

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.968182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.237802Z digest=sha256:ca9f82298922f91514a3fd87b4b3b2715730bc9d0dd6249f7f9fbba76f8179ec

Observation d56dc019-7e13-4916-a1b9-0fc5d2e6b9de · outbound

This paper cites Optimal protocols for continual learning via statistical physics and control theory.

Optimal Training-Time Scaling in Gradual Adaptation Optimal protocols for continual learning via statistical physics and control theory

Reference 86

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.957586Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.241255Z digest=sha256:804e36fa58461664ba48ee7df36d8723abb7d524fd7ba221531b2fe760767e60

Observation 4c125bbc-2d11-4560-9ae1-8d9735659eec · outbound

This paper cites Efficient test-time model adaptation without forgetting.

Optimal Training-Time Scaling in Gradual Adaptation Efficient test-time model adaptation without forgetting

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.946293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.244578Z digest=sha256:228f35d4087c11f4de2cb3ce64bee53aa917f57ddc626be5567fa319dad7f36a

Observation 9024b7fd-7cfb-4e77-bffb-8f0dee92606a · outbound

This paper cites Towards stable test-time adaptation in dynamic wild world.

Optimal Training-Time Scaling in Gradual Adaptation Towards stable test-time adaptation in dynamic wild world

Reference 88

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.935798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.247976Z digest=sha256:912d195c296ac7a320e409300faab7eb6c8cb4549ab409af4d8d5e21ec031d85

Observation caa34e78-1ca6-4068-9c94-08a7db404630 · outbound

This paper cites Parisi, Ronald Kemker, Jose L.

Optimal Training-Time Scaling in Gradual Adaptation Parisi, Ronald Kemker, Jose L

Reference 89

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.251415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.251415Z digest=sha256:f30188413c7a1e23930dd1796bb092833592f87ea4542bce801e402dfe83d84a

Observation b67db30f-b1f7-4a22-aeaf-11f79d86a641 · outbound

This paper cites Wainwright, and Bin Yu.

Optimal Training-Time Scaling in Gradual Adaptation Wainwright, and Bin Yu

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.923543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.255359Z digest=sha256:72c44745e3cad617c90a6e37aeade1c75de8bbc4e8b8974af3363c3104459691

Observation faae7793-d04e-4acb-80ce-fd7d74b5252b · outbound

This paper cites Online structured laplace approximations for overcoming catastrophic forgetting.

Optimal Training-Time Scaling in Gradual Adaptation Online structured laplace approximations for overcoming catastrophic forgetting

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.911521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.258513Z digest=sha256:a4aee0668b91fbca0de9cb3715c82dc79c3e1324105a3eef3eb47f10d4a35bd8

Observation 3308c1a8-2a0b-4782-8e7b-a878b8be4ffe · outbound

This paper cites Maximum classifier discrepancy for unsupervised domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Maximum classifier discrepancy for unsupervised domain adaptation

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.899814Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.261834Z digest=sha256:3b05bf13c4929e0011ff383b068aa6e3f6d558101b0b3af1f8cc80d16bb8b6e7

Observation 5c8119ec-326f-4eed-a9f3-daba2c567ee2 · outbound

This paper cites Test-time Adaptation in the Dynamic World with Compound Domain Knowledge Management.

Optimal Training-Time Scaling in Gradual Adaptation Test-time Adaptation in the Dynamic World with Compound Domain Knowledge Management

Reference 93

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:23:55.407862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.265348Z digest=sha256:d1e8d322b4bf8ca2c4edcad83d0d5847fa5951d25b1b307549bac5cc60a4c66f

Observation e85034c2-b5d1-4039-b690-45b880216757 · outbound

This paper cites Efros, and Moritz Hardt.

Optimal Training-Time Scaling in Gradual Adaptation Efros, and Moritz Hardt

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.887877Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.269323Z digest=sha256:575bf3373fcd32cdf4b391d504d448dd30522940353d51da07c111ef46ed1d44

Observation bcf1c10a-402f-4a9c-800b-24b2f3135d74 · outbound

This paper cites Ward, Mark Kong, and Halyun Jeong.

Optimal Training-Time Scaling in Gradual Adaptation Ward, Mark Kong, and Halyun Jeong

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.875838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.272837Z digest=sha256:dfe1bdc19a3fb9dc1f1739109fd5aa561bc034a66d45b6a65ff7080fb94ad5e6

Observation 26f07ae3-e241-4874-b052-b950743b43b7 · outbound

This paper cites van de Ven, Tinne Tuytelaars, and Andreas S.

Optimal Training-Time Scaling in Gradual Adaptation van de Ven, Tinne Tuytelaars, and Andreas S

Reference 96

Resolution
unresolved
no resolver link, observed 2026-08-08T17:23:55.276142Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-08T17:23:55.276142Z digest=sha256:600b4dba74e150baf3299299ed3ce6a51764c03c890cfff35f1729a011fda46f

Observation a33c56d7-0b13-40a7-8a33-66bafc677b7c · outbound

This paper cites Tent : Fully test-time adaptation by entropy minimization.

Optimal Training-Time Scaling in Gradual Adaptation Tent : Fully test-time adaptation by entropy minimization

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.864000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.278984Z digest=sha256:35a0136644cd954cab44d2db20bb626460b9b9277a1b4e0c591569e7a09f6808

Observation 10bad797-3a3c-4d99-8dee-aeef6e47b5be · outbound

This paper cites Understanding gradual domain adaptation: Improved analysis, optimal path and beyond.

Optimal Training-Time Scaling in Gradual Adaptation Understanding gradual domain adaptation: Improved analysis, optimal path and beyond

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.852148Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.282327Z digest=sha256:4dc67fbca5d79ba83c03df116be62bab75975444c71719eede0a05817e49dc42

Observation 11742923-6c24-40fd-85ab-99d85f0cfec5 · outbound

This paper cites Continual test-time domain adaptation.

Optimal Training-Time Scaling in Gradual Adaptation Continual test-time domain adaptation

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.840391Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.285133Z digest=sha256:32f0982e8717ae99b1b6b030e1b00fe78f6b1ad61b7b02b238f4f28211c6b8e3

Observation fb8385d6-3e23-422f-9920-78cd65e8e0d7 · outbound

This paper cites Curriculum learning by transfer learning: Theory and experiments with deep networks.

Optimal Training-Time Scaling in Gradual Adaptation Curriculum learning by transfer learning: Theory and experiments with deep networks

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-08T17:23:55.828939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.288069Z digest=sha256:236e1967c29266ce9894f22312720b3852639fe540ce55ebf5454750dfefdb9e

Observation 1d3e465a-9131-4ccf-b05b-63a3a0103293 · outbound

This paper cites From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning.

Optimal Training-Time Scaling in Gradual Adaptation From Order to Distribution: A Spectral Characterization of Forgetting in Continual Learning

Reference 101

Resolution
verified exact
local_arxiv, observed 2026-08-08T17:23:55.748087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=arxiv_source observed=2026-08-08T17:23:55.291013Z digest=sha256:700c737fb8c47357a2379bd0eb6a3f7ccffaeb010f1329056ac3e1f9b4efa833

Pith citing papers

No inbound Pith citation observations are available.