Pith. sign in

Paper Citation Record · LEDGER

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining

As of 10 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19893.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19893 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:51.114106Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c5e4565a-faa7-4ee1-b74f-98c0c8a8b8f0 · outbound

This paper cites A Survey on Data Selection for Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Survey on Data Selection for Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:44.390210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:44.390210Z digest=sha256:6122f3356680a01c6c368b7ca4b725836f46f8beb829b5232ef288cf9fa2d8b9

Observation 22613bfc-4a69-440c-a367-0dd88544161b · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.859226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.493543Z digest=sha256:f6c7ac12acc56bd361a3c2ab06745b57bee4b6f2f3fc382fe519ded6abe52c8a

Observation 41a33a96-71a8-4589-96cf-cafbfb7f13cb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.617607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.602907Z digest=sha256:a742f0e307ffb0f3d2cb9a115f8cfc85598c5c98e2b7f9e3377ccd137669a0b0

Observation 6385af8d-a6d1-45f2-9e05-cfea51f8e003 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.436143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.698697Z digest=sha256:aa8f697482a013ec010143d1e7705031206538383b3a3f1e7c9842212894959f

Observation 7ccbd500-5237-46e3-b4fc-794627f6b1af · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:57.223478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.821958Z digest=sha256:cc142e9f72a19c9551e2f3a94fc8b9325bd96d2dc6cd6f86df140c1b1304778e

Observation f016df29-0942-466c-abcd-bea4ff47626a · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.099697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.948297Z digest=sha256:000a11e612008adec4bf5b83e3a72b3d2acd668d7f2d55ff5f428c07ec8fcda6

Observation 04200d5a-070a-4966-ab4a-f4170eb4c6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.914899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.043318Z digest=sha256:498f2ba480fe449b96ae4a27a75b1f618ed2a99d466884d9fdd4cfe02ba24d26

Observation 5abf293f-23b7-41ad-801e-d5bc083f1928 · outbound

This paper cites W., Sutton, C., Gehrmann, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining W., Sutton, C., Gehrmann, S., et al

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.759624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.094544Z digest=sha256:64c20b1d7d3b72bc676eba2b41b18d76cb8f5576abe56f6f925d19139864f57c

Observation 84469261-e86d-4aac-a1c0-eb5e97ff0784 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.215581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.215581Z digest=sha256:0e450ca8c6f247f3549a44e578f857683de903bf534b8c67710f38ad62375b4f

Observation 11b824f8-893b-429e-9bcc-ade2347145a8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.580390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.331575Z digest=sha256:a3f3e9cd6c05d327a98bbbc198c96445e66fb16c3d1dc09ed43381f46f676b26

Observation 6ccd29a6-41ae-4e3d-93d3-daa5aa0ed334 · outbound

This paper cites Y., Jegelka, S., and Krause, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Jegelka, S., and Krause, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.454906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.511080Z digest=sha256:f4876fcbd9e24e405e37ee6d87bfb31eb6464cb8a70544edaecb1c9ff7c79d48

Observation a3f0f741-599d-4cd8-8ee9-9136b1a2ec0a · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.618244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.618244Z digest=sha256:59c64ee36c5bd64f8dc24219e86bf75d28d8998446f2d02208bbfc9ba763ec88

Observation 20c14fee-ce53-41ab-ab0b-a495f8a75a7c · outbound

This paper cites and Namkoong, H.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Namkoong, H

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.317730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.723599Z digest=sha256:bb4256438ef0b7d478f21c29c59c2194254310fe4640fe35c58cd1dc3d5af75f

Observation 0216d8e9-b967-4926-bfba-5ea684dea030 · outbound

This paper cites Irreducible Curriculum for Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Irreducible Curriculum for Language Model Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.871063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.871063Z digest=sha256:2154b35eabe61e3153363208440a7b039701f377be98970d3f1d625613344408

Observation eb274786-b2dd-4323-ac48-42b72e660c48 · outbound

This paper cites DoGE: Domain Reweighting with Generalization Estimation.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining DoGE: Domain Reweighting with Generalization Estimation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.950397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.950397Z digest=sha256:3aaf70b6dc4c7f17732a8c7ab1318c5bbd473c789977a1abdfbd8c9df08dcf74

Observation 93504c1f-4915-4a9e-bf7f-2902b47c8450 · outbound

This paper cites and Dayan, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Dayan, P

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.154815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.064723Z digest=sha256:0cdac6a9b4ac229257bb591aaa72501e236faaac734a47b92f9032bb5e5ab584

Observation 3614366c-30a7-4828-934b-cf16d6ae7b00 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.163491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.163491Z digest=sha256:9a62e96f8ee8ee535a0db57581d47ba6c9540505165c4190481a3f59032b90eb

Observation c51a7a9a-970e-4891-8482-a96b972bd708 · outbound

This paper cites and Cohen, V.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Cohen, V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.029767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.335902Z digest=sha256:b39fa220942d3951b1797dc3c6b17d253733bf611ff433bf6d5e090efe2bb0de

Observation 163cdd05-cc9b-4096-b917-be7194f12bd5 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distilling the Knowledge in a Neural Network

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.458825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.458825Z digest=sha256:33d3026143959e351190168b5acb354bde663dcf48dd5ecc546a60432d9fd969

Observation 01b09708-2514-4827-8aed-eeba5c7b2e94 · outbound

This paper cites Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.823606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.587696Z digest=sha256:67e48edf92b411a744587ea12071d4e621c1626930d8b5d741d2d357783189ac

Observation 53d10d34-f94a-47f0-a91e-55cb4955404b · outbound

This paper cites and Waegeman, W.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Waegeman, W

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.713461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.713461Z digest=sha256:647140103876cd09566420b93db91992ef41489c966ecc80d1e2e4c5a4d6a292

Observation b20dc414-e6ce-4934-bfc4-b9b438aaf92f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.665120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.822031Z digest=sha256:f231af13d48b67f26749c5bd425e541f52d6435d42e5fb4758cc5b170b7057b4

Observation 5ed818e2-10a8-43cd-b75e-02687ffcc794 · outbound

This paper cites Accelerating Deep Learning by Focusing on the Biggest Losers.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Accelerating Deep Learning by Focusing on the Biggest Losers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.918110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.918110Z digest=sha256:6dce8a28c0ccc79b1b39ba6cb232959edb91baef4552dfecffc608323c582388

Observation 85097b5e-ad96-4b9d-889e-cb1715872737 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.012470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.012470Z digest=sha256:b18303114a94aa7d84a5a6f8ac5fd8189921d399bda1678dca4eaf91aa71b3e3

Observation 0a94f43e-7514-4392-9c0c-d167e1717feb · outbound

This paper cites Scaling Laws for Neural Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Scaling Laws for Neural Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.106847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.106847Z digest=sha256:d06d461a0facbf0b72145c944fbc65dff93623ac02af5e7f4602c8a712df6506

Observation b6f1a256-586c-4997-9bac-88b6d80a36fb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.474628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.274894Z digest=sha256:04726743294098dcb0e90a9d5fd8ea8eccfc8c4053dff6b8d1849b11baedae70

Observation a39e50b7-01c8-40ed-872f-321d1da91158 · outbound

This paper cites and Fleuret, F.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Fleuret, F

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.305333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.388011Z digest=sha256:e2a22538ccdf0257ad91b5ed4ccb663a9a9dc8ae6f6e26f6c14b4605c71e0638

Observation 7d6b3ba7-5fb1-4090-8fb2-74f628c16113 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.101657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.497294Z digest=sha256:f6444d929615fa4ecb1c9a15c517cf25c2eff396121e628a6ed0182ee57998b0

Observation 2b0f463c-7082-4204-bf9b-f3f0b6ff0715 · outbound

This paper cites and Rush, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Rush, A

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.950509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.643905Z digest=sha256:225bf8e8adb283c8ea401fd3c1c37f5e0ae44249dffd459819744bcb988037da

Observation c4b5ab01-bbb7-4392-9d9b-ba946e937fd5 · outbound

This paper cites Distributionally Robust Optimization.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.837432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.837432Z digest=sha256:769aeef8a09758eaad31cd1fc287601ce375e7a33d02b869c824705697fbd476

Observation e59e25b1-d230-4505-b4dd-5b381d0f9aad · outbound

This paper cites Rho-1: Not All Tokens Are What You Need.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Rho-1: Not All Tokens Are What You Need

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.959331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.959331Z digest=sha256:3a2d8d4a4ac5534adf2536ec692476077a41c81885cc7ffff61f60d775e667c7

Observation d40f7a1d-b4bb-4e0e-98f7-9b80fde34623 · outbound

This paper cites Online Batch Selection for Faster Training of Neural Networks.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Online Batch Selection for Faster Training of Neural Networks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.082527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.082527Z digest=sha256:ad844d4a55f23b046441321379d77a6e2a6ca45d70416f4adc4611b587376be3

Observation 0d4dafe6-fe6e-433a-abf7-9a58b36b6db5 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.743834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.197734Z digest=sha256:520d626071d411ecc6b499ea5d1734ae1dfb3357e01d865c3ba013a2374fc724

Observation ea4f463e-8f5a-43cf-a59f-4cf597277efc · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.284205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.284205Z digest=sha256:90ce270949d82fe9145a15a721b92fa578519990af8fdaa3228a5452967074d9

Observation 7fcf85e1-298e-4c01-8cb1-422e54e41ba0 · outbound

This paper cites LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.398461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.398461Z digest=sha256:1c4cec97423600340d3d10d2dbc1cb2878e9b4d39bf93475b2c0aea7c6f3caec

Observation a08605b6-4ef4-4c92-93e3-f997ae50026e · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.521179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.521179Z digest=sha256:0f27ef5baebf67550b2a8880a11bef02c2470ce7b4de863e94239644887c0131

Observation 00e1647d-6d87-4f44-a0c2-21413d663041 · outbound

This paper cites M., Razzak, M.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Razzak, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.610272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.648493Z digest=sha256:a830aa031512982b960782594506cb50d7717a6388d50355a3481538783b9946

Observation 2d482e1e-045c-4946-b5d6-fea337de7d4c · outbound

This paper cites Distributionally Robust Language Modeling.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Language Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.756530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.756530Z digest=sha256:f2fe13765a833e61e7d266353f82d3284e2b2325a5a845f12c4834763e95f7bd

Observation 84d475bd-9238-4f02-ae19-0b95e44c49b0 · outbound

This paper cites Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.449216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.909920Z digest=sha256:094a31c7298e85479ba25956c70e8f64397ebf1882e20a2ee7d5124ab1c36f71

Observation 48ae6b37-6f3a-4ac5-af48-a32d7dc34571 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.285715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.043623Z digest=sha256:422bcf58afb1b7c5672ba68532da04cb3a4ae744259dfd10db3fdb8471593515

Observation 2f9ecefc-1c9e-4d9a-9e06-67ac405c031d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.175622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.175622Z digest=sha256:1d13c5f97ee7fe8d52a19bfe06badcc8417b4db50944145b61d516a45cc67b4a

Observation 2d93fa79-547f-4789-a82b-dc4b0615d23d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.132135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.330327Z digest=sha256:c953ab9d64e36c581a8b547c1a491ec297647667d65ed5428d3e37ae626b6117

Observation 70e14c5c-ba69-43c5-9b61-afb4f22427b7 · outbound

This paper cites A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.458727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.458727Z digest=sha256:99c681fba8a6fd84f91985067845d828267ee615d73f2c28521f8ab5e24f2b05

Observation 270d6c88-c103-4b63-bee1-10adb4d7679f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.931666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.569016Z digest=sha256:8639d6d0fde3390cf49b9b378da4a00c1a6775c6e938c5a00e9fe6a7f3084dff

Observation ade882c0-2caf-4109-a564-bdd621b92250 · outbound

This paper cites T., Uryasev, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Uryasev, S., et al

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.660065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.660065Z digest=sha256:9743ce1e144a2fb677150801bd6a77ea2a9493a8f31b6cc40e5e19b1789e178f

Observation 7be11a93-fc4f-456a-b1ce-c09431db1d5d · outbound

This paper cites How to Train Data-Efficient LLMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining How to Train Data-Efficient LLMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.740308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.740308Z digest=sha256:358cfcd94eafb3cae2ad5ec158af8d4238c7b0c6162ef1e63012379235970d07

Observation 85769876-7f40-42db-93e9-6a6007ecca82 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.774772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.807202Z digest=sha256:ca0607ec905ead1ee28687897f73856b1ec50d9512f45b2f3acb4d926205c9fa

Observation d6d22732-e402-401b-b4a1-1c824496e6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.596269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.891517Z digest=sha256:c8704ec9a7c39fc5816e4884fa70bfb5b5b2cde207d4e694a8b837a581a32762

Observation 5bd8cad6-0366-4a12-884c-72c61a5d7d73 · outbound

This paper cites R., Hestness, J., and Dey, N.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining R., Hestness, J., and Dey, N

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:53.460624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.980510Z digest=sha256:6d1d9ea5f3f26b175d7272e7483dfe92dab3eeeb5baef0c26f13938d11bb8b1b

Observation 457b51ea-2390-4ed0-8f05-fefe992ccef0 · outbound

This paper cites Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:09:51.406696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.090492Z digest=sha256:81c695f52b7127fda2d32747b3b93360961f8cdee107a8e953326ed13805be52

Observation 27c94bad-4712-4b5e-b319-730577f541bb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.320047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.182360Z digest=sha256:4ebb33ecd30a611ee28e356c707307262969e314091d1273ce9d97d82e1da0a5

Observation 4fe5b468-ad36-4369-9851-690c6dd90719 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.099622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.257296Z digest=sha256:435c6a5319af43c50d16290dbccbe2ace32be4b727d2fbfe2592ed2b7dc5dfcc

Observation 31b392eb-f669-404e-9718-ca12259c90d1 · outbound

This paper cites T., Wu, T., Song, D., Mittal, P., and Jia, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Wu, T., Song, D., Mittal, P., and Jia, R

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.930135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.366434Z digest=sha256:d79c23000b97a1c7893bd136d8afc025c099d5f8e616704b8e83208fa6c1a474

Observation b3c893cf-3193-42c0-a86a-28d011a4c289 · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Crowdsourcing Multiple Choice Science Questions

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:50.466643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:50.466643Z digest=sha256:b92923233aefc8a4345dab77927d4b249d37113309eaa45d7867f712ca1b8f4b

Observation 9b377001-9843-44c5-8f08-6186730925c2 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.752364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.566769Z digest=sha256:4e0c8a73181f4755654e8d876e74ebb48903a654e91258ba4e2e92e53bd9ba23

Observation 6cdf1910-301f-4cf2-8fa3-950e08d3ead0 · outbound

This paper cites and Menon, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Menon, A

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.566666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.667952Z digest=sha256:fde2f55579b69c7768d137bd254e73616bbfce4b3eb2ba7c080b6ff13db035fb

Observation 0d5f2cf0-b35a-49d8-8c0c-859f1d51c1b8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.388361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.756681Z digest=sha256:1efda55656a4bc5716bc6e0dbcac044b8b1b119e084488fa8a51ec0d7c93e94e

Observation 02ba56f4-1ed8-4022-b299-a43b1f23055e · outbound

This paper cites M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.227595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.817490Z digest=sha256:e6c3b4ed051d70483ecb9da5015c2f187e40537853c40ab4f8dce2eff5557f95

Observation a11d0f2b-811d-4186-bc3f-cf0a2362e079 · outbound

This paper cites M., Santurkar, S., Ma, T., and Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Santurkar, S., Ma, T., and Liang, P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.067412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.936488Z digest=sha256:741653b69c068769ce2f9bc9a0477c82065aa0b7853e2073a18a41ded6622863

Observation f8237699-5b1f-4ffb-93f9-943734fc649c · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:51.888153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-08-07T14:09:51.032016Z digest=sha256:edd45d10574a1ec936ac020f3b206191817cc0c8f20601eadac82f6c7157ed96

Observation dea4ee42-38ff-4b32-9c67-a4b012fa56b0 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining HellaSwag: Can a Machine Really Finish Your Sentence?

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:51.114106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:51.114106Z digest=sha256:194d20dd3224d33dea485d7a79348e05d5486f7270af33428c9540503d9e37fe

Pith citing papers

No inbound Pith citation observations are available.