Pith. sign in

Paper Citation Record · LEDGER

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining

As of 8 August 2026, this Paper Citation Record lists 61 of 61 outbound references and 0 inbound Pith citation observations for arXiv:2505.19893.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19893 v1

Coverage vector

measured 61 of 61 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:09:51.114106Z

measured 61 of 61 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

61 of 61 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved44
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c5e4565a-faa7-4ee1-b74f-98c0c8a8b8f0 · outbound

This paper cites A Survey on Data Selection for Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Survey on Data Selection for Language Models

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:44.390210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:44.390210Z digest=sha256:a1b292e2f373cb77650ef1b27425c237431890218a137673df6e4fbb87abedfc

Observation 22613bfc-4a69-440c-a367-0dd88544161b · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.859226Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.493543Z digest=sha256:d9ff970b4cb86b91ec5ec300ac4bc8351d4096cd5ddf5e9671481c3f351646de

Observation 41a33a96-71a8-4589-96cf-cafbfb7f13cb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.617607Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.602907Z digest=sha256:16176be67a59064c5a60e63254018a10f5c9364b8a45f5ce027862e096c4af33

Observation 6385af8d-a6d1-45f2-9e05-cfea51f8e003 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.436143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.698697Z digest=sha256:a2a20fbd7d917cb8586f2f19e08838186b31fe4f8ec1be7f32fe3fcecc49b89f

Observation 7ccbd500-5237-46e3-b4fc-794627f6b1af · outbound

This paper cites D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining D., Dhariwal, P., Neelakantan, A., Shyam, P., Sastry, G., Askell, A., et al

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:57.223478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.821958Z digest=sha256:7c3d3589d83d498005dff9de11323967f549f16c74f774c13c75426609d6d8b5

Observation f016df29-0942-466c-abcd-bea4ff47626a · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:57.099697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:44.948297Z digest=sha256:b83a3cc5b8df6ab8dd319da5d1a68f96a5d6273405f65fa1f22ae28ff9865e7a

Observation 04200d5a-070a-4966-ab4a-f4170eb4c6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.914899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.043318Z digest=sha256:2599099280ad488fed2320ed3535b72b88fe1e61e6075911fb7bf336e20be428

Observation 5abf293f-23b7-41ad-801e-d5bc083f1928 · outbound

This paper cites W., Sutton, C., Gehrmann, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining W., Sutton, C., Gehrmann, S., et al

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.759624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.094544Z digest=sha256:24d5537a28ef5bb4c38aaa4a1367f8a8d8151f84b0f4251964a9089d3db34737

Observation 84469261-e86d-4aac-a1c0-eb5e97ff0784 · outbound

This paper cites Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Think you have Solved Question Answering? Try ARC, the AI2 Reasoning Challenge

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.215581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.215581Z digest=sha256:82c4401f0db9c308bf4f207d2909045b4a0e74484998b074475c212b1964c064

Observation 11b824f8-893b-429e-9bcc-ade2347145a8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:56.580390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.331575Z digest=sha256:ae9ab3467abbcd5dc61c1bfc8b645d49a65748cd3ad431457ca8b271e31d3ae2

Observation 6ccd29a6-41ae-4e3d-93d3-daa5aa0ed334 · outbound

This paper cites Y., Jegelka, S., and Krause, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Jegelka, S., and Krause, A

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.454906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.511080Z digest=sha256:7e4282286b4a17734606e87fbb62dc9a80519f11e0c2a1ce6ad0ea31c6e6de19

Observation a3f0f741-599d-4cd8-8ee9-9136b1a2ec0a · outbound

This paper cites FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining FlashAttention-2: Faster Attention with Better Parallelism and Work Partitioning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.618244Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.618244Z digest=sha256:6c5d6f9600ef0697375468c1ad5d4fa9fd09e47c255fbb44e64962be8a70c406

Observation 20c14fee-ce53-41ab-ab0b-a495f8a75a7c · outbound

This paper cites and Namkoong, H.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Namkoong, H

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.317730Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:45.723599Z digest=sha256:467a3914fb747838656797519374176339b5325afb7b57cce048281cfe8f64cb

Observation 0216d8e9-b967-4926-bfba-5ea684dea030 · outbound

This paper cites Irreducible Curriculum for Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Irreducible Curriculum for Language Model Pretraining

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.871063Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.871063Z digest=sha256:1742495c6ce293b69354aef536513be755633eba89e8d04c8cad90f1890cc617

Observation eb274786-b2dd-4323-ac48-42b72e660c48 · outbound

This paper cites DoGE: Domain Reweighting with Generalization Estimation.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining DoGE: Domain Reweighting with Generalization Estimation

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:45.950397Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:45.950397Z digest=sha256:f9563b2410fd3ca5c600dbb7a75a4183d4f57c8d1afd9c2bd964ec978cc70f5d

Observation 93504c1f-4915-4a9e-bf7f-2902b47c8450 · outbound

This paper cites and Dayan, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Dayan, P

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.154815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.064723Z digest=sha256:20a726b580e9c8bc9c4f28ecf54c04c770f4ceaa4094ad35ae974e160f2600b2

Observation 3614366c-30a7-4828-934b-cf16d6ae7b00 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.163491Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.163491Z digest=sha256:756af37df79a55b30811532072fb4888257fb5c050eba8f3a488ad2c0bdab3e2

Observation c51a7a9a-970e-4891-8482-a96b972bd708 · outbound

This paper cites and Cohen, V.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Cohen, V

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:56.029767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.335902Z digest=sha256:f0e75b1994e0a4be62a7d62e654a8018299495493df2637ee0bac72cc2701353

Observation 163cdd05-cc9b-4096-b917-be7194f12bd5 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distilling the Knowledge in a Neural Network

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.458825Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.458825Z digest=sha256:698a5de4ffe96febe7f482e8e072015b735faa0e5c343bcf59d42455df8d24d7

Observation 01b09708-2514-4827-8aed-eeba5c7b2e94 · outbound

This paper cites Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Y., Zhou, T., Wu, Y., Song, X., Song, X., and Zhou, D

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.823606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.587696Z digest=sha256:725524e2586189018f590e8928daf784d073cf94f8b2f85378c6af9a07363b4f

Observation 53d10d34-f94a-47f0-a91e-55cb4955404b · outbound

This paper cites and Waegeman, W.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Waegeman, W

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.713461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.713461Z digest=sha256:eb233c46da49ccedac078f28919e13f9aea7d77390375d3ec91cba0109e5f288

Observation b20dc414-e6ce-4934-bfc4-b9b438aaf92f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 22

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.665120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:46.822031Z digest=sha256:4eb663010267a95da1001e7540e490e55f76a4464d817cfaf84d0159636b3bf1

Observation 5ed818e2-10a8-43cd-b75e-02687ffcc794 · outbound

This paper cites Accelerating Deep Learning by Focusing on the Biggest Losers.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Accelerating Deep Learning by Focusing on the Biggest Losers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:46.918110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:46.918110Z digest=sha256:54b44e7097c5ffe605a718e81d09a1135772c737019c7346d4bc688369886c69

Observation 85097b5e-ad96-4b9d-889e-cb1715872737 · outbound

This paper cites TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining TriviaQA: A Large Scale Distantly Supervised Challenge Dataset for Reading Comprehension

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.012470Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.012470Z digest=sha256:f9d85b4d47eae53478cdd7e96302345361e4cf791b076e78006d127233ca5a5b

Observation 0a94f43e-7514-4392-9c0c-d167e1717feb · outbound

This paper cites Scaling Laws for Neural Language Models.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Scaling Laws for Neural Language Models

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.106847Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.106847Z digest=sha256:8addbf161e6473666b7a80cd1a8612dab3ae7b0dfbe09dc39997299c78b8366e

Observation b6f1a256-586c-4997-9bac-88b6d80a36fb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.474628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.274894Z digest=sha256:982f3ce9dac135e14cee0927ea5b9ad05bc25604b33b899d52f15a9082804850

Observation a39e50b7-01c8-40ed-872f-321d1da91158 · outbound

This paper cites and Fleuret, F.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Fleuret, F

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:55.305333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.388011Z digest=sha256:e190f854b483d785dd595d5a0d2f3f0af925a9ed1d8eb22c8f00b06feb4a20bc

Observation 7d6b3ba7-5fb1-4090-8fb2-74f628c16113 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:55.101657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.497294Z digest=sha256:d9097bdc842e543a5a70c5978337d2153e571eebbd51deca2de541680e614689

Observation 2b0f463c-7082-4204-bf9b-f3f0b6ff0715 · outbound

This paper cites and Rush, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Rush, A

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.950509Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:47.643905Z digest=sha256:38c6ca46bffcf1b237b3851f9acbaf0df134b315ea71ad9ccd1b83b7ab900e27

Observation c4b5ab01-bbb7-4392-9d9b-ba946e937fd5 · outbound

This paper cites Distributionally Robust Optimization.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Optimization

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.837432Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.837432Z digest=sha256:ce9d5f768bf19599ba08fd134f5ff3d1eda44debbd1a65681daaf3a1cdc855ba

Observation e59e25b1-d230-4505-b4dd-5b381d0f9aad · outbound

This paper cites Rho-1: Not All Tokens Are What You Need.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Rho-1: Not All Tokens Are What You Need

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:47.959331Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:47.959331Z digest=sha256:ba57596057370c4b7d9e0128a2d1624b91359606fd6f430e2814e0f4cb0cc5d9

Observation d40f7a1d-b4bb-4e0e-98f7-9b80fde34623 · outbound

This paper cites Online Batch Selection for Faster Training of Neural Networks.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Online Batch Selection for Faster Training of Neural Networks

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.082527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.082527Z digest=sha256:5d79718f8a26550ed87ca43e17f9d108cf11de62ee41d0462d40e940adc8d2db

Observation 0d4dafe6-fe6e-433a-abf7-9a58b36b6db5 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 33

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.743834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.197734Z digest=sha256:132554e500303caaccb929eee84d1713ee0d81cac8db33ff63470d6dcf54bc39

Observation ea4f463e-8f5a-43cf-a59f-4cf597277efc · outbound

This paper cites When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining When Less is More: Investigating Data Pruning for Pretraining LLMs at Scale

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.284205Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.284205Z digest=sha256:5a1a3bd6fc6c434d4cc26462f7f5f7183e7a85977be78cf04850be998214bcb2

Observation 7fcf85e1-298e-4c01-8cb1-422e54e41ba0 · outbound

This paper cites LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining LLMs on the Line: Data Determines Loss-to-Loss Scaling Laws

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.398461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.398461Z digest=sha256:cae38c1bf3d6fb9f8e2c3292e91c8c4310e7fe6abeeb8322e9c419eba43063e4

Observation a08605b6-4ef4-4c92-93e3-f997ae50026e · outbound

This paper cites Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Can a Suit of Armor Conduct Electricity? A New Dataset for Open Book Question Answering

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.521179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.521179Z digest=sha256:c900515d070cf161f28bd45c4e885463489319a2a896f8dfe16018f1108885a0

Observation 00e1647d-6d87-4f44-a0c2-21413d663041 · outbound

This paper cites M., Razzak, M.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Razzak, M

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.610272Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.648493Z digest=sha256:c1e43a497fe0baeec38dae6052836cbf695ebe04e3d29ea820dcd3f6831bc602

Observation 2d482e1e-045c-4946-b5d6-fea337de7d4c · outbound

This paper cites Distributionally Robust Language Modeling.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Distributionally Robust Language Modeling

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:48.756530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:48.756530Z digest=sha256:e2b9960f148cc30af795094b0bd043efdc094ab68bf63ec92b947c1cae6ddf4b

Observation 84d475bd-9238-4f02-ae19-0b95e44c49b0 · outbound

This paper cites Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Q., Bernardi, R., Pezzelle, S., Baroni, M., Boleda, G., and Fernandez, R

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:54.449216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:48.909920Z digest=sha256:0291e2aec2f09f266ad269dda0b59613af15e9fc77a587a5431825e10d82b5fb

Observation 48ae6b37-6f3a-4ac5-af48-a32d7dc34571 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.285715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.043623Z digest=sha256:3e50666b7cf3920162967256eb70d9c9b461e1ef39c40686d5b6df2090c3f8e9

Observation 2f9ecefc-1c9e-4d9a-9e06-67ac405c031d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.175622Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.175622Z digest=sha256:a30207fc3582e56e1f8dbe44f2f534f254d513b7c217d053325b031d51f73f42

Observation 2d93fa79-547f-4789-a82b-dc4b0615d23d · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:54.132135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.330327Z digest=sha256:e1f3b374db2843d22b088c8f60fbb6ceb546e3e19ca96303554f60f4d8b467cd

Observation 70e14c5c-ba69-43c5-9b61-afb4f22427b7 · outbound

This paper cites A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining A Little Help Goes a Long Way: Efficient LLM Training by Leveraging Small LMs

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.458727Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.458727Z digest=sha256:08cc40170d8bfb12054061e94f59de7a9875602404b8d0ee63a3f536ac269e34

Observation 270d6c88-c103-4b63-bee1-10adb4d7679f · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.931666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.569016Z digest=sha256:4897de7be23694c26c52b99c608f1334558d10b78569f9b1a6af6c18d4f5bf28

Observation ade882c0-2caf-4109-a564-bdd621b92250 · outbound

This paper cites T., Uryasev, S., et al.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Uryasev, S., et al

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.660065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.660065Z digest=sha256:e928b3aac25316db621e95e3790634d0bdd924d86a21ff09608558992efa24f4

Observation 7be11a93-fc4f-456a-b1ce-c09431db1d5d · outbound

This paper cites How to Train Data-Efficient LLMs.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining How to Train Data-Efficient LLMs

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:49.740308Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:49.740308Z digest=sha256:ff4d9eba259b7fd029689767fdc63123dd83fb362157b98aee0a1c05b44c3560

Observation 85769876-7f40-42db-93e9-6a6007ecca82 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.774772Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.807202Z digest=sha256:c182e59cb1425cedc24d14758a7df96950b16e4bedea625d1827d8f91c42237b

Observation d6d22732-e402-401b-b4a1-1c824496e6af · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.596269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.891517Z digest=sha256:f7c105e1af239023cb4ba23f7dce81dd871a18c00d65821f5a19f90728c5b3ee

Observation 5bd8cad6-0366-4a12-884c-72c61a5d7d73 · outbound

This paper cites R., Hestness, J., and Dey, N.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining R., Hestness, J., and Dey, N

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:53.460624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:49.980510Z digest=sha256:2731f12a71f5c5020f30e5ddadf457042abe646d1a5b47c834f680d004a58c6b

Observation 457b51ea-2390-4ed0-8f05-fefe992ccef0 · outbound

This paper cites Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Dynamic Loss-Based Sample Reweighting for Improved Large Language Model Pretraining

Reference 50

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:09:51.406696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.090492Z digest=sha256:bd454527543ab84217291278aa9d7420581a8f32018137a768df1f8b5d65e2f7

Observation 27c94bad-4712-4b5e-b319-730577f541bb · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.320047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.182360Z digest=sha256:7ad8bb8983fb5612afa36f514b3eb0ca0d0a478f4811b6bc88ffd4cb09ea1e37

Observation 4fe5b468-ad36-4369-9851-690c6dd90719 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 52

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:53.099622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.257296Z digest=sha256:a6659ccb3589ca7615a9c32243daed4772eaa49b34c733cd8dbff422a7c71c59

Observation 31b392eb-f669-404e-9718-ca12259c90d1 · outbound

This paper cites T., Wu, T., Song, D., Mittal, P., and Jia, R.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining T., Wu, T., Song, D., Mittal, P., and Jia, R

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.930135Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.366434Z digest=sha256:a41ae7e3fddb9e172c1a129a28ed5b2b65969144eb39e3391c5ffa9ed1e80c6c

Observation b3c893cf-3193-42c0-a86a-28d011a4c289 · outbound

This paper cites Crowdsourcing Multiple Choice Science Questions.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Crowdsourcing Multiple Choice Science Questions

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:50.466643Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:50.466643Z digest=sha256:67929bd8f4d5c1f6cc0c3f20b48ed4cf82a5b8b65bb4b73020e06a7261787616

Observation 9b377001-9843-44c5-8f08-6186730925c2 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.752364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.566769Z digest=sha256:6e9f9dc025d5aff3ed9ff17dd3d5b6d5f8810c7c15763f224659ed2deb6bf914

Observation 6cdf1910-301f-4cf2-8fa3-950e08d3ead0 · outbound

This paper cites and Menon, A.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining and Menon, A

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.566666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.667952Z digest=sha256:ba75e47214f3b882d22fd574b9e79cc7710b327d20a406dcbc4ca90b00775875

Observation 0d5f2cf0-b35a-49d8-8c0c-859f1d51c1b8 · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:52.388361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.756681Z digest=sha256:adedad34503f908a97e3b6c73b7d91ea9ff41bb2c652667ac0547a4fb444fb0d

Observation 02ba56f4-1ed8-4022-b299-a43b1f23055e · outbound

This paper cites M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Pham, H., Dong, X., Du, N., Liu, H., Lu, Y., Liang, P

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.227595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.817490Z digest=sha256:bec44b1b83f74da196aa0cc7cac67f7651e2d1702143c3f35c105050329189ad

Observation a11d0f2b-811d-4186-bc3f-cf0a2362e079 · outbound

This paper cites M., Santurkar, S., Ma, T., and Liang, P.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining M., Santurkar, S., Ma, T., and Liang, P

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:09:52.067412Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:50.936488Z digest=sha256:4e65438b64cace6ea79430e0108e76c521cdf30ddc773ccfb00dae15d64ee37f

Observation f8237699-5b1f-4ffb-93f9-943734fc649c · outbound

This paper cites an unresolved cited work.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:09:51.888153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-07T14:09:51.032016Z digest=sha256:04d20b67683e0140181f9a385602b9b3c7185246d089c798b60d0178c65d9adb

Observation dea4ee42-38ff-4b32-9c67-a4b012fa56b0 · outbound

This paper cites HellaSwag: Can a Machine Really Finish Your Sentence?.

ESLM: Risk-Averse Selective Language Modeling for Efficient Pretraining HellaSwag: Can a Machine Really Finish Your Sentence?

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-07T14:09:51.114106Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:09:51.114106Z digest=sha256:74fa664ceecc19b3d7eb1da6db48d902353f1c2f7a493bd9306fdada34c83d7c

Pith citing papers

No inbound Pith citation observations are available.