Pith. sign in

Paper Citation Record · LEDGER

Model Connectomes: A Generational Approach to Data-Efficient Language Models

As of 22 August 2026, this Paper Citation Record lists 67 of 67 outbound references and 0 inbound Pith citation observations for arXiv:2504.21047.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2504.21047 v1

Coverage vector

measured 67 of 67 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-16T05:38:37.529590Z

measured 67 of 67 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

67 of 67 outbound references displayed

  • verified exact3
  • verified fuzzy42
  • unresolved22
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4d110ecf-2242-4105-9e6b-b208315bbc22 · outbound

This paper cites Love, Christopher J.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Love, Christopher J

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.247779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.420243Z digest=sha256:6bc716e9788accaf54d53b8e5f4b41d4db95258866143664b03b44ceb99cbd6f

Observation 9d9e3e3b-4847-4fb9-bf21-ddf2ef912d38 · outbound

This paper cites Language in brains, minds, and machines.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Language in brains, minds, and machines

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.235303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.540128Z digest=sha256:6156122d921f7b6b9fb02ced0f0e7f1ca68e37de98752f1197972067c1f5ce71

Observation f3c0130b-74fc-43f3-a6d2-25488aa5f18c · outbound

This paper cites Catalyzing next-generation artificial intelligence through neuroai.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Catalyzing next-generation artificial intelligence through neuroai

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.221108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.545447Z digest=sha256:c807ca9d476c02104bf39e0625031b894f3b384a94dabec7a554736b133a8e08

Observation 753ef396-f075-482e-ab59-4b48b438a06f · outbound

This paper cites A deep learning framework for neuroscience.

Model Connectomes: A Generational Approach to Data-Efficient Language Models A deep learning framework for neuroscience

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.105481Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.552767Z digest=sha256:e7a8764515f9ed4b1c54f645bd46a320f00de2c3ca837e6930b7d910ec9933df

Observation 2f68d6b6-ab98-4dba-b41c-bce855bbbe8a · outbound

This paper cites How learning can guide evolution.Complex Systems, 1(3):495–502, 1987.

Model Connectomes: A Generational Approach to Data-Efficient Language Models How learning can guide evolution.Complex Systems, 1(3):495–502, 1987

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.051332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.558090Z digest=sha256:93dddb143a764aacf3b0f0d0fefdd2fd1f8caaf1905637a0fe55ae1d5a23a4fa

Observation bc4cc8da-6aeb-4394-98af-a7acd95ef290 · outbound

This paper cites A critique of pure learning and what artificial neural networks can learn from animal brains.

Model Connectomes: A Generational Approach to Data-Efficient Language Models A critique of pure learning and what artificial neural networks can learn from animal brains

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.563530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.563530Z digest=sha256:53db8318a8efecfa9bd72bb6c5e266659b8fbc88e89fe1b79871046b77538165

Observation da27f7a5-dd02-463a-8263-9d1fc5cb8ac0 · outbound

This paper cites Direct fit to nature: an evolutionary perspective on biological and artificial neural networks.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Direct fit to nature: an evolutionary perspective on biological and artificial neural networks

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:40.019141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.570269Z digest=sha256:5bdd505406375fd1ab8c2b60081d2e81c8ff21b363d1ee29578ccd566bf7f9ba

Observation 88aea36e-2f99-4dac-a493-1c66489ba031 · outbound

This paper cites Gomez, Lukasz Kaiser, and Illia Polosukhin.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Gomez, Lukasz Kaiser, and Illia Polosukhin

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.921884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.576243Z digest=sha256:c31d81fe93def2665c951f6d973d239bc197f22a03354c5f0863addfdbd8b0fb

Observation d9908253-d2c3-40de-829a-e9363bfe50ca · outbound

This paper cites What artificial neural networks can tell us about human language acquisition.

Model Connectomes: A Generational Approach to Data-Efficient Language Models What artificial neural networks can tell us about human language acquisition

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.823314Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.580285Z digest=sha256:2995490057217134f1b8e67ea2120695063e42e21814694b98dcd2fe13e9f8df

Observation 08308dcd-dfbd-46f0-9853-da31175175f5 · outbound

This paper cites Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Call for Papers -- The BabyLM Challenge: Sample-efficient pretraining on a developmentally plausible corpus

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.585620Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.585620Z digest=sha256:502a1da7b7aba1c7db5c964fc9ae8941d041b513fdafbef8f8c5d23662b6b634

Observation c72c09ff-d846-4c8d-a138-2ab5634dd399 · outbound

This paper cites Improving language understanding by generative pre-training.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Improving language understanding by generative pre-training

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.591064Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.591064Z digest=sha256:a59524a603743d3833a209526c9f928599f0f37ccd4e072df3b8598c58a665fe

Observation 3d9cf620-6544-48d9-94a9-d4912dce77a1 · outbound

This paper cites Llama v3: Next-generation foundation language model, 2023.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Llama v3: Next-generation foundation language model, 2023

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.804369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.596106Z digest=sha256:9abd45bfc6c3784e00ba0bfde354e0ab98c8f6176a989f5b569b66ffda2890ed

Observation 55e3afe8-31e4-44d0-9d45-f3ea5bb30ce2 · outbound

This paper cites Deepseek llm: Advancing deep information retrieval with large language models, 2023.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Deepseek llm: Advancing deep information retrieval with large language models, 2023

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.791343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.600328Z digest=sha256:369367b9f54f31768331add20f8dc1beffb6fe94be22539c122a02cd026f1e0f

Observation d2090d3c-a12d-468d-8fb1-815e8621c80b · outbound

This paper cites Scaling laws for neural language mod- els, 2020.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Scaling laws for neural language mod- els, 2020

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.622434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.604836Z digest=sha256:46c127a1fe6184bfb9c4c8222a3fe677c27e5ebdf8d919db95524c4daa12886c

Observation c5264a8f-23a5-4f38-a242-0b978cead91c · outbound

This paper cites Incorporating context into language encoding models for fmri.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Incorporating context into language encoding models for fmri

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.609632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.609632Z digest=sha256:5271a478c0717416f71a006a89361692e8ab464fa1f9483bd7ee2da7cf890c69

Observation 2b132da2-3b35-46b6-a6d8-7c687630ddaf · outbound

This paper cites Interpreting and improving natural-language processing (in machines) with natural language-processing (in the brain).

Model Connectomes: A Generational Approach to Data-Efficient Language Models Interpreting and improving natural-language processing (in machines) with natural language-processing (in the brain)

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.613989Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.613989Z digest=sha256:19959a638b9bbd27ea98e313af45f456d3bf5891bd181e700a54e6cd26f2ae01

Observation 82448664-fc0b-415b-bcdb-15e898c57658 · outbound

This paper cites an unresolved cited work.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:38:39.550873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.617907Z digest=sha256:bd7cba04ea355409d9ba6f139556123cd8adf621d1bf415cd621fce484f35ae1

Observation abcc7647-97a8-4e40-b0c9-acce367a1d11 · outbound

This paper cites Brains and algorithms partially converge in natural language processing.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Brains and algorithms partially converge in natural language processing

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.538166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.622563Z digest=sha256:7070d1a3bd019e3156b0979b06f6887f82dd3f3adb0b24dc9fe20e4cbea487cd

Observation 2c8b4712-e539-44f2-a3d5-96b69d627e31 · outbound

This paper cites Shared computational principles for language processing in humans and deep language models.Nature neuroscience, 25(3):369–380, 2022.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Shared computational principles for language processing in humans and deep language models.Nature neuroscience, 25(3):369–380, 2022

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.526094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.627319Z digest=sha256:cf441aae2730827b2d7e60708455cba9d487d336d8821240a8faa4f1989c358e

Observation 28edc966-e31f-491a-a822-086dcf49fb6e · outbound

This paper cites Gptq: Accu- rate post-training quantization for generative pre-trained transformers, 2022.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Gptq: Accu- rate post-training quantization for generative pre-trained transformers, 2022

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.512898Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.632326Z digest=sha256:4d43ed4d4f5e71578433cd914fe24fca1a883bfc5bba6fa8cd1af708b7849e19

Observation b642b35f-a803-4871-a8d7-aed2b4319e68 · outbound

This paper cites Movement pruning: Adaptive sparsity by fine-tuning.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Movement pruning: Adaptive sparsity by fine-tuning

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.361650Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.637305Z digest=sha256:af1ea83e6cf1a7890f05e1b1d4366749dcd2fe1933874a10a044ccb2e2213d66

Observation a4b9809c-6559-487e-8b0d-bc197a322a83 · outbound

This paper cites Le, Geoffrey Hinton, and Jeff Dean.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Le, Geoffrey Hinton, and Jeff Dean

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.249143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.642097Z digest=sha256:6ee1b04dfa7dba18d3251e46bb1796d28e8a9a4460c3b9abab11bc65d12540c8

Observation 3e952e49-9bba-460d-8688-654f9f702a41 · outbound

This paper cites The lottery ticket hypothesis: Finding sparse, trainable neural networks.

Model Connectomes: A Generational Approach to Data-Efficient Language Models The lottery ticket hypothesis: Finding sparse, trainable neural networks

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.236828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.677212Z digest=sha256:a9a217c2fb37b3758bcf0d25e4c79b88277b2516e4d88ff4c0f526c503410443

Observation 6bfdd6fb-c911-4407-9c98-66f8344f728a · outbound

This paper cites Unmasking the lottery ticket hypothesis: What’s encoded in a winning ticket’s mask? In Proceedings of the International Conference on Learning Representations (ICLR), 2023.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Unmasking the lottery ticket hypothesis: What’s encoded in a winning ticket’s mask? In Proceedings of the International Conference on Learning Representations (ICLR), 2023

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.224674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.712872Z digest=sha256:787a7e2fb2104c4089059d4292f237d5cd320e5b251e7788c8696de722790d77

Observation 752bce38-be6d-43b6-975a-312552e59892 · outbound

This paper cites Lottery ticket adaptation: Mitigating destructive interference in llms,.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Lottery ticket adaptation: Mitigating destructive interference in llms,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.211602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.770876Z digest=sha256:ea6a27c378aafa36c0357aa18e56d348c5179f2fc9ea4e3bc3080841ac6416a2

Observation 8671004e-9552-460d-947a-5075dfc47a99 · outbound

This paper cites Lottery tickets in llms: Robustness to adversarial examples via binary masking.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Lottery tickets in llms: Robustness to adversarial examples via binary masking

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.021129Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.781083Z digest=sha256:79bcaa795863a4191b4492170e657ca3949d074bd5b9a3eb42000701f7fdb3dc

Observation 2e2c4ac0-bf54-40b8-8214-6af3c756bd95 · outbound

This paper cites Sparse winning tickets are data-efficient image recognizers.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Sparse winning tickets are data-efficient image recognizers

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:39.007385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.786648Z digest=sha256:1f82faaccbf4397ecc8df75f140b5f6946bf95aacae15d8e0316c774a53cc1ae

Observation 742e369a-c452-478c-bf6e-695260a217b9 · outbound

This paper cites Pruning neural networks without any data by iteratively conserving synaptic flow.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Pruning neural networks without any data by iteratively conserving synaptic flow

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.793208Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.793208Z digest=sha256:e26f48e534168abc08c665fa08994496c8b3558073fc742ba125b5dacd08ee5a

Observation e68d300f-80e5-4d0f-a432-d30459a21f65 · outbound

This paper cites Deconstructing lottery tickets: Zeros, signs, and the supermask.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Deconstructing lottery tickets: Zeros, signs, and the supermask

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.981519Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.800131Z digest=sha256:e75f43c31676cf977cec73dab0745c9fa078f85f8221886a08f49bce7f2ee450

Observation 837cedff-ab58-4cdc-96ef-101ee91b6335 · outbound

This paper cites Complex com- putation from developmental priors.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Complex com- putation from developmental priors

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.964389Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.809794Z digest=sha256:7657de0550683504ef3e1d90ebf7a3d58220f0176b17a3fca0fba56bcb506b2f

Observation 767c8859-9eac-4353-b896-ae8f3f747d0e · outbound

This paper cites Encoding innate ability through a genomic bottleneck.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Encoding innate ability through a genomic bottleneck

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.854759Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.854759Z digest=sha256:a27c7bc3bb35982fc2b88c382e09634efcd4fe50082f8be8671defaa4d2d2766

Observation 6a0d292c-1290-40b0-8d0f-2d059dd7fb94 · outbound

This paper cites Enhancing Interpretability using Human Similarity Judgements to Prune Word Embeddings.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Enhancing Interpretability using Human Similarity Judgements to Prune Word Embeddings

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-16T05:38:37.882693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:36.955859Z digest=sha256:a249f7e2cf400fa7c8dd9708ead11e682e3f14ab1d438dd908ccfbcbde2c4cd3

Observation 81a7c370-8c94-416e-b6f1-4b2f2e023429 · outbound

This paper cites Pruning sparse features for cognitive modeling.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Pruning sparse features for cognitive modeling

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.936037Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.040238Z digest=sha256:45acbe2d999a7306c49a853165e4b27ed718bcb557a8b121faf9d22b023caa5c

Observation 8b7f8df6-4d05-4ea5-82be-ef36ae693424 · outbound

This paper cites American parenting of language-learning children: Persisting differences in family-child interactions observed in natural home environments.

Model Connectomes: A Generational Approach to Data-Efficient Language Models American parenting of language-learning children: Persisting differences in family-child interactions observed in natural home environments

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.846134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.053054Z digest=sha256:b4892557fbf19ab51ccfbeb05bd186b378ed6f73273f25367aaa17dbee9dda06

Observation ad1ad28f-c94c-4d7f-b403-02b818144886 · outbound

This paper cites The fineweb datasets: Decanting the web for the finest text data at scale.

Model Connectomes: A Generational Approach to Data-Efficient Language Models The fineweb datasets: Decanting the web for the finest text data at scale

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.807054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.059027Z digest=sha256:bc8713f0b966f4fe40fdbada12b81446db7867e270c3892168ae039550dc7b87

Observation 0baff369-1489-43e2-a941-bd67b0aeddc8 · outbound

This paper cites modded-nanogpt: Speedrunning the nanogpt baseline, 2024.

Model Connectomes: A Generational Approach to Data-Efficient Language Models modded-nanogpt: Speedrunning the nanogpt baseline, 2024

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.063219Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.063219Z digest=sha256:a8a71419cabf143cac4e8cf8eeaf915451ee2e2f27eceb36c8f24c3618f02c82

Observation aed2d034-f9c5-4ff1-b6f5-37b5ceae8c62 · outbound

This paper cites Hellaswag: Can a machine really finish your sentence? In Proceedings of the AAAI Conference on Artificial Intelligence, pages 11867–11878.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Hellaswag: Can a machine really finish your sentence? In Proceedings of the AAAI Conference on Artificial Intelligence, pages 11867–11878

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.775155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.068439Z digest=sha256:0200474919ab0f812f048b3bbb1ac7963a2e259e0e0005420a960ca511ad5bcb

Observation c07068fe-7f11-4a98-9164-6fb764ed2d39 · outbound

This paper cites Measuring massive multitask language understanding, 2021.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Measuring massive multitask language understanding, 2021

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.719658Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.072697Z digest=sha256:5e3c0eb7dc65b785ec46ee92579c173c46284ee3b3eb69c3cd29145052900c7c

Observation 3893c908-c4f7-4e8d-b696-046adee30bf2 · outbound

This paper cites On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior.

Model Connectomes: A Generational Approach to Data-Efficient Language Models On the Predictive Power of Neural Language Models for Human Real-Time Comprehension Behavior

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.077764Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.077764Z digest=sha256:86835a891633642adeeae81b9b5181a20daf805df459308accd47f909d50b707

Observation e1643366-2c6a-4774-ab3f-fac6818b2efe · outbound

This paper cites an unresolved cited work.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-16T05:38:38.583094Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.082646Z digest=sha256:aeedfe4eb890a56be7ea930a9e0450204c1bb4280b48f957c0be8d672286bbaa

Observation 3d433734-0b38-4c10-81cb-f6f6fd608e95 · outbound

This paper cites Word frequency and predictability dissociate in naturalistic reading.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Word frequency and predictability dissociate in naturalistic reading

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.087051Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.087051Z digest=sha256:b334f1611d7edc7e1ad17200f97c88966eca03311865b8aa3c21c16861e3a486

Observation 1b9d2d8a-c70c-47c7-a85e-51d704e6aa22 · outbound

This paper cites A probabilistic earley parser as a psycholinguistic model.

Model Connectomes: A Generational Approach to Data-Efficient Language Models A probabilistic earley parser as a psycholinguistic model

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.472801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.091450Z digest=sha256:da31ba7403e47ea10a72daad03b1283ab9a92f9fbfab9128f93cfb0208a98d11

Observation 650d84ce-6ebf-4182-b207-73e219070ff1 · outbound

This paper cites The effect of word predictability on reading time is loga- rithmic.

Model Connectomes: A Generational Approach to Data-Efficient Language Models The effect of word predictability on reading time is loga- rithmic

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.455863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.097083Z digest=sha256:cf255d0310d4004b399f6587ed377e6f233fd6d9384f32014a5f3432d07f14da

Observation be048b5d-8b92-4f8b-92ca-6ce5087cada6 · outbound

This paper cites The natural stories corpus: a reading-time corpus of english texts containing rare syntactic constructions.

Model Connectomes: A Generational Approach to Data-Efficient Language Models The natural stories corpus: a reading-time corpus of english texts containing rare syntactic constructions

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.440629Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.101855Z digest=sha256:5e85341876f4211bd3a3ce6fd026cdbbd063bd59be48c9419ee9cb43d1cedba8

Observation 7aaf19c9-8c2b-496b-aba6-69af289bd8f9 · outbound

This paper cites Instruction-tuning Aligns LLMs to the Human Brain.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Instruction-tuning Aligns LLMs to the Human Brain

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.106333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.106333Z digest=sha256:67759fc7c47d2c7db4006114e25d7218f59019722556a7acaa192785b5efc865

Observation bf37a86a-c9a9-4d48-a1b6-9d1c67a3cbb6 · outbound

This paper cites Brain-Like Language Processing via a Shallow Untrained Multihead Attention Network.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Brain-Like Language Processing via a Shallow Untrained Multihead Attention Network

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-08-16T05:38:37.794971Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.112078Z digest=sha256:4ec5e5a0a8ce2e3c60a6c52b12bfba177464d234099a567c1f9a6936e629d9eb

Observation 553d5476-c038-4380-bef2-dcb384fba41d · outbound

This paper cites Transformer-Based Language Model Surprisal Predicts Human Reading Times Best with About Two Billion Training Tokens.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Transformer-Based Language Model Surprisal Predicts Human Reading Times Best with About Two Billion Training Tokens

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-16T05:38:37.773705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.175734Z digest=sha256:e2b8886f3b2365ef7f9c355e3085b62a33bda26a82332707a15b5be124581173

Observation 1c02f142-52d3-4add-8f10-4658bd39ad2b · outbound

This paper cites Large GPT-like Models are Bad Babies: A Closer Look at the Relationship between Linguistic Competence and Psycholinguistic Measures.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Large GPT-like Models are Bad Babies: A Closer Look at the Relationship between Linguistic Competence and Psycholinguistic Measures

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.244659Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.244659Z digest=sha256:7546d4404b445198f5f33d945f9a9fa0ef6be70f082266291779039f7d469cb6

Observation 2387bb5e-23d0-46cc-be17-ba414cad24d3 · outbound

This paper cites Frequency Explains the Inverse Correlation of Large Language Models' Size, Training Data Amount, and Surprisal's Fit to Reading Times.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Frequency Explains the Inverse Correlation of Large Language Models' Size, Training Data Amount, and Surprisal's Fit to Reading Times

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.248750Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.248750Z digest=sha256:c08093e8c9ccd8f0f0058ca1be840ff1c0e940270cfc45ac00b9ffab23aee416

Observation 9493fcf3-225d-406b-9ab9-903ef787e29c · outbound

This paper cites Scaling in cognitive modelling: A multilingual approach to human reading times.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Scaling in cognitive modelling: A multilingual approach to human reading times

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.417594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.253457Z digest=sha256:f1499395c070120db64c0090f6fc38b9905411ccc53d35df3727a78c23ce3e68

Observation 41a9b749-1c28-482c-83b7-dbc864e9572f · outbound

This paper cites New method for fmri investigations of language: defining rois functionally in individual subjects.

Model Connectomes: A Generational Approach to Data-Efficient Language Models New method for fmri investigations of language: defining rois functionally in individual subjects

Reference 51

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.341021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.258055Z digest=sha256:c4f6444668f57a6552b8241abf87580303df7287d0321f64a660e8624d0e2688

Observation 7ae4d640-0dc5-4ba1-9b2f-a529a00305a8 · outbound

This paper cites Driving and suppressing the human lan- guage network using large language models.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Driving and suppressing the human lan- guage network using large language models

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.240553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.261904Z digest=sha256:e6acb3e808f2911ef649174153ad94c6f57bab119a8a03e7d5d7fe2edc0f4e77

Observation 24434a2b-9c2e-4632-84d6-1302fafc7148 · outbound

This paper cites The LLM Language Network: A Neuroscientific Approach for Identifying Causally Task-Relevant Units.

Model Connectomes: A Generational Approach to Data-Efficient Language Models The LLM Language Network: A Neuroscientific Approach for Identifying Causally Task-Relevant Units

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.265820Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.265820Z digest=sha256:505ef2e6f19efec9d9c51d1e8077f674218ac7bf8081099854f29f9a1fbe9f56

Observation 486719cd-ea37-4c03-82ce-bd1aa03f9ef4 · outbound

This paper cites Distilling the Knowledge in a Neural Network.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Distilling the Knowledge in a Neural Network

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.270182Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.270182Z digest=sha256:7a9aeab19b7ebce62c7855f632775037d00252c365fead72112f8a8f7d9e9c5f

Observation 03ec4639-98df-49e4-b019-4624cb7722ff · outbound

This paper cites Big self-supervised models are strong semi-supervised learners.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Big self-supervised models are strong semi-supervised learners

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.223962Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.273840Z digest=sha256:d66304e2bfbe5680a5345bd8d3db67ad34a5d9478661ed80338cccf419023aa1

Observation 1a0ec43a-297c-4ad1-90a0-8643f18e0011 · outbound

This paper cites PathNet: Evolution Channels Gradient Descent in Super Neural Networks.

Model Connectomes: A Generational Approach to Data-Efficient Language Models PathNet: Evolution Channels Gradient Descent in Super Neural Networks

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.277239Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.277239Z digest=sha256:2897569e913b771473c7aa85411302c0b6814bb49e4a95e6bed584d84eb56be0

Observation 498fad79-88e2-48e2-9db9-13097181a067 · outbound

This paper cites Learning both weights and connections for efficient neural networks.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Learning both weights and connections for efficient neural networks

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.206720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.281648Z digest=sha256:e2731df9ce33ad5a96c7c0893f92549eba87e045c2b51c4e2108342b167af22d

Observation b2d65e16-965c-4e69-8242-ef900970e782 · outbound

This paper cites Synaptic density in human frontal cortex-developmental changes and effects of aging.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Synaptic density in human frontal cortex-developmental changes and effects of aging

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.191754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.285652Z digest=sha256:639b18431f867c6edef1f27e4ae44090808773c653d1474c6c7f13762c79ed3d

Observation d231055b-5c97-48a2-ae0b-bf09647f582e · outbound

This paper cites Natural evolution strategies.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Natural evolution strategies

Reference 59

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.151551Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.290090Z digest=sha256:ed0110fa7fb3572f9fa0b3738c5111d8b5f48622f7574f85a201e4191b97c5f1

Observation 03864042-9046-44fe-95e7-fb2588e7a8bf · outbound

This paper cites Designing neural net- works through neuroevolution.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Designing neural net- works through neuroevolution

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.058859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.295162Z digest=sha256:fb5caaf5e9dc8f4e79730665e4ed638c9bc0b220cef04fee7b67403ea13ebd35

Observation bc15e902-9dbb-420c-9e39-77a312046b0b · outbound

This paper cites Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Interpretability in the Wild: a Circuit for Indirect Object Identification in GPT-2 small

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.300110Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.300110Z digest=sha256:4124b6334a2032716dd1437acaae619685ebbb0db89f172c6d64492e6e19f8a6

Observation 4ebb35a6-a31e-44d0-9d47-822635009f21 · outbound

This paper cites Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Sparse Feature Circuits: Discovering and Editing Interpretable Causal Graphs in Language Models

Reference 62

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.354373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.354373Z digest=sha256:3e0ff77507cbe23465d03ca500f7fe8b61263c8026a974854bb684920ad80a37

Observation 6006c113-686c-4909-a93f-494deca07ce9 · outbound

This paper cites Root mean square layer normalization, 2019.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Root mean square layer normalization, 2019

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:38.011109Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.426514Z digest=sha256:5707421f818d8bd3c363889f7749e768e4a3d9188c022ba0a47be41c534c63e4

Observation 3b4fc5c3-6706-4a63-8a32-8f1bf7b9ab4a · outbound

This paper cites Roformer: Enhanced transformer with rotary position embedding.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Roformer: Enhanced transformer with rotary position embedding

Reference 64

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:37.998646Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.483685Z digest=sha256:1ef38a02f913fca9ce818c7281950e1e46bcbd47e6729b2e5976314430378eef

Observation 5dac39b9-b80c-4afe-b805-c03c6a7d62cc · outbound

This paper cites Brain-score: Which artificial neural network for object recognition is most brain-like? BioRxiv, page 407007, 2018.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Brain-score: Which artificial neural network for object recognition is most brain-like? BioRxiv, page 407007, 2018

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:37.525087Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:37.525087Z digest=sha256:a4ad5f90fd3f326fc3778334435bb2b8605784594af713fd6b18668b33d4e322

Observation c9f64ae1-0113-4155-9db2-5230a2147c63 · outbound

This paper cites Hosseini, Nancy Kanwisher, Joshua B.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Hosseini, Nancy Kanwisher, Joshua B

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-16T05:38:37.975310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-16T05:38:37.529590Z digest=sha256:6500e3d0a1a66ff5647375bdbf26f7c555d0b46923a82ad58af53e496abef9a6

Observation 84dcf064-9ca6-4074-9e11-4dc7cc52825a · outbound

This paper cites Lottery Ticket Adaptation: Mitigating Destructive Interference in LLMs.

Model Connectomes: A Generational Approach to Data-Efficient Language Models Lottery Ticket Adaptation: Mitigating Destructive Interference in LLMs

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-16T05:38:36.776118Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T05:38:36.776118Z digest=sha256:1fa12181c4730476a7d43d3ab237b556a1cb737afb04efcc639883ec32938a3e

Pith citing papers

No inbound Pith citation observations are available.