Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2411.12372.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:29:09.304268Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
18
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 4824c5f2-d689-44d9-b29b-4d9ac1db0003 · inbound
Diagnosing our datasets: How does my language model learn clinical information? RedPajama: an Open Dataset for Training Large Language Models
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e15629f2-7a49-41eb-85b5-df5ef3621fc5 · inbound
Hardware-Efficient Attention for Fast Decoding RedPajama: an Open Dataset for Training Large Language Models
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37b1cbea-2d9b-41f2-a510-3b6ce62ca490 · inbound
The Common Pile v0.1: An 8TB Dataset of Public Domain and Openly Licensed Text RedPajama: an Open Dataset for Training Large Language Models
Reference 196
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2a633f51-b59d-4224-8915-6a1fad4c7ad1 · inbound
Domain2Vec: Vectorizing Datasets to Find the Optimal Data Mixture without Training RedPajama: an Open Dataset for Training Large Language Models
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 566ef656-7006-4020-8388-65edc1ea2007 · inbound
Essential-Web v1.0: 24T tokens of organized web data RedPajama: an Open Dataset for Training Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8d4a238-444a-4897-981e-b963eab614df · inbound
Rethinking 1-bit Optimization Leveraging Pre-trained Large Language Models RedPajama: an Open Dataset for Training Large Language Models
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 9bbf3e5f-ac51-4f10-b045-a2f423eafe06 · inbound
OLMoASR: Open Models and Data for Training Robust Speech Recognition Models RedPajama: an Open Dataset for Training Large Language Models
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c4457a45-cf0d-497e-b605-89b5accb6f50 · inbound
Router Upcycling: Leveraging Mixture-of-Routers in Mixture-of-Experts Upcycling RedPajama: an Open Dataset for Training Large Language Models
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98500d58-abc6-4f10-a38d-02f91e33dfbc · inbound
Beyond URLs: Metadata Diversity and Position for Efficient LLM Pretraining RedPajama: an Open Dataset for Training Large Language Models
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 6b449b42-ea8b-4d18-969d-63fcdb5c7cf9 · inbound
The Effect of Scripts and Formats on LLM Numeracy RedPajama: an Open Dataset for Training Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dff2e6a1-fa65-4449-b2f9-f9606e2091bd · inbound
A Survey on Evaluating Quality and Trustworthiness in LLM-Generated Data RedPajama: an Open Dataset for Training Large Language Models
Reference 231
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 590a810a-bea8-41b1-9df8-106cc74348e1 · inbound
Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings RedPajama: an Open Dataset for Training Large Language Models
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation f10efb69-6579-48cc-bbba-5edf701d6a07 · inbound
Small edits, large models: How Wikipedia advocacy shapes LLM values RedPajama: an Open Dataset for Training Large Language Models
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 32e897c5-9350-474a-a03d-39cf2adbf4ec · inbound
Small edits, large models: How Wikipedia advocacy shapes LLM values RedPajama: an Open Dataset for Training Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9729b55e-f368-4e28-bc2c-9fd2bdb8eb9a · inbound
Narrative-UFET: Narrative Generation for Ultra-Fine Entity Typing RedPajama: an Open Dataset for Training Large Language Models
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 152f7a68-96f4-42be-912d-4711e32a7b26 · inbound
$\text{Log}_\text{b}$Quant: Quantizing Language Models in Logarithmic Space RedPajama: an Open Dataset for Training Large Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 2a65a684-4ba9-4249-bc4c-5dbb900df86c · inbound
A First-Principles Theory of Slow Thinking and Active Perception RedPajama: an Open Dataset for Training Large Language Models
Reference 167
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 69a6c735-9747-426f-aa56-d90025b34803 · inbound
Libra: Taming Attention Workload Skew in Long-Context LLM Training with Bounded Sequence Pool RedPajama: an Open Dataset for Training Large Language Models
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2acd6937-8313-480a-9af2-15d5e44ef445 · inbound
From Data to Device: ELMOD An Efficient German-First 2.7B Language Model for Mobile Inference RedPajama: an Open Dataset for Training Large Language Models
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.