Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:12:36.460382Z
Paper Citation Record · LEDGER
As of 21 August 2026, this Paper Citation Record lists 19 of 19 outbound references and 0 inbound Pith citation observations for arXiv:2505.01855.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-16T04:12:36.460382Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links
A source-named dated measurement, never combined with another source.
Source: cited_works
19 of 19 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 650b011a-e6ec-4ed4-9724-6b431c81b4e7 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Attention Is All You Need
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 457b876e-4083-45e2-ba3e-0ffd5e664dfc · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Language Models are Few-Shot Learners
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 73369a2f-3d35-4099-8073-6cad5bd82b9d · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Universal Transformers
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0774c4e8-00cc-4e04-a25d-31b31e81b1e0 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Looped Transformers as Programmable Computers
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6054bce5-4cd7-4ef5-9257-1e454c7897d0 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Looped Transformers are Better at Learning Learning Algorithms
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0f7f533e-c6a0-4330-876e-f23935ea085e · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Looped Transformers for Length Generalization
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 303aaed1-9f4c-4b57-b6ad-7d9bd5573bbb · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f8491164-26e3-4e15-ae0a-a04d60463797 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Longshort-termmemory
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 0398b6c7-4092-46b7-8c71-64dfe241238d · outbound
Intra-Layer Recurrence in Transformers for Language Modeling BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d41e9df3-e36b-482b-8fac-fa0751b70b33 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Lan- guage Models are Unsupervised Multitask Learners
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 9bd32c0f-b4c7-41db-8958-50c47f2a6c80 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling LLaMA: Open and Efficient Foundation Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 83f82706-ccc6-44b8-a1d9-7b99ee9393ae · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Revealing the Dark Secrets of BERT
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ce47652-879a-4a71-a231-ef3d072e594d · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Interpreting GPT: The Logit Lens
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 386d908e-aba7-434c-b06b-941389143c0a · outbound
Intra-Layer Recurrence in Transformers for Language Modeling The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c8a0b64a-9052-4ead-9a31-388d13306ebe · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.
Observation 2848ffd2-5aa2-4180-a268-0ce3fe7ab71a · outbound
Intra-Layer Recurrence in Transformers for Language Modeling RoFormer: Enhanced Transformer with Rotary Position Embedding
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8e66d75b-9b58-4270-b52f-2d456d75093a · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Train Short, Test Long: Attention with Linear Biases Enables Input Length Extrapolation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e1634db2-5f44-4a16-b554-c8d24da1ffae · outbound
Intra-Layer Recurrence in Transformers for Language Modeling Training Compute-Optimal Large Language Models
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 48ce3fd7-76b5-4b1b-995d-8d6f441ce078 · outbound
Intra-Layer Recurrence in Transformers for Language Modeling The Impact of Positional Encoding on Length Generalization in Transformers
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
No inbound Pith citation observations are available.