Pith. sign in

Paper Citation Record · LEDGER

Sparse Layers are Critical to Scaling Looped Language Models

As of 5 August 2026, this Paper Citation Record lists 34 of 34 outbound references and 2 inbound Pith citation observations for arXiv:2605.09165.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.09165 v2

Coverage vector

measured 34 of 34 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-01T07:26:37.559459Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T21:34:28.320684Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T17:08:44.006954Z

Reference resolution

34 of 34 outbound references displayed

  • verified exact19
  • verified fuzzy4
  • unresolved3
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch7

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation af9ae7f5-b4c6-43fd-9394-60f4ef6c3fd7 · outbound

This paper cites Universal Transformers.

Sparse Layers are Critical to Scaling Looped Language Models Universal Transformers

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.141390Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:63b76581d12f815fe8013ba18b617389134133bb5794d50065ec9a42e0bd5ee5

Observation 08f630f4-af7e-4624-b4fa-8085d07a2458 · outbound

This paper cites Scaling Latent Reasoning via Looped Language Models.

Sparse Layers are Critical to Scaling Looped Language Models Scaling Latent Reasoning via Looped Language Models

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.143836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:44c374030b254ef5d0311fa20828862898484c01d2965e7fb4ab8a87344bcbfd

Observation 6cddd360-fa1d-4001-8120-9b26c997f321 · outbound

This paper cites Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach.

Sparse Layers are Critical to Scaling Looped Language Models Scaling up Test-Time Compute with Latent Reasoning: A Recurrent Depth Approach

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T07:35:29.150966Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:e13c469db93c0ce9508005ba25b1f2038ba1600a23c3a68ecb3f011796d6583e

Observation ebb975bf-74c0-48fa-b4e2-f7c2f1b49285 · outbound

This paper cites Scaling Laws for Neural Language Models.

Sparse Layers are Critical to Scaling Looped Language Models Scaling Laws for Neural Language Models

Reference 4

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T07:35:29.146101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:13be3da56b6d11b68734e8bdd4f4aebab0299dd5181ea03f5cc3ddd6f0ac52e0

Observation 96cdd294-7655-4f0a-a445-07e850fc2128 · outbound

This paper cites Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer.

Sparse Layers are Critical to Scaling Looped Language Models Outrageously Large Neural Networks: The Sparsely-Gated Mixture-of-Experts Layer

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.148683Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:6cf358ef936c44ae91c7e5125ee378f2f178ce50b3bd30727841b6d6ff54a092

Observation fc17ab31-d334-406e-a644-6bbe10371ba7 · outbound

This paper cites MoEUT: Mixture-of-Experts Universal Transformers.

Sparse Layers are Critical to Scaling Looped Language Models MoEUT: Mixture-of-Experts Universal Transformers

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T15:32:41.166003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:89ab56259ff45cf985e0c3941d0603feba274cd035122b7cc135434b30bc627c

Observation 30f10f13-7b88-4333-b5e3-bfc05eb0cfd9 · outbound

This paper cites an unresolved cited work.

Sparse Layers are Critical to Scaling Looped Language Models Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-07-06T15:32:41.175489Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:8804feda2aa57f4c564ab5f1b8014412a0d136191a3023c3038766e4d1cd0ce5

Observation 8f2f846e-f696-481b-9a95-f94bf96ddf75 · outbound

This paper cites URL http://ieeexplore.ieee.org/document/ 7900006/.

Sparse Layers are Critical to Scaling Looped Language Models URL http://ieeexplore.ieee.org/document/ 7900006/

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:28.527666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:96ff3fa984aa6569ad0c8afa65780c89c9877dbcf876c6f88abc0a1e9cf3b458

Observation 12411a4e-616a-47d4-8169-a10b05fa6253 · outbound

This paper cites DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference.

Sparse Layers are Critical to Scaling Looped Language Models DeeBERT: Dynamic Early Exiting for Accelerating BERT Inference

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.162165Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:1c403d57b461c6838bbc3d138ee4768dfc82c4462f2a5e39f349661be6bc6f16

Observation 7b3c31ce-afb1-40e2-bb03-9e3114b67417 · outbound

This paper cites Confident Adaptive Language Modeling.

Sparse Layers are Critical to Scaling Looped Language Models Confident Adaptive Language Modeling

Reference 10

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:35:29.157001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:7421a6b625eeb395b77119144d9bfe9ae89261feabcb0ff8bb0c62598919eec7

Observation 340b20d4-83df-4fb3-a089-43013b248b20 · outbound

This paper cites RoFormer: Enhanced Transformer with Rotary Position Embedding.

Sparse Layers are Critical to Scaling Looped Language Models RoFormer: Enhanced Transformer with Rotary Position Embedding

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.144506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:ab70bbd16b058f4d7a47722d5ece80bc28322a3a2b8f152a2928a9c30eb3250f

Observation 4f4bb4fd-9db1-453f-85dd-52893aceb77f · outbound

This paper cites GLU Variants Improve Transformer.

Sparse Layers are Critical to Scaling Looped Language Models GLU Variants Improve Transformer

Reference 12

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.149543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:f576dd71ef030cc124908539ad3248d3054503b7b64b005855649ee7ff8a6436

Observation 7e856c73-559c-4b1d-bff5-27469d4ea621 · outbound

This paper cites Root Mean Square Layer Normalization.

Sparse Layers are Critical to Scaling Looped Language Models Root Mean Square Layer Normalization

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T15:32:41.171792Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:2f283af3d0545b2f842b009c1025dee6d07976bb706a6fc73cfe9c423e831717

Observation 252f9fdf-be00-47a8-a6d7-a4206bc4ecd0 · outbound

This paper cites Switch transformers: scaling to trillion parameter models with simple and efficient sparsity.J.

Sparse Layers are Critical to Scaling Looped Language Models Switch transformers: scaling to trillion parameter models with simple and efficient sparsity.J

Reference 14

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T07:35:29.131604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:d155cdab2628c3666a2eba40a43670e11ab2361fbdd04c3bb690a9c0093defcf

Observation 6ae657e2-4a9d-45a9-b8fa-3a8e84edd563 · outbound

This paper cites ST-MoE: Designing Stable and Transferable Sparse Expert Models, April.

Sparse Layers are Critical to Scaling Looped Language Models ST-MoE: Designing Stable and Transferable Sparse Expert Models, April

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T15:32:41.168219Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:70f4bbc7626c9b7ef96487dc88d3f68814481a94340ab1a31632ccc914bbf629

Observation a8545f5e-dea0-4647-b36b-3b57af46b590 · outbound

This paper cites ST-MoE: Designing Stable and Transferable Sparse Expert Models.

Sparse Layers are Critical to Scaling Looped Language Models ST-MoE: Designing Stable and Transferable Sparse Expert Models

Reference 16

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T07:35:29.151888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:b570c2a7d4af0bb0844037a120aead149a0bb022e9800ec78394a02f5598690e

Observation ace89dca-a025-4f73-a417-b3b1e17513b4 · outbound

This paper cites Mixtral of Experts.

Sparse Layers are Critical to Scaling Looped Language Models Mixtral of Experts

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.139849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:a2f9b4a045905aecda212f475f81a3c9f477af86c5545bf858c67632e5b4c5b3

Observation 8e1563b3-a639-4f80-b3cb-96f1a24bc552 · outbound

This paper cites Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer.

Sparse Layers are Critical to Scaling Looped Language Models Tensor Programs V: Tuning Large Neural Networks via Zero-Shot Hyperparameter Transfer

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T07:35:29.134236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:db939d14663b7bdfdb2c9940778163f8f005125d889fe225f3e97f1483b06175

Observation 8d7e4e47-4e11-4c1f-9fc7-d0c71b3b12b6 · outbound

This paper cites µ-parametrization for mixture of experts.

Sparse Layers are Critical to Scaling Looped Language Models µ-parametrization for mixture of experts

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.137134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:13c451fc2e02ae446bcea04961993663d5568f6fcc3d31f3ec2761772a87f298

Observation 78d1846f-9335-4d9d-9343-6f08a72c761e · outbound

This paper cites Training Compute-Optimal Large Language Models.

Sparse Layers are Critical to Scaling Looped Language Models Training Compute-Optimal Large Language Models

Reference 20

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T07:35:29.154592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:ea68ca842d89ca2cb3304076051d86a743b2f2ed38b3b9e32bb83e565edf53d8

Observation 0c737458-8940-4103-a9d9-b1a1f75a123a · outbound

This paper cites The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale.

Sparse Layers are Critical to Scaling Looped Language Models The FineWeb Datasets: Decanting the Web for the Finest Text Data at Scale

Reference 21

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.164556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:0573abcb90d1fe3ae9f369a411161da7dc88e60280c4e8baf46d5535960cf0bf

Observation 0a558ddc-639f-478e-a3e9-510344677a4e · outbound

This paper cites MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies.

Sparse Layers are Critical to Scaling Looped Language Models MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.142171Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:b837a7676926452bb5d77ddf40faf531f074632d62cc0e6f46b8bb78e2fad96c

Observation 912ed991-2177-4a6b-84d1-51791829ffd9 · outbound

This paper cites Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler.

Sparse Layers are Critical to Scaling Looped Language Models Training Dynamics of the Cooldown Stage in Warmup-Stable-Decay Learning Rate Scheduler

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.147307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:8b8acbf02b69d97df45f5b307ac92f1de5048ab051fe3a2b9f07a4b6dbd884d9

Observation 19f4615f-ead4-43ca-8814-52ef30da456b · outbound

This paper cites OLMES: A Standard for Language Model Evaluations.

Sparse Layers are Critical to Scaling Looped Language Models OLMES: A Standard for Language Model Evaluations

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.159520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:abd92e0a461ed48a2593d4ead010c33795b7304e43aaec83c2480eece9f6d1ce

Observation 1678ac11-6a11-4607-b776-7df0c13b453f · outbound

This paper cites an unresolved cited work.

Sparse Layers are Critical to Scaling Looped Language Models Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-07-06T15:32:41.170030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:3daf36c5b87ae6bbd8b96a7309d3cfdaad79310cfbb4a4c9dadc97c9d34634df

Observation 8ea638cf-0782-4ebb-b23c-db06c9603c45 · outbound

This paper cites interpreting GPT: the logit lens — LessWrong.

Sparse Layers are Critical to Scaling Looped Language Models interpreting GPT: the logit lens — LessWrong

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-07-06T15:32:41.179388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:d1d4e6a9612549dd5d4ec9736698f46f54b62dd9ccff415949402caef82ff21f

Observation 66c4dcfe-c260-41f7-bad6-0cee2d519e0a · outbound

This paper cites an unresolved cited work.

Sparse Layers are Critical to Scaling Looped Language Models Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-07-06T15:32:41.177501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:34062ced5880c03ef2267f773b78de31a19011f0ef67e8b709ad3d1bc2f11410

Observation 012e7808-59e0-460c-b772-3acd216a5213 · outbound

This paper cites Approximating Two-Layer Feedforward Networks for Efficient Transformers.

Sparse Layers are Critical to Scaling Looped Language Models Approximating Two-Layer Feedforward Networks for Efficient Transformers

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.121225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:a84168080c2d2eb8ce77b36a74d5b914815684c203e697251326264ac970466b

Observation 82b6ef2c-655b-4686-ac88-54e268505ad6 · outbound

This paper cites SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention.

Sparse Layers are Critical to Scaling Looped Language Models SwitchHead: Accelerating Transformers with Mixture-of-Experts Attention

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.123697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:dd437219839488c921a1ecbc35351ab144a688d7f6e49f117282f4b71cb6ee3a

Observation 16c6fe74-1d7f-4843-a6e4-2b631db4cdb7 · outbound

This paper cites Layer Normalization.

Sparse Layers are Critical to Scaling Looped Language Models Layer Normalization

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.126150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:95166fd39cc5a4d260694d3577bd1fa88633f0be87af25cf6d3bbc46c2744062

Observation 19915ffc-48c3-483d-833a-1635d2273075 · outbound

This paper cites L ayer S kip: Enabling Early Exit Inference and Self-Speculative Decoding.

Sparse Layers are Critical to Scaling Looped Language Models L ayer S kip: Enabling Early Exit Inference and Self-Speculative Decoding

Reference 31

Resolution
metadata mismatch
doi, observed 2026-07-01T07:35:28.530348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:1c623ee99c741e50d00e15be36c992b06163ab01afeea9f11f3baa821cd57b01

Observation 1de4db74-4a21-4dd7-b6b2-3b4b7cb1b546 · outbound

This paper cites Mixture-of-Depths: Dynamically allocating compute in transformer-based language models.

Sparse Layers are Critical to Scaling Looped Language Models Mixture-of-Depths: Dynamically allocating compute in transformer-based language models

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-07-01T07:35:29.113057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:bc7d60c4a11f2457bae2f6aab9a39cc43b515833d17dae95f7127b48ed54023c

Observation 59aaef27-6760-4cdf-a2ec-0da6d6763c27 · outbound

This paper cites Mixture-of-recursions: Learning dynamic recur- sive depths for adaptive token-level computation.arXiv preprint arXiv:2507.10524.

Sparse Layers are Critical to Scaling Looped Language Models Mixture-of-recursions: Learning dynamic recur- sive depths for adaptive token-level computation.arXiv preprint arXiv:2507.10524

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.115807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:580d9c31c481c62211d73c38dbde3e915b3ce8b1eecd800ee0567bd99ed86a1e

Observation f9b073a9-3594-4bd7-a615-fd79224e0d35 · outbound

This paper cites Don’t be lazy: Completep enables compute-efficient deep transformers.

Sparse Layers are Critical to Scaling Looped Language Models Don’t be lazy: Completep enables compute-efficient deep transformers

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-01T07:35:29.118338Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-07-01T07:26:37.559459Z digest=sha256:441055d4101cf268f8ad55926faf47a82295b6b7a90b76a095e99dcde825871c

Pith citing papers

Observation 956782ee-58a1-44a1-bbe2-b9b7db90e466 · inbound

Dense Supervision Is Not Enough: The Readout Blind Spot in Looped Language Models cites this paper.

Dense Supervision Is Not Enough: The Readout Blind Spot in Looped Language Models Sparse Layers are Critical to Scaling Looped Language Models

Reference 10

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T17:08:44.008245Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=pdf_text observed=2026-06-27T04:27:48.590008Z digest=sha256:6ef06280e3eb9b20fa0d9c643c1644b636859b5062681f8f11898baac647dd8c

Observation 3f0ffe69-439e-470a-86ee-8a5bbddbc55f · inbound

Loop the Loopies! cites this paper.

Loop the Loopies! Sparse Layers are Critical to Scaling Looped Language Models

Reference 127

Resolution
unresolved
no resolver link, observed 2026-08-01T21:34:28.320684Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T21:34:28.320684Z digest=sha256:4e2a4b7527664eb7a3f02bea4d0bf1adfa5584af62cb37b8247d863a1801f418