Pith. sign in

Paper Citation Record · LEDGER

ImageFolder: Autoregressive Image Generation with Folded Tokens

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2410.01756.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2410.01756 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 24 of 24 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-08T14:25:55.738517Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T14:38:28.817961Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 30286c70-f0d9-41bc-bf97-62ed05f26b43 · inbound

EVEv2: Improved Baselines for Encoder-Free Vision-Language Models cites this paper.

EVEv2: Improved Baselines for Encoder-Free Vision-Language Models ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-08T14:25:55.738517Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T14:25:55.738517Z digest=sha256:94c6eabf85111274b9e19b2c0d27965acf323b01c4e7e7aa65181e4d46310f8b

Observation 1e4b70dc-e21f-4f7f-ad8e-b513230f332b · inbound

High-Fidelity Functional Ultrasound Reconstruction via A Visual Auto-Regressive Framework cites this paper.

High-Fidelity Functional Ultrasound Reconstruction via A Visual Auto-Regressive Framework ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-07T14:40:16.912616Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:40:16.912616Z digest=sha256:bd60b534c440728ef7bff090b5a27328f3616c78b51d00391316ca1c6cdaf575

Observation e7402fdf-de54-43c8-b179-ba2822a82963 · inbound

HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation cites this paper.

HMAR: Efficient Hierarchical Masked Auto-Regressive Image Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T10:47:52.342304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:47:52.342304Z digest=sha256:43710f18672adca24d2827b10408d89783c64d026817b070b84bf27b64ce881b

Observation 94441927-3fba-4097-aa24-4d85323b66dd · inbound

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective cites this paper.

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-06T20:53:51.001183Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:53:51.001183Z digest=sha256:2136ab110978806d4dc82e87225bc02e3fec0d84871c161cec5b514e63f37438

Observation c129526d-d8b1-42c3-a695-7b51aa3944d5 · inbound

Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation cites this paper.

Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T17:00:24.110592Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T16:57:09.843691Z digest=sha256:b2b809ffc487eade71a41472ed9e6fd58c220f975e61b5fd29b5383813ee3ba3

Observation 7c9fdf2a-d4cc-48b3-a9dd-d4d28bebc405 · inbound

Language-Guided Transformer Tokenizer for Human Motion Generation cites this paper.

Language-Guided Transformer Tokenizer for Human Motion Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T03:25:01.015059Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:25:01.015059Z digest=sha256:9fae5edb6d0a586cbf7b7ca7187f1cebfe17a24531925563181263c92f582a26

Observation 09949609-7993-4ae1-8e2a-af78e381ccc7 · inbound

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction cites this paper.

Representations Before Pixels: Semantics-Guided Hierarchical Video Prediction ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 47

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T08:45:58.558880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T16:30:53.578491Z digest=sha256:1c43bdffc48445b233d0d1a057a08053c43416fc95c6242f8aeac30c1b4f36c2

Observation 8da6bca7-b891-4169-a2cc-b839f7494d03 · inbound

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations cites this paper.

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:11.526347Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T04:12:58.100610Z digest=sha256:9e5af59bc5c534bc840722e86336eb654b8f7bd6c2e631622f84e75c0e4070c0

Observation f1559262-224d-44f1-9e05-fb003bd5a13c · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:51:06.106220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-08T13:52:42.834440Z digest=sha256:8b4837aca4033d39d4fe8f9ce8ede6ef7a301aa83796ba1ca51266875c947bc2

Observation 37377297-0e32-4a31-8956-939fae63d2b0 · inbound

Autoregressive Visual Generation Needs a Prologue cites this paper.

Autoregressive Visual Generation Needs a Prologue ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-06-30T23:45:08.405282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-30T23:36:51.148162Z digest=sha256:6fb28ec647e4bd5c6c55750966bbd729dbacda446281493fd6901b1ff3234be9

Observation 17912036-6ec8-43dc-bb3a-864a1572f0ce · inbound

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion cites this paper.

What Matters for Diffusion-Friendly Latent Manifold? Prior-Aligned Autoencoders for Latent Diffusion ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T04:05:57.596052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-11T01:57:24.033068Z digest=sha256:a7ac5ab9b1697208cc45fcf298873de043052a560eea7818766c398d82b982ac

Observation 8683721b-f159-42e7-87b9-ead524fecf49 · inbound

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens cites this paper.

WinTok: A Win-Win Hybrid Tokenizer via Decomposing Visual Understanding and Generation with Transferable Tokens ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:08:15.766012Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T12:04:19.761430Z digest=sha256:9a5604a24235cc6f9fcd598da7231b4e19b000a4af7b5b6b4e712057804ae27d

Observation 35a3c3f3-f2dc-45c9-a751-5897e4c373f8 · inbound

Vision Foundation Models as Generalist Tokenizers for Image Generation cites this paper.

Vision Foundation Models as Generalist Tokenizers for Image Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:03:13.505362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T11:01:24.738195Z digest=sha256:44568afd7d4426c6151fa0d22cd9041876c757c27563bf1dda10666289e3ad2d

Observation 2b07b738-19ed-4e91-a393-bfced9d837b2 · inbound

Structure over Pixels: Learning Variable-Length Visual Programs cites this paper.

Structure over Pixels: Learning Variable-Length Visual Programs ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:13:49.038250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T18:06:01.684713Z digest=sha256:a7fb4fe556fcbff8f93e066bbc8c76bb711047d8ddb0c61b9c66d9698f543a68

Observation 23d2b2c6-191c-44b0-8b2d-7c738ba24df3 · inbound

Fixed-Point Masked Generative Modeling cites this paper.

Fixed-Point Masked Generative Modeling ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 44

Resolution
verified exact
arxiv_id, observed 2026-06-28T23:22:47.319491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T23:18:27.050697Z digest=sha256:1afe5db465e7fec2c4420205f61067f72b33311d3c7f250c8668ac9c9d1b33bd

Observation 426bcb2f-4166-4dee-94a5-724f01eeb82f · inbound

Balancing Image Compression and Generation with Bootstrapped Tokenization cites this paper.

Balancing Image Compression and Generation with Bootstrapped Tokenization ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:55.467937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T03:07:33.054518Z digest=sha256:03cac5db189f8404f44496c4dc6404dab872cf217461ee8b0bc7234adf10cec4

Observation 1cc79708-28e0-4ccd-9a75-e6951b1117d6 · inbound

AdaTok: Self-Budgeting Image Tokenization with Quality-Preserving Dynamic Tokens cites this paper.

AdaTok: Self-Budgeting Image Tokenization with Quality-Preserving Dynamic Tokens ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 29

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T16:27:09.117096Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T22:43:09.489524Z digest=sha256:cf7faddf6ce92ae8baa806cbebad20ca871cd7906beb4cefed3a3140fa12756e

Observation 280769d0-f1ae-473a-925a-d3bc569611a0 · inbound

IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder cites this paper.

IDEAL: In-DEpth ALignment Makes A Discrete Representation AutoEncoder ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-03T05:37:40.280769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T13:10:14.308216Z digest=sha256:5f16c92f6239491b7a94d9af603083e7c8df148afbf4afbdc4134b0f33a7a69d

Observation e3b60c85-bd9d-43fc-9357-5d022a1a207b · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T14:38:28.819341Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:fea41ecc3501146aca42080e80d8ecf5676c75cc6e7c9c2f1c8811b8bd79a3ee

Observation d8be1fea-6bcb-401a-bc7d-e3da85105ca0 · inbound

MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixture of Experts cites this paper.

MEPA: Multi-Scale Representation Alignment for Visual Autoregressive Modeling with Mixture of Experts ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 36

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T15:17:07.209220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-02T15:14:36.946247Z digest=sha256:4eda4b851d63d45718b2372b4210ac58b03ffd7e1c998af077bc352e2cae02f8

Observation e11fcc61-85fd-46de-92a8-72ab290e30e0 · inbound

Orbis 2: A Hierarchical World Model for Driving cites this paper.

Orbis 2: A Hierarchical World Model for Driving ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-01T22:05:03.320207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:05:03.320207Z digest=sha256:40ca0be0cfaec2daabe48d3f04fe929eecc2adef104144c8f2c86aba3a1fad48

Observation a75947e9-8cd5-4698-a61a-9947aef30afb · inbound

Twins: Learn to Predict Unified Representations with Focal Loss cites this paper.

Twins: Learn to Predict Unified Representations with Focal Loss ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 55

Resolution
unresolved
no resolver link, observed 2026-08-01T04:29:47.341170Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T04:29:47.341170Z digest=sha256:b62a596b90c9c784904ec72698908829f20d44e7a219bdd4ea24db6ef2197664

Observation cd358879-c884-476a-978b-545f7e2af462 · inbound

UniGen-AR: Unifying Visual Generation with Auto-Regressive Modeling cites this paper.

UniGen-AR: Unifying Visual Generation with Auto-Regressive Modeling ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-31T22:25:12.380740Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-31T22:25:12.380740Z digest=sha256:8806a7644b5eeac6cfc02dd82313b4ca655a13e54459406d63d9cde5954ef1e6

Observation fbe590f5-d3f7-4fde-81cd-e79630b097cd · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation ImageFolder: Autoregressive Image Generation with Folded Tokens

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.733716Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.733716Z digest=sha256:dcd328d57128a159c0a0ac68c16dab8a8b060eaba1127146e3cd995e06459bb4