Pith. sign in

Paper Citation Record · LEDGER

Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 46 inbound Pith citation observations for arXiv:2404.02905.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2404.02905 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 46 of 46 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T23:28:07.776955Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

7
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d91e267a-b3a2-4d8c-8cf8-9ca90a95bf04 · inbound

Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation cites this paper.

Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 33

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T22:09:17.082429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T22:09:16.622717Z digest=sha256:81881149158c511786bed5d7dc01e0dd3c9405fc137c0b0ec44fd5290a8cd4d5

Observation aa59adfe-778c-44d2-9b3f-206b7dbd019e · inbound

Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation cites this paper.

Janus: Decoupling Visual Encoding for Unified Multimodal Understanding and Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 79

Resolution
verified exact
arxiv_id, observed 2026-05-15T22:09:16.295516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T22:09:16.001309Z digest=sha256:37414d4e66c343e6e0a02aa2696633dbf60efdab7112aa7912201ce95b595f59

Observation 1c6270e0-1e58-47e8-9950-89e91be7415e · inbound

Autoregressive Video Generation without Vector Quantization cites this paper.

Autoregressive Video Generation without Vector Quantization Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 24

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T15:07:39.912883Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-17T15:07:39.718555Z digest=sha256:b3826bb9b0054bec5e8bf8077dc119587dc35bf7b00a3a22003a9a83bb58fc8f

Observation e1c44baf-a60b-4c92-a621-f698ecaee904 · inbound

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation cites this paper.

WISE: A World Knowledge-Informed Semantic Evaluation for Text-to-Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T16:24:27.751864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-15T16:24:27.407376Z digest=sha256:4cb9cc6efab438049522ec3164a2d9ef0c290e9a1459adcf1073cda2c9a4b2e5

Observation 6c6b19f1-d144-4ac2-baae-41fc252e9ec4 · inbound

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation cites this paper.

ShareGPT-4o-Image: Aligning Multimodal Models with GPT-4o-Level Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T23:28:07.776955Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:28:07.776955Z digest=sha256:45c5dfce402a7b060d59557b89caa8b2f9e81730102510823d3f03b9c23f65d3

Observation 7e1f0a75-9876-4c36-8caa-c1527bff8700 · inbound

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation cites this paper.

Instella-T2I: Pushing the Limits of 1D Discrete Latent Space Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T22:43:04.234499Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:43:04.234499Z digest=sha256:78634b616fae2045576cf0f58aac73edcc4327b545871f9f793c946cb79c5ff5

Observation 9dc51c88-7357-4f98-a4a9-1de3b1eb3ffd · inbound

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation cites this paper.

CycleVAR: Repurposing Autoregressive Model for Unsupervised One-Step Image Translation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-06T21:49:44.581169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:49:44.581169Z digest=sha256:0657680f88b473171c07e8d7e2c5eca093f35ea760a15b61eaf283c6b6054fd9

Observation e6940ed3-9957-47f0-88c4-acdfa6af7c69 · inbound

Transition Matching: Scalable and Flexible Generative Modeling cites this paper.

Transition Matching: Scalable and Flexible Generative Modeling Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-06T21:46:29.478571Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T21:46:29.478571Z digest=sha256:c6e8ba83b79745bad33e01f54ae3fd0d8f9821795968d8d80d68904dc0f2e181

Observation 72462e34-fbb2-48d0-9d92-2f6d9989153a · inbound

LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs cites this paper.

LLaVA-SP: Enhancing Visual Representation with Visual Spatial Tokens for MLLMs Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 52

Resolution
unresolved
no resolver link, observed 2026-08-06T21:19:45.091175Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:19:45.091175Z digest=sha256:2993260f2094eddaec4c6881ec275c5ea8c9eaec9f32fad811e11bdb719f114e

Observation 5cdea1a8-c53a-44ee-8b72-d77010d5c024 · inbound

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective cites this paper.

Autoregressive Image Generation with Linear Complexity: A Spatial-Aware Decay Perspective Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-06T20:53:50.735276Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:53:50.735276Z digest=sha256:95a44fc0da4ed6a2d6382440a717ad5cb73f31eeee6e6da9626ef18eaf058944

Observation 0545d7ab-e8ab-4af0-aa92-0f0e0a056a55 · inbound

Hita: Holistic Tokenizer for Autoregressive Image Generation cites this paper.

Hita: Holistic Tokenizer for Autoregressive Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T20:38:57.032639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:38:57.032639Z digest=sha256:aa1070d987008aad6f7bfc29b2a00a5a7d22b51035cea9fd0c9b28150dd9fe16

Observation c331c4ef-5f0d-4885-8015-dc306921171d · inbound

AIGI-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models cites this paper.

AIGI-Holmes: Towards Explainable and Generalizable AI-Generated Image Detection via Multimodal Large Language Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 85

Resolution
unresolved
no resolver link, observed 2026-08-06T20:32:59.851177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:32:59.851177Z digest=sha256:ea9d789615e3b0d921cfb0a322ac0d6e1f8820779694fd958494c917fbb1fa83

Observation a7b8f6bf-c5f6-41d4-af45-62a8b04c1406 · inbound

Cautious Next Token Prediction cites this paper.

Cautious Next Token Prediction Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T20:39:13.582808Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:39:13.582808Z digest=sha256:9d6a355a390ea14db1ada553ade2baec3fffa872905c216cb1f71f191bf93bf3

Observation 6922dc7f-bc2e-4614-b45c-672fe72dca08 · inbound

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization cites this paper.

MambaVideo for Discrete Video Tokenization with Channel-Split Quantization Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T19:52:28.093409Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:52:28.093409Z digest=sha256:cdd02e81926d6133f7d428dff4fa7a172596ef6a9d4750ab1bb766941cd99feb

Observation b085723f-8da0-459c-b4b6-83be057144a1 · inbound

Implementing Adaptations for Vision AutoRegressive Model cites this paper.

Implementing Adaptations for Vision AutoRegressive Model Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T17:12:34.614860Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T17:12:34.614860Z digest=sha256:147ba4912c74a3168b3c479deaf4a35216f31b73b9c3808e81f090547b7e50bf

Observation 7e04678f-8b4e-4b4a-af7b-5d204de7e867 · inbound

HRVVS: A High-resolution Video Vasculature Segmentation Network via Hierarchical Autoregressive Residual Priors cites this paper.

HRVVS: A High-resolution Video Vasculature Segmentation Network via Hierarchical Autoregressive Residual Priors Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T11:38:36.823952Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:38:36.823952Z digest=sha256:520d314189d066246534c69b5a1ce7cccab9a12378f3d4e90d8965fa13a18eca

Observation e041e701-8573-4ea0-8ccf-a6072d7f67d8 · inbound

Joint Lossless Compression and Steganography for Medical Images via Large Language Models cites this paper.

Joint Lossless Compression and Steganography for Medical Images via Large Language Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T05:31:33.619804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:31:33.619804Z digest=sha256:d2988435450305c5a0b09fd5889a20e4c1371409acb0a6de5581fab14fc4cb29

Observation 6adb60b0-7778-4b42-bb90-895e2d417c12 · inbound

StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion cites this paper.

StarPose: 3D Human Pose Estimation via Spatial-Temporal Autoregressive Diffusion Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T05:19:29.750778Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T05:19:29.750778Z digest=sha256:01786e4bf4573bd8b9fe707fbcd866279030104233c25d7e53e8f5593765fc8a

Observation b364c00e-154f-4734-abc1-0f1e1dc0afa7 · inbound

FuXi-\beta: Towards a Lightweight and Fast Large-Scale Generative Recommendation Model cites this paper.

FuXi-\beta: Towards a Lightweight and Fast Large-Scale Generative Recommendation Model Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-05T20:25:12.063091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:25:12.063091Z digest=sha256:c37fae8c95f9f63912ac1b91c7159920134015a2ebdb3c51c7296c054169266b

Observation 03ea97dd-cb5b-469d-87a8-c4b36faeda82 · inbound

Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing cites this paper.

Discrete Noise Inversion for Next-scale Autoregressive Text-based Image Editing Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:36.753406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:36.753406Z digest=sha256:e632821157bac1caa1063d340008762d6b4bf14738e0f42307747ff645610336

Observation e72b48e1-f76a-42fb-86b3-dcad5db7bf7e · inbound

Scalable Training for Vector-Quantized Networks with 100% Codebook Utilization cites this paper.

Scalable Training for Vector-Quantized Networks with 100% Codebook Utilization Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-04T18:12:47.607231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T18:12:47.607231Z digest=sha256:13434714fb3f4468990911bc96ec6f13f7bb2891ad99e4a56e26b076c023939b

Observation 3f8d344b-98b5-4ca1-9b5b-65080d7c805d · inbound

InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames cites this paper.

InertialAR: Autoregressive 3D Molecule Generation with Inertial Frames Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-04T07:22:17.259251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T07:22:17.259251Z digest=sha256:b57d7e5e6fed62d2f9b24e71b003952f267e596d62dc6e21dc4c2e246a2ffed6

Observation d9c1e3ad-5dd4-4956-8844-f9f8c7b7c7f7 · inbound

REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion cites this paper.

REGLUE Your Latents with Global and Local Semantics for Entangled Diffusion Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-03T15:34:59.272712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T15:34:59.272712Z digest=sha256:ee27e23a2c7fbbea068ba53f1510f14f9c15a7c94c6ab2dbe4bc7fcc1a2c46bd

Observation 2c51f95a-e5c3-48c6-850b-70aafb6c13f4 · inbound

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models cites this paper.

CG-MLLM: Captioning and Generating 3D content via Multi-modal Large Language Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-05-21T14:50:14.905925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-21T14:48:21.787919Z digest=sha256:61eabcb9b5d4a315a531952d75204cc5f3b4a608704c6b4c499364d3e7ef352a

Observation fa7ffe04-5f60-43c7-8e39-b2322a2433d5 · inbound

Language-Guided Transformer Tokenizer for Human Motion Generation cites this paper.

Language-Guided Transformer Tokenizer for Human Motion Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-03T03:25:00.856914Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T03:25:00.856914Z digest=sha256:d19ac6faf93737b6bda56fed1ceaec17abe050da11b5c13cc66c4a722a477372

Observation 8599a202-8d46-4591-8938-d4fbc97564a3 · inbound

Pinterest Canvas: Large-Scale Image Generation at Pinterest cites this paper.

Pinterest Canvas: Large-Scale Image Generation at Pinterest Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-15T13:50:51.666642Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T13:50:51.666642Z digest=sha256:e06bd2857c206feacee6348eab10d13bd719274389c51a3c2b82fc90d0cee506

Observation deddcb0a-0fe0-49c6-8f6f-fab0a95d9d8c · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:01:03.170979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T15:41:23.057949Z digest=sha256:cf92efe6385a7830b809b60d08d6f55f0b8881899d10b136dcede60c40177d00

Observation 46f1a57f-98b3-4157-bcd9-012fd358eb4f · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-12T21:00:35.123495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:00:35.123495Z digest=sha256:18b8d983eafa29226fd790f4a455b6487bf99d600dca38fd9253bee30f70861d

Observation 0469d808-47ad-4bbe-ae39-dfda3f3e0aeb · inbound

Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation cites this paper.

Text-To-Speech with Chain-of-Details: modeling temporal dynamics in speech generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 28

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T01:04:50.133561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T01:01:06.094276Z digest=sha256:24decab8e925a79a04b3968f768ff28f991b5e6e6c8a72089fb89d6c4305b3db

Observation af125df3-6c61-462c-946a-bf973da75c54 · inbound

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings cites this paper.

MMCORE: MultiModal COnnection with Representation Aligned Latent Embeddings Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 34

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T03:29:21.706947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-10T03:27:50.144706Z digest=sha256:2c8df1470a7cd03ab58e578752ba8010e9af08af89579f22dfddfdec1d143988

Observation 4dd55285-e8f5-4959-be4a-0c54419c273a · inbound

Normalizing Trajectory Models cites this paper.

Normalizing Trajectory Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T04:10:55.014588Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-11T01:55:30.852030Z digest=sha256:4463db1bfb9369b414ab4624d9b4d130c0127c2af4b323ffe8ee61654027e0de

Observation 085bce16-a77c-4529-92ba-498749d3a624 · inbound

Normalizing Trajectory Models cites this paper.

Normalizing Trajectory Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 19

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T21:17:58.733808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-14T21:16:53.641457Z digest=sha256:9d7bf13c963e1e79bad48b122ff642b36d8328df3c18a8156791bdda5c3400e5

Observation 5b4b231a-8f83-4a8d-84d3-6bbda913b35f · inbound

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning cites this paper.

TRACE: Temporal Routing with Autoregressive Cross-channel Experts for EEG Representation Learning Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-13T02:27:07.499894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-13T02:22:59.247850Z digest=sha256:72f7e2dd47f5486e6f1cb1c7f477ff34ac7af1258f1cb1a1c3bdaa0d8dc49286

Observation 74e2ea72-7abf-4b0b-9e0f-873877f3d55a · inbound

Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice cites this paper.

Mutual Enhancement Between Global Tokens and Patch Tokens: From Theory to Practice Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 55

Resolution
verified exact
arxiv_id, observed 2026-05-20T22:43:51.045051Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-20T22:41:44.510546Z digest=sha256:78800e6ea06decafa4baa01fce506f31a795b678121b55e40b8b00764d42a2e8

Observation 6e2dedba-bbfb-49dd-8411-78c31672b540 · inbound

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation cites this paper.

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T12:03:15.386946Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T11:59:54.139888Z digest=sha256:173288915e079369874a6faea13fdddcc76e14a8e01d4fd038cbfd630fbbaff0

Observation 24f58dbd-1847-4812-a2d1-7477efe571c5 · inbound

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation cites this paper.

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T18:55:00.880061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-30T18:39:40.667006Z digest=sha256:3a69a705fd427846d397856d3a42b0c3ba1faaf8ca37b3837987fbb2c292a502

Observation c0e8b435-e4df-4f72-a77b-518e4cab39ef · inbound

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation cites this paper.

SRC-Flow: Compact Semantic Representations Enable Normalizing Flows for Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-02T13:49:19.647602Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T13:49:19.647602Z digest=sha256:3cf45c9a35b35bd89ee5bdab7c5ed3854313f58b696f0e6e817c18db07ebf3b0

Observation 6a3834f3-bfc7-4c5b-bde3-a61da64e68a9 · inbound

Vision Foundation Models as Generalist Tokenizers for Image Generation cites this paper.

Vision Foundation Models as Generalist Tokenizers for Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 75

Resolution
metadata mismatch
arxiv_id, observed 2026-05-20T11:03:13.431357Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T11:01:24.738195Z digest=sha256:99fe82bf167e1182f75c403d2b080f3ea1ad6302f5d245405ae5db478a820fc8

Observation dd5b4dfa-05d3-4d4a-b4e3-ea7130a7e7f3 · inbound

Structure over Pixels: Learning Variable-Length Visual Programs cites this paper.

Structure over Pixels: Learning Variable-Length Visual Programs Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T18:13:48.992424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T18:06:01.684713Z digest=sha256:73840a72fc41a42b22c90e3e5bfb2ae44b7938f68e9e8ce1530e998c62338889

Observation 5f206bdb-b56a-4463-90cf-951d1da76bde · inbound

GPIC: A Giant Permissive Image Corpus for Visual Generation cites this paper.

GPIC: A Giant Permissive Image Corpus for Visual Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T07:43:14.161443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-29T07:36:21.262064Z digest=sha256:872bf685212f6d5104e2d8194436de834f731a084e3a70b9dcadbdded2bb894c

Observation 7a9eac4a-74fb-4c34-8054-3bb219dba706 · inbound

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models cites this paper.

Coarse-to-Control: Action-Token Planning for Vision-Language-Action Models Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-02T17:37:14.616849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T21:57:19.195413Z digest=sha256:d4891d033a462128c2843422710f181f748c6cbf45cb77cf030c0c7e706768ff

Observation ea1f55ad-4600-458d-be96-a18ba8af89bf · inbound

OmniGen-AR: AutoRegressive Any-to-Image Generation cites this paper.

OmniGen-AR: AutoRegressive Any-to-Image Generation Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 69

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:47:29.665114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-27T17:05:16.883488Z digest=sha256:54fefc0fb859b6af9085600fa3183d39500e2633a6278518a74202c9fe246673

Observation 9c7ac308-bf41-4d0b-8395-357a56231a49 · inbound

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision cites this paper.

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 40

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T13:39:51.006842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-06-26T05:01:44.347859Z digest=sha256:6dcd93d671e8ec2f269680e34fd9c9ec8c804aeb8d1b165c447b004a8a7673c0

Observation 1c819c78-e0aa-405f-9ed6-28be8e9a3637 · inbound

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision cites this paper.

SubdivAR: Autoregressive Next-Scale Prediction for Neural Mesh Subdivision Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-02T10:07:06.115453Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T10:07:06.115453Z digest=sha256:31907f2bcfb0c1578013b1eec828c061c265f40117095802817a7bdc12f7946b

Observation ba787460-7882-4003-a7cb-b410f43850c5 · inbound

Amplifying Membership Signal Through Chained Regeneration cites this paper.

Amplifying Membership Signal Through Chained Regeneration Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 35

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:45:39.297920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-07-01T06:22:08.140403Z digest=sha256:e73a2156346ab1c4d197c4197dcf71ebbe097eac1ded0473d72f14726dfe2131

Observation 98f86f12-26fd-4466-b869-57bf08f09e9b · inbound

MentalThink: Shaping Thoughts in Mental SVG World cites this paper.

MentalThink: Shaping Thoughts in Mental SVG World Visual Autoregressive Modeling: Scalable Image Generation via Next-Scale Prediction

Reference 41

Resolution
unresolved
no resolver link, observed 2026-07-12T01:50:59.184754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-07-12T01:50:59.184754Z digest=sha256:ab88832d951a162901b74512d0f9846512c8f5093027306d3751b9ef4b154c00