Pith. sign in

Paper Citation Record · LEDGER

NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 23 inbound Pith citation observations for arXiv:2508.10711.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10711 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 23 of 23 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:27:44.293829Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T00:39:17.418338Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation edc6200a-5307-46af-9c92-3cfdb25c77d6 · inbound

Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models cites this paper.

Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-04T08:37:08.967195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T08:37:08.967195Z digest=sha256:7a519ae5d45036c015bea5ee76c4709163bb369d557dcd20488cdbebbfbc5f46

Observation bfdd9846-457f-4c2b-a6a0-fa98b8e20eea · inbound

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks cites this paper.

PlanViz: Evaluating Planning-Oriented Image Generation and Editing for Computer-Use Tasks NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-16T07:00:42.786426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T07:00:18.807922Z digest=sha256:60b5d7dbba05d2b2bf87ce099c596c7c3b09accea30900521cedeba673e29906

Observation 0e7c4310-9f2e-41f7-9796-830036eaa138 · inbound

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens cites this paper.

LLaMo: Scaling Pretrained Language Models for Unified Motion Understanding and Generation with Continuous Autoregressive Tokens NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 62

Resolution
verified exact
arxiv_id, observed 2026-05-16T05:02:19.780996Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-16T05:01:11.880003Z digest=sha256:26217634237457262c1de44149bec160d3272883b7cb38c71b5f63e78879ba9b

Observation e985abad-9503-4258-b108-83fe222dc60d · inbound

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting cites this paper.

Drift-AR: Single-Step Visual Autoregressive Generation via Anti-Symmetric Drifting NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-05-14T22:08:04.859807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-14T22:03:27.981438Z digest=sha256:2898517fca6c6e133d6348119771acabbd72a8220dd921cbb4a60ea791342aae

Observation f9844b41-8d17-4928-ba98-c896658bde83 · inbound

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation cites this paper.

From Broad Exploration to Stable Synthesis: Entropy-Guided Optimization for Autoregressive Image Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-15T12:50:37.073902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-05-15T12:50:13.764159Z digest=sha256:6b6bdb59f2eee6bf999e0a75c0597374b5d606ea94a32aacdb0f1e902a28bb59

Observation 6f6ef27b-52ba-4d5a-9d5b-b6b30c7b23f8 · inbound

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation cites this paper.

MAR-GRPO: Stabilized GRPO for AR-diffusion Hybrid Image Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-05-11T05:40:58.113056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:00:50.105629Z digest=sha256:16fce922a276c9f0d08e614bab8eaf59d58038f66c9465017a4d6e9b689ef183

Observation 4fa18364-f55a-4476-9e81-3af2ae313dab · inbound

Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator cites this paper.

Uni-ViGU: Towards Unified Video Generation and Understanding via A Diffusion-Based Video Generator NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T00:20:55.596964Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T18:34:06.558546Z digest=sha256:30734b9a957caa682b8d911d7be390ce022e55f5408263a3af8abfc1a8e91428

Observation 19bbb28e-7f47-4dba-b76f-fbf1bb4c1b2c · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-05-11T10:01:03.137066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-10T15:41:23.057949Z digest=sha256:4a4333a1ed5470c9043c003c26ced0d30a573a52e31d605c56bd3e1a195f8f8b

Observation 6472e8a1-d8f2-4a8e-8dc2-e3189779db30 · inbound

Generative Refinement Networks for Visual Synthesis cites this paper.

Generative Refinement Networks for Visual Synthesis NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-12T21:00:35.123495Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:00:35.123495Z digest=sha256:20beae6354e64c34125953351411d9aa2981c509386beec4c6f9b16400ea76a9

Observation 26a7f1ea-5c29-474e-8958-77e998b00eb3 · inbound

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations cites this paper.

VibeToken: Scaling 1D Image Tokenizers and Autoregressive Models for Dynamic Resolution Generations NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:51:11.560529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T04:12:58.100610Z digest=sha256:b6f7b6ed88d2fe00dc95eba81cf848d6b4a93e516e79fcbc16b985761df98758

Observation 5cce2100-9b19-4717-9a7d-20fa9edbd5f3 · inbound

Steering Visual Generation in Unified Multimodal Models with Understanding Supervision cites this paper.

Steering Visual Generation in Unified Multimodal Models with Understanding Supervision NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-11T18:41:10.111591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-08T14:48:22.805268Z digest=sha256:2b16ab16e5bc7a6b70e5486df70fb69ec0f4ff702e6bd4920bcc8448f6b4bf91

Observation 757e39f1-ce56-4b56-9bac-b5e9d9e723e8 · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-12T02:21:16.523440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T02:20:25.071428Z digest=sha256:2cb36b069a609850e9eb178827a84f57a5fac1fe90ea80d8e367d3dde5e79c5a

Observation 36307d51-2891-4ee9-a280-0fe07ea3e4e9 · inbound

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation cites this paper.

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 17

Resolution
verified exact
arxiv_id, observed 2026-05-13T06:02:23.624188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-13T05:54:24.248910Z digest=sha256:7cae4f095a4975b99aa0100bbc1c466bf95414a1848df37400a8f33cfe40b50a

Observation 06b86e01-f01e-4af9-962d-200098a59b44 · inbound

Generate "Normal", Edit Poisoned: Branding Injection via Hint Embedding in Image Editing cites this paper.

Generate "Normal", Edit Poisoned: Branding Injection via Hint Embedding in Image Editing NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-12T06:01:23.823688Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-12T04:42:31.547395Z digest=sha256:60e1e005a5c9e7e2e7ca83c7ba83feeffee08656c9fc513eeec84b02042d4bc6

Observation e03e09b7-4b9d-4290-a315-7a79397e72cf · inbound

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning cites this paper.

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-21T04:33:57.668057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-21T04:33:51.535737Z digest=sha256:d468845c0bb171ec54b587996dd88ba4e210b9cb964210a3a681f2c60268f642

Observation acf8a6b2-3271-4ca0-bdcb-b9482baf02d0 · inbound

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning cites this paper.

Uni-Edit: Intelligent Editing Is A General Task For Unified Model Tuning NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-25T05:45:24.244834Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-25T05:43:38.328972Z digest=sha256:f0e1916e1877d31958bfd5decd5e09293d9588ebbaaab8d84d5929020fd9d0fd

Observation 3a776b73-fb74-47fe-80ef-1cbbabe84746 · inbound

DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement cites this paper.

DIVA: Harnessing the Representation Divergence in Unified Multimodal Models for Mutual Reinforcement NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-06-29T23:14:01.723038Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T23:08:57.793923Z digest=sha256:3564341f2ec95ddc4557fa1db3fdbcf4775ba4833ce9f8f521c44711fd681027

Observation 19dfa07d-2dbc-49ff-a6ca-17a5f9b46735 · inbound

Channel-wise Vector Quantization cites this paper.

Channel-wise Vector Quantization NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 36

Resolution
verified exact
arxiv_id, observed 2026-06-29T22:44:01.484864Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T22:40:05.761148Z digest=sha256:973b6b564f57315ee50d0bfac3e36c888914cd1675019b09c8cf1b176456f3a9

Observation 42847af6-2c04-4cfa-a5c8-fda992bb1e6d · inbound

F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation cites this paper.

F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-02T15:47:06.011188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T23:36:27.369551Z digest=sha256:f08ea86f2d4bda24d8300ea12e2c818950a0b38e50940d1eda0374f917adc366

Observation 91eb9690-007e-4ce1-8383-78c7d784bd25 · inbound

ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? cites this paper.

ImageWAM: Do World Action Models Really Need Video Generation, or Just Image Editing? NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-04T00:39:17.419935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T21:02:24.792139Z digest=sha256:d6fcb3fda34dfde139d4eb6cdd563cd4b211ac540803dcd9f17cc4ca15f1abec

Observation e17df1ca-5629-489d-bf29-6cb66225d845 · inbound

Editing Everything Everywhere All at Once cites this paper.

Editing Everything Everywhere All at Once NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 94

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T09:55:40.104960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=arxiv_source observed=2026-07-01T06:10:12.350268Z digest=sha256:656b9fc9a4999f8b775ba5be05b291c3a2339e54c93086f2b07abfcf7e9b0f64

Observation 5245c4f7-8e54-4fe0-ac9f-9dec98ba3ebd · inbound

Generative Models: Principles, Architectures, and Applications cites this paper.

Generative Models: Principles, Architectures, and Applications NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-12T00:31:34.721694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T00:31:34.721694Z digest=sha256:9e4b92ad574ea6eaa3b752e2558c0186eda1600ed581ffcb72cf21caa8ec5423

Observation 0da7dfb6-da9a-4a76-b492-3e0c477d86fd · inbound

On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image Generation cites this paper.

On the Limitations of Cross-Lingual Consistency in Multilingual Text-to-image Generation NextStep-1: Toward Autoregressive Image Generation with Continuous Tokens at Scale

Reference 63

Resolution
unresolved
no resolver link, observed 2026-08-12T12:27:44.293829Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:27:44.293829Z digest=sha256:2429ee7d6f06853c50d161336ba8982ccda2922d63782b4cafb70acad334dd00