Pith. sign in

Paper Citation Record · LEDGER

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation

As of 11 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 0 inbound Pith citation observations for arXiv:2605.30317.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.30317 v1

Coverage vector

measured 52 of 52 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-29T08:00:16.005187Z

measured 52 of 52 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

52 of 52 outbound references displayed

  • verified exact25
  • verified fuzzy0
  • unresolved24
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ed3b6c15-2b04-4260-bcac-4123f1d00003 · outbound

This paper cites GPT-4 Technical Report.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation GPT-4 Technical Report

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.053299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:983fda364f58676cdb1890b66a1c79efe11e8eadde485a67a3f91923fdc3ad34

Observation 93b16958-4035-4284-ab1a-64a6bcaf27dc · outbound

This paper cites Self-rectifying diffusion sampling with perturbed-attention guidance.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Self-rectifying diffusion sampling with perturbed-attention guidance

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:40f5d88059da2383ee01afa4f1588ce25267d1baf83cc7fc41d7697ec362b050

Observation 3f27767c-42e8-4eea-8fd4-6c305beda22c · outbound

This paper cites Findings of the Association for Computational Linguistics: ACL 2022 , publisher =.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Findings of the Association for Computational Linguistics: ACL 2022 , publisher =

Reference 3

Resolution
metadata mismatch
doi, observed 2026-06-29T08:03:13.618766Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:2710a6b6f546f201ba8d636dd1978c858a54d8f0e5179892972da3db6ba0b981

Observation 95dc5809-f98c-4ea4-993c-0d9e3324bd7a · outbound

This paper cites Scheduled sampling for sequence prediction with recurrent neural networks.Advances in Neural Information Processing Systems, 28, 2015.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Scheduled sampling for sequence prediction with recurrent neural networks.Advances in Neural Information Processing Systems, 28, 2015

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:afac7d0de6a61b3a0a8bde9870e5b5eaeb4e72685cfbd350d687e02aaa7ecc17

Observation 98c96b3d-7a37-4884-8ca2-439d4a4de3a6 · outbound

This paper cites Generative pretraining from pixels.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Generative pretraining from pixels

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:55e9bb399a92a34a6e53e1c3d6711ca30a54c1164d6629982382d743ec04a31c

Observation 6b471dcf-9997-4c40-9171-3ea36359ac57 · outbound

This paper cites Emerging Properties in Unified Multimodal Pretraining.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Emerging Properties in Unified Multimodal Pretraining

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.050583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:be17b9cd90bb10c2cb775de1f16d53eb72f35ec8b2593362b15af720f3cb29d7

Observation a4c74d2e-963c-48bc-b7c8-d6cbfe9d16d4 · outbound

This paper cites Imagenet: A large-scale hierarchical image database.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Imagenet: A large-scale hierarchical image database

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:7c149c802e0e59ca578f20a5c450474fd9a0d04bcf25e3134555ad90d97c5b89

Observation 400e1312-d99b-42a3-86e5-8dd993ef4428 · outbound

This paper cites Taming transformers for high-resolution image synthesis.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Taming transformers for high-resolution image synthesis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:46cae7dc7c2bad89fc09ec53726f33e097ab29cebbd3aea260429e078abd26fb

Observation 73b74f10-6c50-46b8-abab-0853bdd662b1 · outbound

This paper cites Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Geneval: An object-focused framework for evaluating text-to-image alignment.Advances in Neural Information Processing Systems, 36:52132–52152, 2023

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:1238f39161b2243717ec610b2383a4767c60b7b962acb3f51038606609da0eae

Observation 0b7a5381-863c-4e21-8f0c-aa7c8c0a3322 · outbound

This paper cites Deep autoregressive networks.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Deep autoregressive networks

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:e0b08d4cf55ebaf880aa9a0cf07642f0e5f4e7162da83bcd596f4e95debe243b

Observation 0e3b39a0-77ec-48da-a31f-42593d997426 · outbound

This paper cites Suboptimal behavior of bayes and mdl in classification under misspecification.Machine Learning, 66(2):119–149, 2007.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Suboptimal behavior of bayes and mdl in classification under misspecification.Machine Learning, 66(2):119–149, 2007

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ff44abd4bfba239f38f277a6a28cb7a8b4ed27a1e7a42f07e858dc1a4d8e63f1

Observation 31124dfc-aa24-450a-b2dd-c8e59e2e13a3 · outbound

This paper cites Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infinity: Scaling bitwise autoregressive modeling for high-resolution image synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:fe591ff38bcb0166697e379efc36eabdab7f3269eeaee359a21ae4edcf7c9d3f

Observation a3df8ecd-32d0-4cc0-ba45-436ff05cbab3 · outbound

This paper cites Conceptrol: Concept Control of Zero-shot Personalized Image Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Conceptrol: Concept Control of Zero-shot Personalized Image Generation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.006326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:d7e3b759d1c62fc9fac2a4a9973950785dd78ede94c5b4c3b68b14725a45e828

Observation f607d092-18b6-4404-b2c6-6e747271918c · outbound

This paper cites AID: Attention Interpolation of Text-to-Image Diffusion.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation AID: Attention Interpolation of Text-to-Image Diffusion

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.041712Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4a134dc1f720a84d756cd4b4f5df268a664874c3b941ff29646d5d6b7b3fa2e3

Observation 9083b459-91cf-404e-9352-98adfe80ea31 · outbound

This paper cites REAR: Rethinking visual autoregressive models via generator-tokenizer consistency regularization.arXiv preprint arXiv:2510.04450, 2025.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation REAR: Rethinking visual autoregressive models via generator-tokenizer consistency regularization.arXiv preprint arXiv:2510.04450, 2025

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.991902Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:5cae4a0fc941f67d7cd918a3294309d846eba7c165a346762719acbb42e76985

Observation 75051b49-dd15-4a8d-92f0-1338ca6f9fb7 · outbound

This paper cites Classifier-Free Diffusion Guidance.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Classifier-Free Diffusion Guidance

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.003382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:832d24215e85b349fba423a1472736ac5e4d89f0921a6737fa4d727551eb0da9

Observation cbf2529d-be3f-4743-9b91-f1f78e4566c5 · outbound

This paper cites Smoothed energy guidance: Guiding diffusion models with reduced energy curvature of attention.Advances in Neural Information Processing Systems, 37:66743–66772, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Smoothed energy guidance: Guiding diffusion models with reduced energy curvature of attention.Advances in Neural Information Processing Systems, 37:66743–66772, 2024

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ea2317535f561b9e06edf495704a31eb7965647904ea824d187348a058d74b5d

Observation 6057887a-a83a-4bf6-be73-32b9d480e92b · outbound

This paper cites ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation ELLA: Equip Diffusion Models with LLM for Enhanced Semantic Alignment

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.026465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:dabd899977913325c802ac8c89557bb8848f841115cf98ee9c0b3cfdc053202e

Observation 2c046fa5-3350-4f41-b1e6-556a71750e6f · outbound

This paper cites Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Self Forcing: Bridging the Train-Test Gap in Autoregressive Video Diffusion

Reference 19

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:13.986844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:398f526bbe85a5f2286b8b55c30c652a5bfc8d58b8ee28509d0ba18de2b61922

Observation 07eb17f9-7820-462b-b61b-488dff237447 · outbound

This paper cites Vbench: Comprehensive benchmark suite for video generative models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Vbench: Comprehensive benchmark suite for video generative models

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:3b663a0d3d7bbdd9a3b5d8c6aa79d7754ff18badc275dde422494286b34b54a4

Observation c6c66a4c-7d7d-4b70-b362-002311b3800d · outbound

This paper cites Guiding a diffusion model with a bad version of itself.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Guiding a diffusion model with a bad version of itself

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b98f17d04754f1de18ae2e5922d5a5c3ff57856e66f8d5c5b42e9f17495dd123

Observation bef1f905-a5e6-4dca-8756-6ba81abe32d8 · outbound

This paper cites HunyuanVideo: A Systematic Framework For Large Video Generative Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation HunyuanVideo: A Systematic Framework For Large Video Generative Models

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.000811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4440d8838b3400e0953aa8e1626e67b5f69d711cd385a2e9fbccdb1a3d281773

Observation 2c2eadf8-e9c0-40ae-b4fe-2014487d94d9 · outbound

This paper cites FLUX.2: Frontier Visual Intelligence.https://bfl.ai/blog/flux-2, 2025.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation FLUX.2: Frontier Visual Intelligence.https://bfl.ai/blog/flux-2, 2025

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:f61005cf616b898bb7f69e83c8d6efaa407fa07b3dcb592ff38ac8aae6d5e1de

Observation 5d139193-acd6-4058-8e74-0999205b0ab0 · outbound

This paper cites Lamb, Anirudh Goyal, Ying Zhang, Saizheng Zhang, Aaron C.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Lamb, Anirudh Goyal, Ying Zhang, Saizheng Zhang, Aaron C

Reference 24

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:eef9cb7fdc14cb8b0ff014642b6fb2cf27219e1f3ce4adf22612e12fdffa478c

Observation a4e4e4e6-d8eb-4a08-8fc3-28f786052ca0 · outbound

This paper cites Autoregressive image generation without vector quantization.Advances in Neural Information Processing Systems, 37:56424–56445, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Autoregressive image generation without vector quantization.Advances in Neural Information Processing Systems, 37:56424–56445, 2024

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:2c46113ed0c312b96e62e2c3cf819bd3f03d32c54d6bf86725c55da09e9fa219

Observation 8a83ec89-a026-414e-8ce8-b7400950421a · outbound

This paper cites arXiv:2512.19680 (2025) 5.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation arXiv:2512.19680 (2025) 5

Reference 26

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T08:03:14.044788Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:175126bb37a1c6db4f44a1b63599e39d88a38ed89dddb2cafdbb9dc43117bd53

Observation e2f44700-b3bf-4fee-b294-a22df8a5e414 · outbound

This paper cites Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Step-level Reward for Free in RL-based T2I Diffusion Model Fine-tuning

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.984555Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:c1bbf4e56fa28ad8ff214f08b948849cc9873de6889e960665c78482724eac29

Observation 3e2fce92-ecd4-4df1-ace6-002cd328c80c · outbound

This paper cites Infini- tystar: Unified spacetime autoregressive modeling for visual generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infini- tystar: Unified spacetime autoregressive modeling for visual generation

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:13.998565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:213b089a5cfbcf036b0a00305e088139e523860d2f775b6e303bc656bb3497d8

Observation 4db29d96-217c-45f7-a594-85352b5f3947 · outbound

This paper cites Interp3d: Correspondence-aware interpolation for generative textured 3d morphing.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Interp3d: Correspondence-aware interpolation for generative textured 3d morphing

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:5757b567cb7a6ba8475268f2ef4c46b25012b4a69f891173218df69203c83d03

Observation 90e4471b-9750-40f8-a803-89c407d567c9 · outbound

This paper cites Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Unitok: A unified tokenizer for visual generation and understanding.arXiv preprint arXiv:2502.20321, 2025a

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.038473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:05f1e8b36562b2787e772bec012f68e358e8eaca6c4cfdec69a35488074bc187

Observation b4fce583-aa95-4a38-ac42-7b53996283d3 · outbound

This paper cites Finite Scalar Quantization: VQ-VAE Made Simple.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Finite Scalar Quantization: VQ-VAE Made Simple

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.035462Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:da92b42bdaa8d45c6ba036e7df8cce7dd1ae2c9af866418f991826dea275abd7

Observation 18777129-4c80-40b6-aa14-d1d680a1f58b · outbound

This paper cites Image transformer.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Image transformer

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:f2540611033ab634c00a07003b13e0bd4f9bef3f6df1d246a9a4081dcf3edd9f

Observation 2ef82b06-ad04-43e7-b1e7-e8bf54ea51ba · outbound

This paper cites Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Beyond Next-Token: Next-X Prediction for Autoregressive Visual Generation

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.032693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:d5ca1131446e6d1c27b9678a1e0bfc68abcaeee65e40ec3553e305d2cb2512e7

Observation 2e13bba3-f311-43e2-b66f-5e8e7ace93e9 · outbound

This paper cites A reduction of imitation learning and structured prediction to no-regret online learning.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation A reduction of imitation learning and structured prediction to no-regret online learning

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:bfe9fe3b341590717af7d8abbe858de34f5e12f4f83b69a5709a8cf2594b7c9b

Observation 1b532297-0cc7-4230-9e72-7bd436286f91 · outbound

This paper cites Generalization in generation: A closer look at exposure bias.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Generalization in generation: A closer look at exposure bias

Reference 35

Resolution
verified exact
doi, observed 2026-06-29T08:03:13.622848Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:f14baf9233e5ff764863aaebf62e9d9a61ca714439e5f805e36d2f45ac410214

Observation a0dab1f5-8885-4b0c-985f-dad850f4a79d · outbound

This paper cites SSG: Scaled spatial guidance for multi-scale visual autoregressive generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation SSG: Scaled spatial guidance for multi-scale visual autoregressive generation

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b64a7362afa3566e965b96201146f8dad0e88410adcfdd8d08ed0ce45fc6e605

Observation 6fea1418-0209-443c-9c0b-52757ead38d6 · outbound

This paper cites Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Autoregressive Model Beats Diffusion: Llama for Scalable Image Generation

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.020689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:14be52183b9a92cdb879c8b2753c2834320673d0cf3443f1718af768308f3fdf

Observation 4991b45d-f9ca-4900-a43a-82df66351c11 · outbound

This paper cites Chameleon: Mixed-Modal Early-Fusion Foundation Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Chameleon: Mixed-Modal Early-Fusion Foundation Models

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.011397Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:f4831bbaa6f64d2835bcdf62120e1b53c381aa466d4dd327604cc279dbb8a447

Observation fa2401a5-0ed9-4ae7-a9aa-c4162a9fbe12 · outbound

This paper cites Gemini: A Family of Highly Capable Multimodal Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Gemini: A Family of Highly Capable Multimodal Models

Reference 39

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.005561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:c3641c3b252242c5a86361f8c20a04ba2fb0c5d8c04e9075463dfe9ee350bbde

Observation c0a7076f-c789-4122-9f2e-e96ed35b0d93 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Visual autoregressive modeling: Scalable image generation via next-scale prediction.Advances in neural information processing systems, 37:84839–84865, 2024

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:d463513cf8c8afe97c0da621fc0d6f156237bcd6826998ce8585eead7e836500

Observation 6c1dd962-3c73-45f3-80bc-6ef55b7e4142 · outbound

This paper cites Neural discrete representation learning.Advances in neural information processing systems, 30, 2017.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Neural discrete representation learning.Advances in neural information processing systems, 30, 2017

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:d1cb2a0cde95208bd1abeb78c667f14bdceda0a40ea13a9fe8dae0af5d1247a3

Observation 1151c0a7-cda7-4d45-8b4e-cc39787496fb · outbound

This paper cites Wan: Open and Advanced Large-Scale Video Generative Models.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Wan: Open and Advanced Large-Scale Video Generative Models

Reference 42

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.029311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:76dba9107bf0162085c8118aa39f42d12b6f3bc0cd334fe51b87853002519a52

Observation 8c1d0f35-8b15-4af2-8133-d6d0c427f24b · outbound

This paper cites On Exposure Bias, Hallucination and Domain Shift in Neural Machine Translation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation On Exposure Bias, Hallucination and Domain Shift in Neural Machine Translation

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.047424Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:aa26c3b0bf6d705fc8aacbb065134e3d8c22fcf1af01b65d24f99ad4dbccdfaf

Observation 0ed9c76e-7e73-46df-9462-2063178dcb34 · outbound

This paper cites Blaschko.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Blaschko

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:524040226df0c4e3fc08dda71dc45eeb2a1e314f345032b4d67f7b638d9eb831

Observation c7df9d30-7ee3-4de0-a9a7-7c3d1b2aa685 · outbound

This paper cites an unresolved cited work.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Unresolved cited work

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:1527df5aa2afdcce39ac2ff59489a1e746cf6433a276d8f398149bf070e372d8

Observation 72d0987f-c2ca-4c5e-a5a5-0a17ef30a38b · outbound

This paper cites Infotok: Adaptive discrete video tokenizer via information- theoretic compression.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Infotok: Adaptive discrete video tokenizer via information- theoretic compression

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.024593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:7616beb4c2881e9a141de622c580126d65b4b3fb07b4bd4a9faa1f06e0c37081

Observation 601e44bc-c052-4d4b-89f9-e011dd6a7e21 · outbound

This paper cites Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Language Model Beats Diffusion -- Tokenizer is Key to Visual Generation

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.019333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:b849ccff4641fd40766533285469d69b8c5ddd0920d31ca247373eec2f4bec16

Observation d2b5d881-b1e2-413b-b412-6f87a6623a31 · outbound

This paper cites An image is worth 32 tokens for reconstruction and generation.Advances in Neural Information Processing Systems, 37:128940–128966, 2024.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation An image is worth 32 tokens for reconstruction and generation.Advances in Neural Information Processing Systems, 37:128940–128966, 2024

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-29T08:00:16.005187Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:cf2228a8e82b910b108a5d895c9864596fcfc1f6c35a25a3d18f56c9a59108c8

Observation f6202b98-316e-46fa-b1a7-5f1084975064 · outbound

This paper cites Guiding a Diffusion Model by Swapping Its Tokens.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Guiding a Diffusion Model by Swapping Its Tokens

Reference 49

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.023553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:df8f81edebdc00dfb92183afe626262f7d830f0b8681ec002a65be5283218d15

Observation 02c37d58-e6fb-466a-b1e0-f19f5d5dcb95 · outbound

This paper cites author Feng, Y.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation author Feng, Y

Reference 50

Resolution
metadata mismatch
doi, observed 2026-06-29T08:03:13.620657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:4117bf3d6133960710ef5439b0b604a3d086df86d0e6110aed62cae2b64c610a

Observation 5195f449-27e6-47ae-aa2a-7076b077e4c8 · outbound

This paper cites Image and Video Tokenization with Binary Spherical Quantization.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation Image and Video Tokenization with Binary Spherical Quantization

Reference 51

Resolution
verified exact
arxiv_id, observed 2026-06-29T08:03:14.008832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:567afd149b28e7a8290fb65f6c3b6dbc84fb4f9a643101f7a2f738dfa7b2070b

Observation 14e6710d-d70c-4bd7-a6ca-16c18b370ed5 · outbound

This paper cites RelaxFlow: Text-Driven Amodal 3D Generation.

VPG: Visual Prefix Guidance for Autoregressive Image and Video Generation RelaxFlow: Text-Driven Amodal 3D Generation

Reference 52

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:03:14.017742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T08:00:16.005187Z digest=sha256:ad0389b7df40b350def53ff869c710846a9f496fdf927d9056f3a9c8565023f4

Pith citing papers

No inbound Pith citation observations are available.