Pith. sign in

Paper Citation Record · LEDGER

Autoregressive Models in Vision: A Survey

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 18 inbound Pith citation observations for arXiv:2411.05902.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.05902 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 18 of 18 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T04:56:00.670435Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T12:49:52.859276Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation bce9d780-065b-4b7d-85bf-3a1fd0ee2013 · inbound

PanoLlama: Generating Endless and Coherent Panoramas with Next-Token-Prediction LLMs cites this paper.

PanoLlama: Generating Endless and Coherent Panoramas with Next-Token-Prediction LLMs Autoregressive Models in Vision: A Survey

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-12T13:54:33.859863Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T13:54:33.859863Z digest=sha256:012c8f4eee92762644143ab8f40e85460836acedcad113b7bcfd34c08bb5b7bc

Observation 3b0035df-556a-42d0-a748-2c8e81ab871e · inbound

Identity-Preserving Text-to-Video Generation by Frequency Decomposition cites this paper.

Identity-Preserving Text-to-Video Generation by Frequency Decomposition Autoregressive Models in Vision: A Survey

Reference 75

Resolution
unresolved
no resolver link, observed 2026-08-12T12:10:27.372712Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:10:27.372712Z digest=sha256:cb6a54fbd493be26283650fc3171fe8f92b28e9e6d6f7130a1685e328a4b3ad4

Observation ffca2200-517f-416f-a903-5e3ab4cda93e · inbound

[MASK] is All You Need cites this paper.

[MASK] is All You Need Autoregressive Models in Vision: A Survey

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T19:29:21.205007Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T19:29:21.205007Z digest=sha256:1a2ab12c8b98d9f0fe262ebf0109a4bfaf3d4f4c9140c32e3a2996e08e8cf23f

Observation 69d20af0-c0a3-486f-8321-fc1f868b1e71 · inbound

A Survey of Interactive Generative Video cites this paper.

A Survey of Interactive Generative Video Autoregressive Models in Vision: A Survey

Reference 71

Resolution
unresolved
no resolver link, observed 2026-08-16T04:56:00.670435Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T04:56:00.670435Z digest=sha256:01fe28b806b9f62924c3be23d0107e2c7fa8765ebe56ef301851e3c46db8fb33

Observation f6a1f174-c10f-490f-b299-990ce099dca8 · inbound

FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge cites this paper.

FastCar: Cache Attentive Replay for Fast Auto-Regressive Video Generation on the Edge Autoregressive Models in Vision: A Survey

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-15T20:54:59.048394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:54:59.048394Z digest=sha256:a037fb26c5fc9939868d31a27a4abd2cd4740e08d8fd35f3c61192a3afbcc2f0

Observation b9586c45-5905-4b34-b0da-3be87381ae11 · inbound

Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward cites this paper.

Multimodal Generative AI with Autoregressive LLMs for Human Motion Understanding and Generation: A Way Forward Autoregressive Models in Vision: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:06:23.551098Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:06:23.551098Z digest=sha256:1416b2d5c90c01698a5ebcc5045475727d2bfe22232ff3279104cb49d94ee9f0

Observation 78fdd273-787d-4a46-bdf7-d02c1498ef5f · inbound

BulletGen: Improving 4D Reconstruction with Bullet-Time Generation cites this paper.

BulletGen: Improving 4D Reconstruction with Bullet-Time Generation Autoregressive Models in Vision: A Survey

Reference 80

Resolution
verified exact
arxiv_id, observed 2026-05-19T08:03:01.094244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T08:02:14.123151Z digest=sha256:cef50d5074b8c23244ea8c76d6c31b804c9bd28a3b0d8308c8c0491bb0db63ce

Observation 5b25831e-5295-43d6-83c5-bfc7c54cf5a3 · inbound

A Survey on Vision-Language-Action Models for Autonomous Driving cites this paper.

A Survey on Vision-Language-Action Models for Autonomous Driving Autoregressive Models in Vision: A Survey

Reference 137

Resolution
unresolved
no resolver link, observed 2026-08-06T21:31:04.806091Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T21:31:04.806091Z digest=sha256:ce153c085a4b2e1692bba1b7a08178865d35308c8cc5738d1f340f9214a0edb0

Observation b0b0ce77-cd97-4ce9-b230-1575e9c3cc68 · inbound

A Survey on Training-free Alignment of Large Language Models cites this paper.

A Survey on Training-free Alignment of Large Language Models Autoregressive Models in Vision: A Survey

Reference 92

Resolution
unresolved
no resolver link, observed 2026-08-05T21:18:49.044177Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T21:18:49.044177Z digest=sha256:dee11b4d7fccba254c8b885781ddf2c105b586d44888452c8ea056f93d57ef12

Observation 5ccadcb1-b9c3-4187-8069-e6163e26585e · inbound

Reinforced Context Order Recovery for Adaptive Reasoning and Planning cites this paper.

Reinforced Context Order Recovery for Adaptive Reasoning and Planning Autoregressive Models in Vision: A Survey

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T17:23:20.480169Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T17:23:20.480169Z digest=sha256:bf78a7a475cea8ceb6800673e8ba939c564f3b42a031107628ab149c860d96bd

Observation e5ef6492-9158-4167-bb9e-38560576de78 · inbound

A Unified Low-level Foundation Model for Enhancing Pathology Image Quality cites this paper.

A Unified Low-level Foundation Model for Enhancing Pathology Image Quality Autoregressive Models in Vision: A Survey

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-05T13:00:48.189650Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T13:00:48.189650Z digest=sha256:9ee2ca74f9441d3b25eef471dda748d7a80fb56a11bf66c658fe5394095cb09f

Observation b63d9169-def5-4af6-ab4a-891d9b6a018c · inbound

Generative AI Meets 6G and Beyond: Diffusion Models for Semantic Communications cites this paper.

Generative AI Meets 6G and Beyond: Diffusion Models for Semantic Communications Autoregressive Models in Vision: A Survey

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-17T23:40:31.101868Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-17T23:40:01.988138Z digest=sha256:98d5ec03d656c13b2c42730486ddf9576622d9d01f2278ae8ff6f3bdf9c7836c

Observation 32756b07-7c5a-4a08-8075-2cd47f5efb29 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs Autoregressive Models in Vision: A Survey

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-11T16:01:22.428411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T18:53:06.494640Z digest=sha256:67f5d06cd54cf450d2d512a62b841ee89a58c28dea8ab159fbacc25c8fc608c4

Observation feacafef-7fb4-4d97-94b4-479f6d129e74 · inbound

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs cites this paper.

Persistent Visual Memory: Sustaining Perception for Deep Generation in LVLMs Autoregressive Models in Vision: A Survey

Reference 85

Resolution
verified exact
arxiv_id, observed 2026-05-11T01:50:51.575223Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-11T01:49:15.136031Z digest=sha256:b11ecf6fdd08aef440987d57015e082a703956156757ac971e0a12605e909acb

Observation a077f488-522c-4d11-a33e-fa4372dce36e · inbound

PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images cites this paper.

PathAR: Structure-First Autoregressive Synthesis of Multimodal Pathology Images Autoregressive Models in Vision: A Survey

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-07-01T22:06:15.894857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-28T15:50:54.950899Z digest=sha256:e823b1692707a94d2bc5c56fbbc6daffe1446b5ca9eb7f0ac36fa67a28f733bd

Observation 23c1441d-9ebf-42aa-9d12-905d76aafc44 · inbound

NEXUS: Neural Energy Fields for Physically Consistent Contact-Rich 3D Object Dynamics cites this paper.

NEXUS: Neural Energy Fields for Physically Consistent Contact-Rich 3D Object Dynamics Autoregressive Models in Vision: A Survey

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-03T17:18:43.309763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-27T04:25:27.067362Z digest=sha256:9a59122068b95fdd965610d02038eab5cdc5ace70045b75cc886a05ce73b8b5d

Observation 4e4dd0b6-d32f-4088-9f05-e56e3726c784 · inbound

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks cites this paper.

Safe Autoregressive Image Generation with Iterative Self-Improving Codebooks Autoregressive Models in Vision: A Survey

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-04T12:49:52.860606Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T05:51:01.446139Z digest=sha256:19960dc226ad4272880faee3e3435a96c156f06a7d25db73718e8a757eb2ff15

Observation 76e2d6b2-bba3-4df6-8425-cba8a7dc73e6 · inbound

Does More Retrieved Evidence Help Visual Retrieval-Augmented Generation with Diffusion Language Models? cites this paper.

Does More Retrieved Evidence Help Visual Retrieval-Augmented Generation with Diffusion Language Models? Autoregressive Models in Vision: A Survey

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-10T16:51:18.564403Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T16:51:18.564403Z digest=sha256:854f1674c2e2c1f9863be189a27cb3a85c0f480fd3e0212136b8ecb05bacea16