Pith. sign in

Paper Citation Record · LEDGER

Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 19 inbound Pith citation observations for arXiv:2503.13436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13436 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 19 of 19 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 19 of 19 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:30:41.596688Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d29f49f-1b5b-4bdc-a518-b8d1ffc5bbed · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.912974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:30ad6febb9316bea88009d0374426db7096a515be4e9663bfcaf51ae63c925ac

Observation 33621fbe-ebb6-43a6-a7a6-bd47ca6318c0 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.596688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.596688Z digest=sha256:0b712ba29c7c425713b38c9028ad6eb56a8c5abc541a3d373649869e99714c67

Observation ae8ad5d0-247d-4c36-9353-0225eb3b27b7 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:16.091886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:02485b3af74131b553cf5d580c4b0c3632e096f27dd8cc00506e425141520845

Observation 254445f6-6f51-42b2-a357-beb673a5d945 · inbound

NeoBabel: A Multilingual Open Tower for Visual Generation cites this paper.

NeoBabel: A Multilingual Open Tower for Visual Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:15:26.807532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:15:26.807532Z digest=sha256:cee23a62049f4e6237d48ea0476973e66ce921ac5ecdade2a1d279edea5747fb

Observation 3e050806-5d68-47c6-a051-591c3ae8a5e5 · inbound

Generative Distribution Distillation cites this paper.

Generative Distribution Distillation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:10:22.068784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:10:22.068784Z digest=sha256:e91864b5499fffd55943936c69e5183c06daaab67e9d1fa7fec1007248f83bec

Observation 4afd479f-c989-4a8c-92b3-2384f48e0880 · inbound

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation cites this paper.

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T04:37:13.531723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:37:13.531723Z digest=sha256:c1d352393ff0bb192b2290f2f59bac75c853d4eeaf16738ec833f693aed8946d

Observation 03e24383-60cf-4717-b0b4-e9cecdc8022f · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:07.931702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:07.931702Z digest=sha256:b9259a9b02171432b4ad5c58c1504262910c6937c75048d0720e0ca0b5d5c0a2

Observation 9a95757f-7988-4315-9698-f93aab884bd3 · inbound

Image Diffusion Preview with Consistency Solver cites this paper.

Image Diffusion Preview with Consistency Solver Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:43:34.705010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T21:42:24.098676Z digest=sha256:2e1fb3bf67c10ccf04f23419e563c3df7f275a8959787b75ef53fe815690d782

Observation bbaeaf67-53dd-49ed-ac6f-cdd0282a82a3 · inbound

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation cites this paper.

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T07:03:12.169209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:03:12.169209Z digest=sha256:c387c281560dbc7eb90f3e58d6d1944c5f9cd954f23f892cf92fdd8ef8f3fa0f

Observation c6306d4c-b4f4-41ef-8581-3d1ac64ef55a · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:52:59.401742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T16:51:48.705876Z digest=sha256:9d0960f5e7248b32cfe5a547e1b65227a87ab64774c8203a8ad234bcd9137a7a

Observation fccf5d1f-9135-4504-a707-6f9f71dea4b6 · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T12:10:53.720348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:10:53.720348Z digest=sha256:43c1d267f9183197da38c876458857c32ce72e3c8366f11eacb9ced0d2a84b28

Observation d273a90e-10cc-40a3-b427-df164a0fbe4c · inbound

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models cites this paper.

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:59.145717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T16:27:14.491492Z digest=sha256:68f74da4cdfb8a35115582e3817a0c6ea4bc0e02c915a09a47744def3ada18a1

Observation 61a52f06-dec5-4951-99fa-95a5ef2d98de · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.823854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-08T04:31:26.325118Z digest=sha256:9c9e7d6532f910d86189a6181355e7cf61f0a2420a7c07355532ccd400a56f76

Observation 1683cb3c-e9e0-4994-9a8a-409e6b24e763 · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:43:51.078460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T23:41:25.275207Z digest=sha256:79da002c55816cc2c48b578b0ce6bf10ea96eb26ad13982ea64d11d1cd196da6

Observation b91df5b1-f000-4f06-85b7-66cb885b08e3 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.918323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:2121c5eb3bdb2229a9003b55b3e9f322efcafb67dd1fc95ea0d75d64ad12b649

Observation d5172ce2-5713-4439-8ec6-0c928cda97c6 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.752143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:c41a0b1f043eb3a2d027506d0d37c00ea3c92cd1dfe50f09776ef4ce7a3bd1c2

Observation fc4b3c3a-b56b-4259-a0e3-9a079cd900d8 · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:32.117532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:c0dc0eb822a7440799b3ebf140709fd1a5f4217fb6a1ea707ea9bd8b7c076797

Observation cc2b5ce7-2ac5-4c70-b38c-bf713142dea9 · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.861564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.861564Z digest=sha256:686e07b1f2f6ffd9d74b1cd1c73a5e1813b4f1e3e7efe00ec92162df2d29357e

Observation 951e616b-0053-4d9b-a2ec-a51aa180cb9d · inbound

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications cites this paper.

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T01:51:57.970306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:51:57.970306Z digest=sha256:84f520427bfe24ea8a8891a9eb35fbcbe35e46d51148c541540a9ec4c42e9965