Pith. sign in

Paper Citation Record · LEDGER

Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 20 inbound Pith citation observations for arXiv:2503.13436.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2503.13436 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 20 of 20 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 20 of 20 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:16:36.394421Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 0d29f49f-1b5b-4bdc-a518-b8d1ffc5bbed · inbound

Step1X-Edit: A Practical Framework for General Image Editing cites this paper.

Step1X-Edit: A Practical Framework for General Image Editing Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-11T14:36:41.912974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-11T14:36:41.467429Z digest=sha256:c3d8849884d00c148e1a9f56f915d38b49b8784ebbc33daa4006dff12cfeaae2

Observation 1b0cb649-3dc8-4e6e-87d6-6df936d3e7e7 · inbound

VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation cites this paper.

VTBench: Evaluating Visual Tokenizers for Autoregressive Image Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T20:16:36.394421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:16:36.394421Z digest=sha256:bf3c71e38d7e7338411aadde2c9af6ef4a243b206a25173d96ce5c0af085cc11

Observation 33621fbe-ebb6-43a6-a7a6-bd47ca6318c0 · inbound

Fake it till You Make it: Reward Modeling as Discriminative Prediction cites this paper.

Fake it till You Make it: Reward Modeling as Discriminative Prediction Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T00:30:41.596688Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:30:41.596688Z digest=sha256:6e07b384987fcf7fb6bd87685d0ce0be97aee53e4cf23bb7cd071723b852337c

Observation ae8ad5d0-247d-4c36-9353-0225eb3b27b7 · inbound

Show-o2: Improved Native Unified Multimodal Models cites this paper.

Show-o2: Improved Native Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-12T18:51:16.091886Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-12T18:51:15.428692Z digest=sha256:87458d76233d3032f0f85b76ee9785d007e9cfb6e0be4a9338c74ce40e6b2c2c

Observation 254445f6-6f51-42b2-a357-beb673a5d945 · inbound

NeoBabel: A Multilingual Open Tower for Visual Generation cites this paper.

NeoBabel: A Multilingual Open Tower for Visual Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-06T19:15:26.807532Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:15:26.807532Z digest=sha256:bee0b9adc1a7dfa0df3cfd8fdc74e3063db70d5dd859e092f9b8c547b10f92d5

Observation 3e050806-5d68-47c6-a051-591c3ae8a5e5 · inbound

Generative Distribution Distillation cites this paper.

Generative Distribution Distillation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T16:10:22.068784Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:10:22.068784Z digest=sha256:18fd906de2aced237ce589e07794b2f0687eb7e7a23caf029a629c94c5272aaa

Observation 4afd479f-c989-4a8c-92b3-2384f48e0880 · inbound

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation cites this paper.

Skywork UniPic: Unified Autoregressive Modeling for Visual Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T04:37:13.531723Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T04:37:13.531723Z digest=sha256:2ea52884d8511943e43a5003be9067828322f4720fa30b3bb25e2b55eea18f06

Observation 03e24383-60cf-4717-b0b4-e9cecdc8022f · inbound

Reconstruction Alignment Improves Unified Multimodal Models cites this paper.

Reconstruction Alignment Improves Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-04T22:36:07.931702Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T22:36:07.931702Z digest=sha256:43b1f7a920851bab24c9724128110017279e9f38f8495404a5560b1a50a6d70f

Observation 9a95757f-7988-4315-9698-f93aab884bd3 · inbound

Image Diffusion Preview with Consistency Solver cites this paper.

Image Diffusion Preview with Consistency Solver Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-16T21:43:34.705010Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-16T21:42:24.098676Z digest=sha256:412cbf31363d865c4ae27403399d655b1af3d75397dfabdfd788e0db1160a89f

Observation bbaeaf67-53dd-49ed-ac6f-cdd0282a82a3 · inbound

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation cites this paper.

Generation Enhances Understanding in Unified Multimodal Models via Multi-Representation Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-03T07:03:12.169209Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T07:03:12.169209Z digest=sha256:81eeddc6762aae8b55d0ae45a91e351bcede5898aad42aa21d09d984f8f5a345

Observation c6306d4c-b4f4-41ef-8581-3d1ac64ef55a · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T16:52:59.401742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-13T16:51:48.705876Z digest=sha256:4d34d2f4d55d17cc82c51d700e0954727ada167de926d142677f5659ccd2916a

Observation fccf5d1f-9135-4504-a707-6f9f71dea4b6 · inbound

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning cites this paper.

FeynmanBench: Benchmarking Multimodal LLMs on Diagrammatic Physics Reasoning Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 14

Resolution
unresolved
no resolver link, observed 2026-07-13T12:10:53.720348Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T12:10:53.720348Z digest=sha256:e1af76a4ca2c5557ed8e8af74aab1ff734b8b6ca067ad26e5e339dba5c720a9d

Observation d273a90e-10cc-40a3-b427-df164a0fbe4c · inbound

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models cites this paper.

Pseudo-Unification: Entropy Probing Reveals Divergent Information Patterns in Unified Multimodal Models Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-05-11T08:50:59.145717Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T16:27:14.491492Z digest=sha256:7edb3987d646de3a63879a8b7a0d332e102fb210573a973fc542d110248dfcd9

Observation 61a52f06-dec5-4951-99fa-95a5ef2d98de · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:41:18.823854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-08T04:31:26.325118Z digest=sha256:40efee050226d532d8a94920522e65d71c0b520b12739c2e917b5d66a0f327fe

Observation 1683cb3c-e9e0-4994-9a8a-409e6b24e763 · inbound

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation cites this paper.

Tuna-2: Pixel Embeddings Beat Vision Encoders for Multimodal Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T23:43:51.078460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T23:41:25.275207Z digest=sha256:e372c4a25d798366e787afb20d6eb2b7cb048e7d1a4ff30bc45487e62d7cce04

Observation b91df5b1-f000-4f06-85b7-66cb885b08e3 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 26

Resolution
verified exact
arxiv_id, observed 2026-05-20T11:48:14.918323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-20T11:46:52.658984Z digest=sha256:04804aea96e1fddd95729ecb8b7f4389dd1e2132cbbf124be2e41a2070e14a0d

Observation d5172ce2-5713-4439-8ec6-0c928cda97c6 · inbound

Lance: Unified Multimodal Modeling by Multi-Task Synergy cites this paper.

Lance: Unified Multimodal Modeling by Multi-Task Synergy Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-21T07:59:50.752143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-21T07:56:34.034047Z digest=sha256:154f2f9ce61c2e6011e40ea0f56c3b92831b29407a2525ecc51b1e1fad7ebea7

Observation fc4b3c3a-b56b-4259-a0e3-9a079cd900d8 · inbound

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers cites this paper.

HYDRA-X: Native Unified Multimodal Models with Holistic Visual Tokenizers Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-07-03T14:28:32.117532Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-27T07:01:07.362430Z digest=sha256:64c43b3df1c18e5e9612411bb0f2d37112bdd001cdd2c76efd7259ebadeb9637

Observation cc2b5ce7-2ac5-4c70-b38c-bf713142dea9 · inbound

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation cites this paper.

Argus-Unified: Towards A Compact and Economical Unified Model for Image Understanding and Generation Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 104

Resolution
unresolved
no resolver link, observed 2026-08-01T02:14:09.861564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-01T02:14:09.861564Z digest=sha256:35f3a363aae97964ce39cf6d157cb35a708c728daa3ac49563067c9866772499

Observation 951e616b-0053-4d9b-a2ec-a51aa180cb9d · inbound

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications cites this paper.

Instruction-based Image Editing: A Survey on Data, Models, Evaluation, and Applications Unified Autoregressive Visual Generation and Understanding with Continuous Tokens

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-01T01:51:57.970306Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T01:51:57.970306Z digest=sha256:f381afc1be9a54cb88d1faf6ded9192884bb65b81fb20bd8ce487dd5bbc07104