Pith. sign in

Paper Citation Record · LEDGER

Auto-Encoding Morph-Tokens for Multimodal LLM

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2405.01926.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2405.01926 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T20:21:30.509148Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T02:10:56.069881Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e58a19f5-ef6e-4519-8d89-47b155f1b43e · inbound

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation cites this paper.

HealthGPT: A Medical Large Vision-Language Model for Unifying Comprehension and Generation via Heterogeneous Knowledge Adaptation Auto-Encoding Morph-Tokens for Multimodal LLM

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-07T20:21:30.509148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T20:21:30.509148Z digest=sha256:21a0d1482807fc9fe52814d9032018b31f077321eb48564e0d191fbee4b28184

Observation 72fcbe35-186c-472a-8f04-0556d9ffd91a · inbound

Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM cites this paper.

Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM Auto-Encoding Morph-Tokens for Multimodal LLM

Reference 52

Resolution
verified exact
arxiv_id, observed 2026-05-22T02:10:56.072744Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-22T02:06:35.204166Z digest=sha256:14ef128c7220c15fb028b02b6090ebcb6782d4920a2fd467513f1ab1d35ad9ac

Observation dca5abc9-c343-408a-8b5e-7b94d32c7641 · inbound

FocusDiff: Advancing Fine-Grained Text-Image Alignment for Autoregressive Visual Generation through RL cites this paper.

FocusDiff: Advancing Fine-Grained Text-Image Alignment for Autoregressive Visual Generation through RL Auto-Encoding Morph-Tokens for Multimodal LLM

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-07T10:23:12.551112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T10:23:12.551112Z digest=sha256:236f383cfeb06313b807ad0f1b7f520a32ac4d1efcb64fabad197965bc566e2e

Observation 17e9e6e2-2b33-4e2e-8b5b-99c661c349f5 · inbound

What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities cites this paper.

What Limits Virtual Agent Application? OmniBench: A Scalable Multi-Dimensional Benchmark for Essential Virtual Agent Capabilities Auto-Encoding Morph-Tokens for Multimodal LLM

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T05:02:29.246334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:02:29.246334Z digest=sha256:008ea562b425d1851a16f146ae0a8a93d0da73c1346b66e68bcd0abd4e8dd5b8

Observation d6e0347b-9722-4670-9626-73ebca9cd5b2 · inbound

Towards Meta-Cognitive Knowledge Editing for Multimodal LLMs cites this paper.

Towards Meta-Cognitive Knowledge Editing for Multimodal LLMs Auto-Encoding Morph-Tokens for Multimodal LLM

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-05T05:12:44.292798Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T05:12:44.292798Z digest=sha256:ea128b18fd6e20ff13dbd1763c60b4473f64398bb94fabc2f0a98e2834db8553