Pith. sign in

Paper Citation Record · LEDGER

OneLLM: One Framework to Align All Modalities with Language

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:2312.03700.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.03700 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:25:52.851020Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-06-29T07:43:13.418796Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 800823cb-a109-4a0f-87e9-f8838e1f2a34 · inbound

Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation cites this paper.

Distill CLIP (DCLIP): Enhancing Image-Text Retrieval via Cross-Modal Transformer Distillation OneLLM: One Framework to Align All Modalities with Language

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T14:25:52.851020Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:25:52.851020Z digest=sha256:b8349789d346301a5a4082eaffd47d529bec84d57fdb55457c486c1881761e9b

Observation 4925e626-4ae7-49a8-b780-4c964a3ff5d8 · inbound

Abstractive Visual Understanding of Multi-modal Structured Knowledge: A New Perspective for MLLM Evaluation cites this paper.

Abstractive Visual Understanding of Multi-modal Structured Knowledge: A New Perspective for MLLM Evaluation OneLLM: One Framework to Align All Modalities with Language

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:51:15.373641Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:51:15.373641Z digest=sha256:529d593738c905cd718e9dc9a455bf9e1e9eb81dde9604ec295a8d05e4368595

Observation 3a331ea2-4e45-46c3-b90b-5d7b8e65f565 · inbound

Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes cites this paper.

Towards Open-Vocabulary Multimodal 3D Object Detection with Attributes OneLLM: One Framework to Align All Modalities with Language

Reference 2024

Resolution
unresolved
no resolver link, observed 2026-08-05T17:13:02.378508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:13:02.378508Z digest=sha256:fd01ccf1e5b95c348abb9e1c7a640db6656da784939413a6958a4da970c5408e

Observation 3aa9f971-1848-4586-9727-d1bb3d7790a3 · inbound

AffectVerse: Emotional World Models for Multimodal Affective Computing cites this paper.

AffectVerse: Emotional World Models for Multimodal Affective Computing OneLLM: One Framework to Align All Modalities with Language

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-20T06:48:05.742234Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-20T06:46:33.612905Z digest=sha256:3def9f7a3cabea08d13f8662bd14e7b2d544c1055b1b27a8dacf530a1e3fa9e3

Observation 55bd66df-eade-4c41-b4fc-22aa053c87a4 · inbound

Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion cites this paper.

Mind-Omni: A Unified Multi-Task Framework for Brain-Vision-Language Modeling via Discrete Diffusion OneLLM: One Framework to Align All Modalities with Language

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-06-29T07:43:13.420351Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T07:41:53.678423Z digest=sha256:203f49fed9b47171fe9574039d657e035941c1cd3635e54c7ce60713fa03597f