Pith. sign in

Paper Citation Record · LEDGER

PixelLM: Pixel Reasoning with Large Multimodal Model

As of 6 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 4 inbound Pith citation observations for arXiv:2312.02228.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.02228 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 4 of 4 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T15:25:03.107930Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T10:48:03.024533Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5cf9db23-a1b3-4486-9bd3-6d6313aaeca9 · inbound

Advancing Visual Large Language Model for Multi-granular Versatile Perception cites this paper.

Advancing Visual Large Language Model for Multi-granular Versatile Perception PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 58

Resolution
unresolved
no resolver link, observed 2026-08-06T15:25:03.107930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:25:03.107930Z digest=sha256:f2320c27363e3b7a816f2e572ceedb57477ca39975bece088dc39a64d30af234

Observation 1a35217e-96be-4a4a-bc27-61c00f6786e8 · inbound

MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes cites this paper.

MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:36:34.027052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-18T15:35:30.549656Z digest=sha256:ece3d26719ad1222d436f00fd75be5b7b7f8d57163dc72902d442a4f1496fec6

Observation cd626809-10df-47f8-9bf5-15e7b2ba0485 · inbound

Vision Harnessing Agent for Open Ad-hoc Segmentation cites this paper.

Vision Harnessing Agent for Open Ad-hoc Segmentation PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 90

Resolution
verified exact
arxiv_id, observed 2026-05-20T05:53:04.475649Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-20T05:52:40.429412Z digest=sha256:78905e84b5dfcb48042a00c77d686dffd0017a790f392ebf7c7f62e014dd32ed

Observation 592de972-91a7-4d8e-b634-46c85a1f6496 · inbound

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning cites this paper.

InternVideo3: Agentify Foundation Models with Multimodal Contextual Reasoning PixelLM: Pixel Reasoning with Large Multimodal Model

Reference 159

Resolution
metadata mismatch
arxiv_id, observed 2026-07-03T10:48:03.025941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-06-27T09:48:27.652901Z digest=sha256:2e1e939c84ce882397c61f09c075835a954f68302c6b41011e89edfdd9811c76