Pith. sign in

Paper Citation Record · LEDGER

u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

As of 18 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2311.05348.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.05348 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T20:16:58.873253Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-18T15:36:34.016359Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d52ffb01-c86d-4286-9866-779efd545e76 · inbound

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation cites this paper.

MPG-SAM 2: Adapting SAM 2 with Mask Priors and Global Context for Referring Video Object Segmentation u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-10T15:47:59.877607Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T15:47:59.877607Z digest=sha256:8867a6e970c7b929b0b7814541028defcc6252553581ad319b8504b648b297ad

Observation bf6f09e9-e25a-4ad4-a635-392d2d4b58c6 · inbound

Ground-V: Teaching VLMs to Ground Complex Instructions in Pixels cites this paper.

Ground-V: Teaching VLMs to Ground Complex Instructions in Pixels u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-15T20:16:58.873253Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T20:16:58.873253Z digest=sha256:6e8346de62b391f2ecd702281b68cf1a134a3565ce66da5d16ed8ef01d5be0a2

Observation 3e0302b2-4c90-4d32-b153-02826c9f8c20 · inbound

CLGRPO: Reasoning Ability Enhancement for Small VLMs cites this paper.

CLGRPO: Reasoning Ability Enhancement for Small VLMs u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T23:27:46.102560Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:27:46.102560Z digest=sha256:492797a369d5d3cd499bf9d1e63330fa95e2f97ef060a156a94cf157dead0516

Observation 67f82c73-78c0-4c7c-a433-0ab6aff10414 · inbound

MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes cites this paper.

MINGLE: VLMs for Semantically Complex Region Detection in Urban Scenes u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-05-18T15:36:34.019765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-05-18T15:35:30.549656Z digest=sha256:a19a7b102f551c8cd23f84d91419da74c3f28771674cd480cc07b2418a8fa5eb

Observation 4eeb3ec6-d20a-4bff-a11d-4d7377ab8468 · inbound

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation cites this paper.

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-06T00:48:45.470485Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T00:48:45.470485Z digest=sha256:2b6a17be9e1d128754f28cbe070d0ef920f3ce3a0631bc58357a2e9f404458ef

Observation f43e1b53-f549-4a85-b80b-000d89afe81f · inbound

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation cites this paper.

CROSS: Cascaded Distillation and Dual-Constraint Grounding for Remote Sensing Referring Segmentation u-LLaVA: Unifying Multi-Modal Tasks via Large Language Model

Reference 65

Resolution
unresolved
no resolver link, observed 2026-08-08T00:49:55.158116Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T00:49:55.158116Z digest=sha256:60e6da9832b9534519123af16151aa57cc2019765014deaa544dcee2114a0844