Pith. sign in

Paper Citation Record · LEDGER

Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language Model

As of 7 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 3 inbound Pith citation observations for arXiv:2312.12423.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2312.12423 v2

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 3 of 3 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 3 of 3 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:49:53.561200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T18:44:54.549994Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f46c4f5b-6f22-451a-b49f-263f50a3fd1c · inbound

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks cites this paper.

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language Model

Reference 110

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:53.561200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:53.561200Z digest=sha256:c80e38b3a6f88bcaa24e9f26b3d50f6117aa71a1b0b0e52761cbcc9b60d04696

Observation 6226041d-62b7-4d13-82ff-75218f735c81 · inbound

EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception cites this paper.

EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language Model

Reference 87

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:11.061406Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:41:11.061406Z digest=sha256:e5d31bb5a6a51c5e6b8e3bb2d504304a99a7e13df1c5fe74b7ac2a0e22d3cbba

Observation aa5a5238-fd1f-413c-a830-951424b64e50 · inbound

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO cites this paper.

Teaching MLLMs to Say No: Generalized Referring Expression Comprehension via Refusal Calibrated GRPO Jack of All Tasks, Master of Many: Designing General-purpose Coarse-to-Fine Vision-Language Model

Reference 34

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T18:44:54.593433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=arxiv_source observed=2026-08-06T18:44:46.348688Z digest=sha256:b59cbd5dfb1c1d6bc07f9d78791eaf104dc303910ec38fcab236dbc2c1307f18