Pith. sign in

Paper Citation Record · LEDGER

Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 8 inbound Pith citation observations for arXiv:2311.06783.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2311.06783 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 8 of 8 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 8 of 8 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:41:45.598200Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-22T19:52:01.984738Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 439d8daa-2fbe-4583-963d-724f143185e0 · inbound

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels cites this paper.

Q-Align: Teaching LMMs for Visual Scoring via Discrete Text-Defined Levels Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 288

Resolution
verified exact
arxiv_id, observed 2026-05-15T16:35:48.052998Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-15T16:35:47.826165Z digest=sha256:7b4e656bde4cec9cb9a0a41b3b36e65e60913c6506cd0152599ad8bd4af29fe5

Observation 62cddf9d-c106-438b-aaa0-d36e497ee0cc · inbound

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs cites this paper.

Cambrian-1: A Fully Open, Vision-Centric Exploration of Multimodal LLMs Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 136

Resolution
verified exact
arxiv_id, observed 2026-05-17T00:05:03.790735Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T00:05:03.547664Z digest=sha256:bee12fe9253324b2cd1d0758819deb7a6f64887d725400903f96d453003638bb

Observation 8a7f8e73-2e58-47a8-aaee-975629daa6a0 · inbound

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning cites this paper.

MetaMorph: Multimodal Understanding and Generation via Instruction Tuning Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 118

Resolution
metadata mismatch
arxiv_id, observed 2026-05-17T07:51:13.317986Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T07:51:12.953777Z digest=sha256:7253a3edad191125552e655f66fd0e796f3af00cace11e869032c06c08a018e0

Observation 802eabf2-ac88-4d93-b729-f46580100167 · inbound

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding cites this paper.

FLARE: Fully Integration of Vision-Language Representations for Deep Cross-Modal Understanding Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 68

Resolution
verified exact
arxiv_id, observed 2026-05-22T19:52:01.986923Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-22T19:49:00.961388Z digest=sha256:8c1bd5e3dbf481f94acbab44994dde326ba443ee0db552a08e19016b012b0cb5

Observation d46953ed-94ab-4b7f-9676-3bdd2604202f · inbound

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset cites this paper.

STORM: Benchmarking Visual Rating of MLLMs with a Comprehensive Ordinal Regression Dataset Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 59

Resolution
unresolved
no resolver link, observed 2026-08-07T11:41:45.598200Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:41:45.598200Z digest=sha256:5c170c2fa8176c6f8f08984508d6215ee93d678377521605cb9235ce2e70196c

Observation 2fd25818-8080-4507-bf87-18931a4ee56d · inbound

Grounding Degradations in Natural Language for All-In-One Video Restoration cites this paper.

Grounding Degradations in Natural Language for All-In-One Video Restoration Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-06T15:50:47.659414Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T15:50:47.659414Z digest=sha256:05a9c14c4f848f277f4bd830c6efe42781f564479809df25ac042bb7916c0eb2

Observation cd0eb886-c7c4-4d11-9518-3a0d4a945e7e · inbound

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA cites this paper.

Parameter-Efficient Adaptation of mPLUG-Owl2 via Pixel-Level Visual Prompts for NR-IQA Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T10:55:55.252815Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T10:55:55.252815Z digest=sha256:d969937bee2ce487c3dde10646b56c4bbc02fc839b0234c4fd02e73fd059fce6

Observation 93f23fb1-bc83-47b0-bb33-4330b1a11487 · inbound

SR-Ground: Image Quality Grounding for Super-Resolved Content cites this paper.

SR-Ground: Image Quality Grounding for Super-Resolved Content Q-Instruct: Improving Low-level Visual Abilities for Multi-modality Foundation Models

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-05-21T05:34:40.116069Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-21T05:34:17.056685Z digest=sha256:0f1529554fa7ef784540f5564628f135f15afb5d63a142a03ac27ef92ece31f9