Pith. sign in

Paper Citation Record · LEDGER

MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2212.10773.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2212.10773 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T22:43:59.481781Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-24T09:14:16.463173Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation fde93e06-3f15-492a-85ab-da365e81ff90 · inbound

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning cites this paper.

The Flan Collection: Designing Data and Methods for Effective Instruction Tuning MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 63

Resolution
verified exact
arxiv_id, observed 2026-05-24T09:14:16.466194Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-24T09:13:30.054153Z digest=sha256:c79f6ab1f3e206157048f6d5a86b9a61156d5a99248d83da399c9be2a5f2bf12

Observation 629f4b41-8689-4908-939a-3d0eb2628467 · inbound

Otter: A Multi-Modal Model with In-Context Instruction Tuning cites this paper.

Otter: A Multi-Modal Model with In-Context Instruction Tuning MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 91

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T02:43:47.887076Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T02:43:47.775691Z digest=sha256:141d379f2f3955c0a8850c51503c3754df6ae9d08e39286fd26e8fb715ac5ebb

Observation e63c9893-e889-401c-b263-54d1317a6dfc · inbound

InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning cites this paper.

InstructBLIP: Towards General-purpose Vision-Language Models with Instruction Tuning MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 48

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T02:13:52.162310Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-13T02:13:52.097263Z digest=sha256:95a499fbee42fcbb92b64487c52e7b68c4ba8199db126c37b2a3e4ac6bc0a842

Observation f346baeb-1a5e-4080-89a4-ace0b5551339 · inbound

MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models cites this paper.

MME: A Comprehensive Evaluation Benchmark for Multimodal Large Language Models MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 51

Resolution
metadata mismatch
arxiv_id, observed 2026-05-10T20:25:34.250884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T20:25:33.854923Z digest=sha256:ebd5dbd34397c7ee60101a297ced3436aff6c33772479c30260ae19d743b3c24

Observation 50137cd8-ac54-4bdc-a85e-125f94bb35cd · inbound

A Survey on Multimodal Large Language Models cites this paper.

A Survey on Multimodal Large Language Models MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 104

Resolution
metadata mismatch
arxiv_id, observed 2026-05-16T02:56:42.675095Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-16T02:56:41.658658Z digest=sha256:04c3b7dfd858dc6ee5ff2689728cbad0f1a64d75d2cd08b6432dbd92c21f1e18

Observation 1f6c1be1-0112-4b96-998d-b0f793db6cd3 · inbound

A Comprehensive Overview of Large Language Models cites this paper.

A Comprehensive Overview of Large Language Models MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 280

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:28:39.244930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-19T20:28:38.900026Z digest=sha256:166e82bdfe1dc721b5fa16ce0a4cc76be3dbaac6c24fd5fd9902c5ddc64d0ca6

Observation 72e38b62-b111-41b5-ae2a-b7cc87228188 · inbound

From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs cites this paper.

From Visuals to Vocabulary: Establishing Equivalence Between Image and Text Token Through Autoregressive Pre-training in MLLMs MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T22:43:59.481781Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T22:43:59.481781Z digest=sha256:9e1dfa68b81473fc16ec2d83eda24672c63601ea6378a090df8e4faec4a74677

Observation 7b2e6457-62fd-42fb-a0e8-3d9c75c8c7c7 · inbound

Instructify: Demystifying Metadata to Visual Instruction Tuning Data Conversion cites this paper.

Instructify: Demystifying Metadata to Visual Instruction Tuning Data Conversion MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-07T14:37:53.282556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:37:53.282556Z digest=sha256:0b17672f47d692e0ccd8777ff84fee73b3d9c1a2a8db2a0c62ba4e3fa28ee6ce

Observation 5185a701-f321-4b12-839f-2bc92edb1fb1 · inbound

Can LLMs Understand Unvoiced Speech? Exploring EMG-to-Text Conversion with LLMs cites this paper.

Can LLMs Understand Unvoiced Speech? Exploring EMG-to-Text Conversion with LLMs MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-07T12:11:52.785933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:11:52.785933Z digest=sha256:a4f5df711226661b1764850b84d14b00424dfe49014556da97087cc42a9e9495

Observation f9340508-6582-4b6f-8286-651ed35a8bdc · inbound

CARE-X: Towards Clinically Useful Radiology VLMs with Auxiliary Supervision, Reward-Aligned Learning, and Tool-Augmented Measurement cites this paper.

CARE-X: Towards Clinically Useful Radiology VLMs with Auxiliary Supervision, Reward-Aligned Learning, and Tool-Augmented Measurement MultiInstruct: Improving Multi-Modal Zero-Shot Learning via Instruction Tuning

Reference 125

Resolution
unresolved
no resolver link, observed 2026-08-05T10:23:59.148488Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T10:23:59.148488Z digest=sha256:696c8fe98d4211f01e4b3e622b666c6ebb6c39b295d1279623eee09c85ffd886