Pith. sign in

Paper Citation Record · LEDGER

LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2306.06687.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06687 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:34:10.895265Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T01:22:04.240381Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45f1048f-03b8-4514-8ac8-f1e834c67083 · inbound

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension cites this paper.

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:59:50.548101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-12T16:59:50.495335Z digest=sha256:73815b5c19f89f58e36794f929bf9b8395b3cc7adeebb1d06bf99168d70abf94

Observation 276392f6-c458-4c68-9c0f-f98b6c696e35 · inbound

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models cites this paper.

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:22:04.243041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-17T01:22:04.035994Z digest=sha256:1ac0010760ff1f6c13b880e9b15e7bb7bde8cbbabf2736acd41dc13f9a3525a2

Observation 5de93d22-69aa-48d6-b8e1-b44b4aba09db · inbound

MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI cites this paper.

MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:37:41.718632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-15T05:37:41.401736Z digest=sha256:568ef4ffffeccde0dcb675c777e495020eb7025b92c59cf4d7765dfb6fe0db13

Observation eb2dfaed-8279-4d42-9cac-c2fa98baea49 · inbound

Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation cites this paper.

Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T04:34:10.895265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:34:10.895265Z digest=sha256:d92552d6b1d65de01fb100777133192b0fb4ee67a9b8a84189883555bc12210a

Observation f852d706-4f23-4ec5-afb2-54e1404b1277 · inbound

Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models cites this paper.

Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:13.260546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:13.260546Z digest=sha256:db96b6ba0a2f174eeedccd9431d6108d04ac697b98f1b03ba06280b7861fb74b

Observation ca4deca7-28f2-4a59-ab18-b6c3b99817f3 · inbound

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning cites this paper.

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:40:19.799648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-10T04:46:22.891034Z digest=sha256:e912ee68dc28b3aaba4cba5e471c5844a496d0b2b24e577cfca0277d4cf2b6bb