Pith. sign in

Paper Citation Record · LEDGER

LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2306.06687.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.06687 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:34:10.895265Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T01:22:04.240381Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 45f1048f-03b8-4514-8ac8-f1e834c67083 · inbound

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension cites this paper.

SEED-Bench: Benchmarking Multimodal LLMs with Generative Comprehension LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-12T16:59:50.548101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-12T16:59:50.495335Z digest=sha256:82775f74c75aa85c53d3a9f650e88b77dc4e0b1deb27b8231376834b842a4c03

Observation 276392f6-c458-4c68-9c0f-f98b6c696e35 · inbound

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models cites this paper.

HallusionBench: An Advanced Diagnostic Suite for Entangled Language Hallucination and Visual Illusion in Large Vision-Language Models LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-05-17T01:22:04.243041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-17T01:22:04.035994Z digest=sha256:b8bae521082388b7b61f103a93ef556eeacaf4462eb677110a9566ec4d2e100f

Observation 5de93d22-69aa-48d6-b8e1-b44b4aba09db · inbound

MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI cites this paper.

MMMU: A Massive Multi-discipline Multimodal Understanding and Reasoning Benchmark for Expert AGI LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 84

Resolution
verified exact
arxiv_id, observed 2026-05-15T05:37:41.718632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-15T05:37:41.401736Z digest=sha256:d714b16238cb683a9fe5c3591d93214733584c6ff6f1b562fb6e73f64571cf56

Observation eb2dfaed-8279-4d42-9cac-c2fa98baea49 · inbound

Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation cites this paper.

Pisces: An Auto-regressive Foundation Model for Image Understanding and Generation LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 78

Resolution
unresolved
no resolver link, observed 2026-08-07T04:34:10.895265Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T04:34:10.895265Z digest=sha256:d92552d6b1d65de01fb100777133192b0fb4ee67a9b8a84189883555bc12210a

Observation f852d706-4f23-4ec5-afb2-54e1404b1277 · inbound

Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models cites this paper.

Argus: Leveraging Multiview Images for Improved 3-D Scene Understanding With Large Language Models LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T16:42:13.260546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:42:13.260546Z digest=sha256:8f8029ae312567530a74dd7a685a586fda7b12f4127ffc829db5facfc414e5f8

Observation ca4deca7-28f2-4a59-ab18-b6c3b99817f3 · inbound

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning cites this paper.

AeroRAG: Structured Multimodal Retrieval-Augmented LLM for Fine-Grained Aerial Visual Reasoning LAMM: Language-Assisted Multi-Modal Instruction-Tuning Dataset, Framework, and Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-05-10T11:40:19.799648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T04:46:22.891034Z digest=sha256:d397939da0923cfc6825ae8c8e9d2e0030f623c9af8cd8cf9c19dee9a9ca4fe9