Pith. sign in

Paper Citation Record · LEDGER

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

As of 9 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 2 inbound Pith citation observations for arXiv:2507.02859.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.02859 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T20:24:57.154109Z

measured 41 of 41 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-07-14T10:06:52.822171Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T11:46:55.443273Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact0
  • verified fuzzy17
  • unresolved20
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation dcf504e3-25e2-4fb6-b3f4-b1afc63c03a0 · outbound

This paper cites Lion: Empowering multimodal large language model with dual-level visual knowledge.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Lion: Empowering multimodal large language model with dual-level visual knowledge

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.982611Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.982611Z digest=sha256:3c7887876f704cf21547f41c264b83096c0ba58c737867b19d2b69208f7f27e9

Observation b9b49cf4-48b3-4927-bedb-62f79bc4c293 · outbound

This paper cites MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation MiniGPT-v2: large language model as a unified interface for vision-language multi-task learning

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.987451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.987451Z digest=sha256:67a77f88ed8e6e657d7e939ee71ecdc0bc2704f4b910fe2590046ab56f2045fb

Observation fddf9d81-e69c-41bc-8704-d7293b380278 · outbound

This paper cites Training Verifiers to Solve Math Word Problems.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Training Verifiers to Solve Math Word Problems

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.992240Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.992240Z digest=sha256:7d42c3dc877712120045b27c771a9182e41a8111ef4b2d0b54b94c46f07b3af9

Observation fde16b29-bd6f-49a1-9fc3-c35633c2d7c0 · outbound

This paper cites The Llama 3 Herd of Models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation The Llama 3 Herd of Models

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:56.997274Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:56.997274Z digest=sha256:83544259810d18ff3063c1d1fd682d1d68e4db6db92b03c4c2274720b9ffc63f

Observation 337842db-ae19-48bd-b2cc-4701c88c8015 · outbound

This paper cites ChartLlama: A Multimodal LLM for Chart Understanding and Generation.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartLlama: A Multimodal LLM for Chart Understanding and Generation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.001879Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.001879Z digest=sha256:7a058ff732fe511ff456a4b7fb43a6718e2592059dfcc27e7e2623cbd95fc6b9

Observation c84e23e9-1b37-4075-a4e2-ba347837b1f8 · outbound

This paper cites Lora: Low-rank adaptation of large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Lora: Low-rank adaptation of large language models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.717781Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.006527Z digest=sha256:2044f4f4d401b681917208f9dc85d2ae8b3f1aa5e425fde58aa2eb9e2a73ae47

Observation 89640aa1-16c9-4a21-981c-4d7c069ee456 · outbound

This paper cites Icdar2019 compe- tition on scanned receipt ocr and information extraction.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Icdar2019 compe- tition on scanned receipt ocr and information extraction

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.010697Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.010697Z digest=sha256:7f34a2637f7c551d578e76881d1f389f4e2133924ed97da196285994c0448b6d

Observation 41ccbb43-9ec3-4e91-bcb1-ccd0d4e2fdf1 · outbound

This paper cites Dvqa: Understanding data visualizations via ques- tion answering.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Dvqa: Understanding data visualizations via ques- tion answering

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.014859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.014859Z digest=sha256:3af78cad6c602671e52a3815b70f3cfdbdc6278df9d4c97f66c2e7725b1ad478

Observation 75e04989-60eb-42bd-990c-34ef9714ebcb · outbound

This paper cites Large language models are zero-shot reasoners.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Large language models are zero-shot reasoners

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.686967Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.019111Z digest=sha256:a7ff6aa33390c243eb57bd74d762b2ef760f5ddac115a385e5105de9e26aa76d

Observation 9b952664-fbd5-4b9f-b041-c23bd4d6d716 · outbound

This paper cites Visual genome: Connecting language and vision using crowdsourced dense image annotations.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual genome: Connecting language and vision using crowdsourced dense image annotations

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.024113Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.024113Z digest=sha256:34f3834c3494d3175f66e89b68145a4386869a4bc77a8808310abf55e800eb19

Observation 70ba3b26-17d3-41d2-8670-94555f69bacc · outbound

This paper cites Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Blip-2: Bootstrapping language-image pre-training with frozen image encoders and large language models

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.028725Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.028725Z digest=sha256:08d938b2ff4257f5dc1ed9fc31a11544701c754959c796c8d100814a67b9e948

Observation e2f59f45-9306-4a10-ae4f-c6dca2908af5 · outbound

This paper cites Deductive verification of chain-of-thought reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Deductive verification of chain-of-thought reasoning

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.656097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.033101Z digest=sha256:7176e67af849d6f0e68dd55b76527e51c4091cdd4597150789a96f44598edcaf

Observation eeefdf7a-eee3-4929-9f3c-0eb9cc76fd04 · outbound

This paper cites Improved baselines with visual instruction tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Improved baselines with visual instruction tuning

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.638991Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.037955Z digest=sha256:004ef90fcf7f83fd51971dc8b37161b1deae6dfd7f444f6088808257f7d175d3

Observation f691e5eb-9093-4fec-abe8-181e7349ec9e · outbound

This paper cites Visual instruction tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual instruction tuning

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.042273Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.042273Z digest=sha256:06ea1d4a50a40f976e8d0ac2defbc1493fc95c570eeed0e6f41259a73769516c

Observation 8d73a0a5-26ab-4176-a6f0-1edefb03280d · outbound

This paper cites Nltk: The natural language toolkit.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Nltk: The natural language toolkit

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.617121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.046456Z digest=sha256:faa73a66aec2cbea43fc1bf70010012335a96cb786ff7b72556ed51726dc4f9c

Observation 730f94d4-a500-453d-b600-77ff873e777c · outbound

This paper cites Decoupled weight decay regularization, 2019.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Decoupled weight decay regularization, 2019

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.051035Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.051035Z digest=sha256:b44be9802297c7a435a68c7ee232606e35716e7bc5d3142daeaad96843fd1400

Observation f67e04b2-c603-40db-b79a-40384cc05c33 · outbound

This paper cites Dynamic prompt learning via policy gradient for semi-structured mathematical reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Dynamic prompt learning via policy gradient for semi-structured mathematical reasoning

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.594983Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.055458Z digest=sha256:8c1aae1a7a4fc1dd4e31431bc715f96be7975f73d796e50ec99ec2aea4c25770

Observation 295cb5a7-f4bc-4a59-97ab-171cb26625f2 · outbound

This paper cites Chartqa: A benchmark for question an- swering about charts with visual and logical reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Chartqa: A benchmark for question an- swering about charts with visual and logical reasoning

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.580807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.060651Z digest=sha256:cef616e10cab42b97d69bd18998c1e9a75c0ff3743122eb6675cb1c58ea941b5

Observation f3f7c609-a6df-4643-a018-ee3e45b957f7 · outbound

This paper cites ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartGemma: Visual Instruction-tuning for Chart Reasoning in the Wild

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.064925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.064925Z digest=sha256:0c9f90a90a6ee99d9f7610e4d417ad103cd97ed4c5216ee316e894556c786345

Observation 196330d8-89f1-494f-bb63-b6a766836d71 · outbound

This paper cites ChartAssisstant: A Universal Chart Multimodal Language Model via Chart-to-Table Pre-training and Multitask Instruction Tuning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation ChartAssisstant: A Universal Chart Multimodal Language Model via Chart-to-Table Pre-training and Multitask Instruction Tuning

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.069320Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.069320Z digest=sha256:ffaac00da3059d90c1c13242cda7b7e80e57c5e3d495571a993a1ad247ad46ff

Observation 5684f8a7-b8f4-4511-a2fa-e21875e76203 · outbound

This paper cites Selfcheck: Using llms to zero-shot check their own step-by-step reason- ing.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Selfcheck: Using llms to zero-shot check their own step-by-step reason- ing

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.567212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.073735Z digest=sha256:85f2037eb24d98fa9f89ef6c1b17744c1e8ab8ca0441f0f41ace3dff8b7ab0b4

Observation 7c4cc44b-90d6-4fef-946c-a6aa0c2d07da · outbound

This paper cites Im2text: Describing images using 1 million captioned pho- tographs.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Im2text: Describing images using 1 million captioned pho- tographs

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.552516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.078227Z digest=sha256:5436d58758e91d7f6fef4b26478c5ca38dd7269a476663ffc5ef2dee6434cc12

Observation 865f1f13-391c-44c2-8d01-0426461d180b · outbound

This paper cites Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Flickr30k entities: Collecting region-to-phrase corre- spondences for richer image-to-sentence models

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.082981Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.082981Z digest=sha256:a99f337c4dc08fec59c50d18ed4aad925876b453238ae0f1804972c8128bfcaf

Observation a3bd51e3-7e11-445c-b490-77522bcb375e · outbound

This paper cites Aligning large and small language models via chain-of-thought reasoning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Aligning large and small language models via chain-of-thought reasoning

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.530060Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.087259Z digest=sha256:fc4db9648e46185fbd8d18bb1c8459e91c2d79efe4c0046901bbe419191e5bfa

Observation de6b8527-f939-4f29-97c2-c8f351b2d215 · outbound

This paper cites Laion-400m: Open dataset of clip-filtered 400 million image-text pairs.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Laion-400m: Open dataset of clip-filtered 400 million image-text pairs

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.516556Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.092576Z digest=sha256:cbead2f049cfb53994f9274415d00a269d4b760f385b8bd3fd7a60d269c2d842

Observation 6127810b-2b5a-498c-bfd3-905d4efbf858 · outbound

This paper cites Visual cot: Unleashing chain-of-thought reasoning in multi-modal language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual cot: Unleashing chain-of-thought reasoning in multi-modal language models

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.503066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.096740Z digest=sha256:1d957d74b3d3ccc113efb9c6e65e9ea3e6aadf8caa51463a2d39a3b6f04f9a82

Observation 3bcdec76-d803-4e29-ad25-d2a84edfcd44 · outbound

This paper cites Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Conceptual captions: A cleaned, hypernymed, im- age alt-text dataset for automatic image captioning

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.488546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.100783Z digest=sha256:737d847e8a9d08ebd430668ac5f6c646b04e9d401b134f6768db904cc4649260

Observation 181ad49c-580b-4edb-9735-fbf457ff94d6 · outbound

This paper cites Mome: Mixture of multimodal experts for generalist multimodal large language models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Mome: Mixture of multimodal experts for generalist multimodal large language models

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.105148Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.105148Z digest=sha256:6518128e58b79044a70d62d1509008420c84f2e09da57aad5430f3db2e1e83f6

Observation 3e38e282-e2d4-488f-aaf8-06f6ad2db1ce · outbound

This paper cites Chain-of-thought prompting elicits reasoning in large lan- guage models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Chain-of-thought prompting elicits reasoning in large lan- guage models

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.109588Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.109588Z digest=sha256:da6f8b33f3145f817e4b70131712c89b326cc44bb06850573f9038b0aad71aed

Observation 33aafc2c-50ed-4f2a-8985-4d80ee31dcc9 · outbound

This paper cites Grounded Chain-of-Thought for Multimodal Large Language Models.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Grounded Chain-of-Thought for Multimodal Large Language Models

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.113550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.113550Z digest=sha256:ca7012e5126bbecb97b5842a6bb59bd09a38dd0cc1d32f1b6649e51189314d3b

Observation 4be37be1-5449-44a0-a5b5-3afe5f5a0f82 · outbound

This paper cites Visionary-r1: Mitigating shortcuts in vi- sual reasoning with reinforcement learning.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visionary-r1: Mitigating shortcuts in vi- sual reasoning with reinforcement learning

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T20:24:57.117687Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:24:57.117687Z digest=sha256:061ffbdd676ccb509ede8e90bfb405e5bbb59e574235381f1de0816218bc16dc

Observation 723aeddf-b3be-4a50-8615-2acb65734afc · outbound

This paper cites Falcon: Resolv- ing visual redundancy and fragmentation in high-resolution multimodal large language models via visual registers.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Falcon: Resolv- ing visual redundancy and fragmentation in high-resolution multimodal large language models via visual registers

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.456072Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.121824Z digest=sha256:b3dc0dc71dea357635e07bba9bd55a692bdf27a32a850e512401390ee7ba7135

Observation 8b606a23-4068-4979-ab81-d52df5b795f6 · outbound

This paper cites Tat-qa: A question answering benchmark on a hybrid of tab- ular and textual content in finance.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Tat-qa: A question answering benchmark on a hybrid of tab- ular and textual content in finance

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.442042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.126531Z digest=sha256:7e47d88b3d2d2fbc87d10863e85828248be82189aa1c6c4e6505225d0f501579

Observation d1ea09ca-e4d5-477a-a6c8-8c56b48f928e · outbound

This paper cites Visual7w: Grounded question answering in images.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Visual7w: Grounded question answering in images

Reference 34

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:24:57.428003Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.130597Z digest=sha256:fd03a2d18032e8fff775c463b0145e45f9fff89207cfddcd49a1c5bb1bc401c6

Observation 1be332be-6cf8-4e31-9040-ebfaa80a41a8 · outbound

This paper cites an unresolved cited work.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:57.398146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.140629Z digest=sha256:869626346b4398dbec61f6da814dfee6b0a967b7e78db40432877f5c328575e4

Observation 83a36de8-3632-41b7-93dc-e3e9ff4fc514 · outbound

This paper cites an unresolved cited work.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T20:24:57.383742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.145279Z digest=sha256:91fd5b20ae55326bc26b632434a18ddac21f912eff4fea8339514564f2850bb3

Observation dc3c4422-2004-4480-b7f5-4459846cb8a9 · outbound

This paper cites The table shows the number of fan letters for each day.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation The table shows the number of fan letters for each day

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.368364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.149476Z digest=sha256:fbbad720adcb34032b0d1d08100b6a8024853b34ea85b04c058e05d4bfa0c8ab

Observation c92ee23a-db34-49bb-b932-050b9e5fa0dd · outbound

This paper cites Opening Balance.\.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Opening Balance.\

Reference 39

Resolution
malformed identifier
raw_fallback, observed 2026-08-06T20:24:57.352221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.154109Z digest=sha256:1d97c642f399d9d92adda1d4c90c173e08621aba51b26fb39442ab01dd0d9b74

Observation bb02a9a4-8ac3-4115-bb41-4aa687633838 · outbound

This paper cites Thus, the total number of fan letters received on Thursday and Monday is: 204 + 271 = 475.

Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation Thus, the total number of fan letters received on Thursday and Monday is: 204 + 271 = 475

Reference 204

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T20:24:57.413154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T20:24:57.135943Z digest=sha256:5d41e4655d73ecec174c0f8d97ebaf55ed79d07930f8645e93f0a155b6a72e56

Pith citing papers

Observation 71c9b1fc-2c69-43b6-8b4d-c87e0865845c · inbound

Balancing Image Compression and Generation with Bootstrapped Tokenization cites this paper.

Balancing Image Compression and Generation with Bootstrapped Tokenization Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

Reference 35

Resolution
verified exact
arxiv_id, observed 2026-07-02T11:46:55.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-06-28T03:07:33.054518Z digest=sha256:fcd6bad923553f8e984591aeb2c31154ed9aee091906dc40a61d7502a0438f7a

Observation e9852caa-6d97-453e-b997-8bb10ba5d27c · inbound

Answer-Conditioned Chain-of-Thought Distillation for Few-Shot Industrial Vision with Small VLMs cites this paper.

Answer-Conditioned Chain-of-Thought Distillation for Few-Shot Industrial Vision with Small VLMs Bootstrapping Grounded Chain-of-Thought in Multimodal LLMs for Data-Efficient Model Adaptation

Reference 16

Resolution
unresolved
no resolver link, observed 2026-07-14T10:06:52.822171Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T10:06:52.822171Z digest=sha256:778cc8ef8dfb9417998c835a86783ce4368a103057963e5ad0beb066770add6a