Pith. sign in

Paper Citation Record · LEDGER

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

As of 11 August 2026, this Paper Citation Record lists 22 of 22 outbound references and 1 inbound Pith citation observation for arXiv:2506.08277.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.08277 v3

Coverage vector

measured 22 of 22 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-22T00:02:41.893373Z

measured 23 of 23 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-29T12:53:45.827516Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-06-29T13:03:26.839969Z

Reference resolution

22 of 22 outbound references displayed

  • verified exact7
  • verified fuzzy11
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch2

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9247353d-2d6d-46c8-a949-2b8c31632d98 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.054934Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:d4e425d98712bfd08ba4463f510ff0c32c1c7612a716b0dd0eb1764c3fb3ee83

Observation e4d9f9e8-464f-419b-9dff-4098bbb04d87 · outbound

This paper cites Mae-ast: Masked autoencoding audio spectrogram transformer.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Mae-ast: Masked autoencoding audio spectrogram transformer

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.287209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:e921f51762ac153c20d9c0620b6f325fd872c3c4e82c51ecb5bb0dbb08944894

Observation 9e0954fe-bac3-4ba0-b162-7f8603e32b93 · outbound

This paper cites Qwen2-Audio Technical Report.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Qwen2-Audio Technical Report

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-05-22T00:04:27.052288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:d89163f3e5784fdd53514bdc1716ae486a341473bc16a1f5a64835165de9830f

Observation efdf35dd-cb56-4ba0-8021-379b9f1dcf0e · outbound

This paper cites What can 1.8 billion regressions tell us about the pressures shaping high-level visual representation in brains and machines? bioRxiv, pp.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli What can 1.8 billion regressions tell us about the pressures shaping high-level visual representation in brains and machines? bioRxiv, pp

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.285189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:0ca8be7b2e6a2439b36e33f92d0c73d9b319414f52d84c6be99e48d63b64af61

Observation 9715bfa2-67b0-41ad-b544-735da853fc9f · outbound

This paper cites Visual representations in the human brain are aligned with large language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Visual representations in the human brain are aligned with large language models

Reference 5

Resolution
metadata mismatch
arxiv_id, observed 2026-05-22T00:04:27.048641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:055652e20c6150203c9a30d49e682ff3e6fdd1f8fdfa0ce03d361401038102f6

Observation 76930e47-ee7f-4ed6-b513-d193e6c655cc · outbound

This paper cites Vision-Language Integration in Multimodal Video Transformers (Partially) Aligns with the Brain.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Vision-Language Integration in Multimodal Video Transformers (Partially) Aligns with the Brain

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:04:27.047326Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:4fb309fe8db8738e1b2fb9782b60c0db0a6a24193226479d29ee652e8f071ec7

Observation 0f735a8a-5188-41ae-932e-3d575835f6fc · outbound

This paper cites Mistral 7B.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Mistral 7B

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.059751Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T22:08:12.954417+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:dae328a8015dca704dced5ca4f6e13e9ba0cb0e85c60f7ad08894aeba19472d1

Observation 37320b77-74e9-4433-921e-1f4164d9d5ac · outbound

This paper cites VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli VideoChat-R1: Enhancing Spatio-Temporal Perception via Reinforcement Fine-Tuning

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.053881Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:c0cacf404dd5d16cc35f53c15a41f8ca48b6ca48aac0b2ca54db98605a318dcc

Observation 3b23392d-797b-4374-89ff-f3631fdbd9e1 · outbound

This paper cites Video-llava: Learning united visual representation by alignment before projection.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Video-llava: Learning united visual representation by alignment before projection

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.274146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:1d8a85cb86dd88853bf09983170bd48db920876632e2f8531c97e3dae7303ad9

Observation 625aa75e-907d-4bf0-8787-59144ae213a6 · outbound

This paper cites Video-chatgpt: Towards detailed video understanding via large vision and language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Video-chatgpt: Towards detailed video understanding via large vision and language models

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.279461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:31240c91764e2b434ef798a89c2e1745c30c5bb84af0718596baf74a38952e81

Observation 7b3b12e9-57e4-4798-bb9d-5fb156e41aa9 · outbound

This paper cites Yuko Nakagi, Takuya Matsuyama, Naoko Koide-Majima, Hiroto Yamaguchi, Rieko Kubo, Shinji Nishimoto, and Yu Takagi.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Yuko Nakagi, Takuya Matsuyama, Naoko Koide-Majima, Hiroto Yamaguchi, Rieko Kubo, Shinji Nishimoto, and Yu Takagi

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.275120Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:79c0ae2bc308911e19c8fcffc02768137982b1159adcd760750339ac7d392fdd

Observation c8c34f79-e349-404d-9333-6913a3b9e154 · outbound

This paper cites The cost of compression: Investigating the impact of compression on parametric knowledge in language models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The cost of compression: Investigating the impact of compression on parametric knowledge in language models

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.283002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:18cb3ffc9a4e692e16f1a97aea48811600124eb818a96cba222f6b793d49c5c6

Observation 4da21162-b1b2-4b45-a792-31febf25b320 · outbound

This paper cites URL https://aclanthology.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli URL https://aclanthology

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.277456Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:1acdfcb8a237d23add0911c569c1e9e1bd9519590623d05bc5811f59cc1053ce

Observation 63b56686-57f7-4834-a598-187e1d725da5 · outbound

This paper cites Tuning in to neural encoding: Linking human brain and artificial supervised representations of language.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Tuning in to neural encoding: Linking human brain and artificial supervised representations of language

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.283506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:741f4f4ace89b19964aef260d30da172d2fa6f476952036877330ddbf959e9b9

Observation 88ad10ad-238a-4de3-b50c-c1ce2d88cae3 · outbound

This paper cites Llama 2: Open Foundation and Fine-Tuned Chat Models.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Llama 2: Open Foundation and Fine-Tuned Chat Models

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.056811Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:7f7ffdebbfbebcd0ec83c980fbbd5efc7739632439de2979c62f9d353e388a8a

Observation b53ecedc-025a-4479-ab5e-9cc7d79a8560 · outbound

This paper cites BrainWavLM: Fine-tuning Speech Representations with Brain Responses to Language.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli BrainWavLM: Fine-tuning Speech Representations with Brain Responses to Language

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-22T00:04:27.036797Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:254aef9ebf1745d4288cfce615c6f50b0f99ab589fa96849de954b55980eebac

Observation 8d32732a-66aa-4ffe-88db-a2de1fe8eab3 · outbound

This paper cites Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Qwen2-VL: Enhancing Vision-Language Model's Perception of the World at Any Resolution

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-05-22T00:04:27.032285Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:89025a15b045fcfd99d2c5c910577aca8ee61427ceb5ccea80cd37ae18cacc71

Observation f055e0a8-b3d8-45bb-bcc6-077daa71a230 · outbound

This paper cites an unresolved cited work.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-05-22T00:04:27.265269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:c49a43736a336285ea1e72edf84a91405ea7dae485c2f1b7be29663fbfc12518

Observation 90792d82-4934-43df-a3a7-db0831f052ac · outbound

This paper cites an unresolved cited work.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-05-22T00:04:27.272127Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:eed27c8aee3b105e2d74e344fdceaa498a0d2f471b827b4ab76944ebe5ef3840

Observation 456d5385-2b71-4c7c-a6f4-ec843d0f4ea4 · outbound

This paper cites The Wolf of Wall Street.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The Wolf of Wall Street

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.263053Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:1a762fa3612d2ddff034150300c512ffe4441a40a75de704c5a9ff4ab89df813

Observation 41a82bbb-fceb-431d-a08c-7b38b63d1e85 · outbound

This paper cites Goodfellas.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli Goodfellas

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.267166Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:d44c1580046847aba239110edb35ec210741a43a05cbcf5c0aacfa4251fd57fc

Observation 580d4fca-ed72-4077-8784-bbefbe91d4ac · outbound

This paper cites The Hangover.

Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli The Hangover

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-22T00:04:27.278693Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-22T00:02:41.893373Z digest=sha256:4a6ea570117b25931a7832a979ea7ab6245cb8e42f1b8c79886449452a92fc51

Pith citing papers

Observation 9b99c8da-1be8-41d9-b1ae-56b31d0727f8 · inbound

Do VLMs Align Better with Humans than LLMs during Natural Reading? cites this paper.

Do VLMs Align Better with Humans than LLMs during Natural Reading? Task-conditioned probing of instruction-tuned multimodal LLMs: Region-specific brain alignment patterns under naturalistic stimuli

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-06-29T13:03:26.841301Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-29T12:53:45.827516Z digest=sha256:b4ef92e42b84dbdcfce0d9cab9a225cb734e5166cc10d8ad01ed5295d298d130