Pith. sign in

Paper Citation Record · LEDGER

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought

As of 6 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 0 inbound Pith citation observations for arXiv:2605.09906.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.09906 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-12T04:33:07.477217Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

44 of 44 outbound references displayed

  • verified exact5
  • verified fuzzy35
  • unresolved1
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch3

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 45764ddc-6dea-4eee-bb4c-79fe2afd35b9 · outbound

This paper cites 2004 , issn =.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2004 , issn =

Reference 1

Resolution
verified exact
doi, observed 2026-05-12T04:36:21.179467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e791c9f9974aff80054515e6bdf2b716e021c66ea482e7626a0e680cf5a5082a

Observation 0ba1a00a-68cd-40f9-82c2-15c8d53a8c9e · outbound

This paper cites Nature Reviews Neuroscience , volume =.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Nature Reviews Neuroscience , volume =

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.478724Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:9ddc22b201b1aee35d457ac057620a37756f1a4c377ea3624bd0444cb36f6ae7

Observation c1473872-cc1d-4610-97d8-4d9d3ca42ba5 · outbound

This paper cites 2018 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2018 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.481954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:175a23a574a4b8a23da7255baf5108501ea1f3e0a76a43939f4c904baffacf5f

Observation 637d90e4-6cb4-4887-9601-278f979901ed · outbound

This paper cites Quantifying uncertainty in answers from any language model and enhancing their trustworthiness.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Quantifying uncertainty in answers from any language model and enhancing their trustworthiness

Reference 4

Resolution
verified exact
doi, observed 2026-05-12T04:36:21.184027Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:085399a8b4b35df5e57d37ad2ae5363ecbe9ff87d820179c3a6749cfe3024756

Observation 81efc321-d08d-431f-bc5d-80f6592c24ba · outbound

This paper cites Proceedings of the 30th ACM International Conference on Multimedia , pages=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Proceedings of the 30th ACM International Conference on Multimedia , pages=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.485288Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:c8463ecbf6afd5a4ecda1e363aab8c1a7ece31935b0ed96515df7d9e46cb9d44

Observation 6fce5ff5-d4c7-4443-84e5-2b4bc0565708 · outbound

This paper cites 2022 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2022 , eprint=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.488860Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:b9193a30baac3fc0e281ac1c3363b44255f0cb06eaff0132b72e7c8069cc1aeb

Observation 784d1f96-9c53-4826-8948-8ec487851706 · outbound

This paper cites 2025 , isbn =.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , isbn =

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T04:36:21.168220Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:bfb8ce5657d625927790c27234598dcb8afcbd3544a357eb5701d30e97ad8ba1

Observation 9b5472ae-fd87-4225-b4c7-ee9bc99016c5 · outbound

This paper cites 2025 , isbn =.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , isbn =

Reference 8

Resolution
verified exact
doi, observed 2026-05-12T04:36:21.172871Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e78b33c23a6cdb78f38002cbc1e580d7c18dd53e95c5fb2e0e0c7abca85d1c9f

Observation 2f40871c-00e0-4560-9acf-372e1e7a5bbc · outbound

This paper cites 2023 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2023 , eprint=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.562941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:9135c156c94df74a9c0e5b61125b55a8f7f59da7504948861a099c3c1bedfde1

Observation cce4da37-e48a-4fd7-971b-bf9192d7c46e · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.548734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:cb9a5d7e1fbff6a454eaa6c160dc6b288683b327cd331fbfa6c18daf807017d3

Observation 563e8657-1180-4c9a-9143-e68c76b584cf · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.559356Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:2dc26ccb6bb9aa61008292d7a0db90a784582a04c3c85302c608c18b6e371a53

Observation 9b2fcdb5-0b37-44a8-b3cc-e25c59b5eb92 · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.508380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:8757e40b7b65d2f07372f2d7956ea2cc4f1f32a40d42094a669156e70bbb3a37

Observation 8e286190-b0a1-4057-95f3-42b9b130d89c · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.538141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:779b8afd5c104ebb82f364c608a46aaf9ccb03a08930872b3d0cf90dcc8d9e31

Observation 4833746b-58bc-40ae-947a-29e90032d0d3 · outbound

This paper cites Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Proceedings of the IEEE/CVF Winter Conference on Applications of Computer Vision , pages=

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.523968Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:2b8ed17b9493df31a57f4c2b24bd4cb5ab4861e650374e6a8cf73be3fa35c262

Observation ffbb2a56-69b6-4fc7-a84e-6d0f7294b082 · outbound

This paper cites 2020 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2020 , eprint=

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.545425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:568c6d11e99712112d2c5216fbf124be97776adbb05fc275f8fa0efcc64e9565

Observation 10a27652-d95f-4a40-a902-fd5b91862d7e · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.495443Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:f4892ff88e2cc27f843984ee25fad11397e8ffe47394469d3996ec7de429ceaf

Observation c941e73d-9f9d-4f02-90f1-435e5887170e · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.492273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:173a3e5acb811f8c76b1650c222b93f5ac87eadbbf4f8c67dbc9b890d5371c79

Observation 04336c9f-6a13-483b-8152-3d6dd7ef0cfc · outbound

This paper cites National Science Review11(12) (2024).https://doi.org/ 10.1093/nsr/nwae403,http://dx.doi.org/10.1093/nsr/nwae403.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought National Science Review11(12) (2024).https://doi.org/ 10.1093/nsr/nwae403,http://dx.doi.org/10.1093/nsr/nwae403

Reference 18

Resolution
metadata mismatch
doi, observed 2026-05-12T04:36:21.141803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:2e34101cea38165af3cb3d3ee84480c2fe4f44440feeccb686075ddf57671cc7

Observation 95ce46cd-8598-4e98-84d9-3c8fe0b0e033 · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.501728Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:79d915935cd4bb3683facd32f1f0e275b73e662896cf2309953106ee580e6676

Observation 7c1f4987-f77a-4b08-874c-2c967ab347c8 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.527304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:17b55344b845a3d09712a963bcf01c723521670a3d78c80c7c7cefec216ea8cb

Observation 535c203b-40b6-4b16-88fe-2caaafcebe06 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.530578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e9b98663053bd53c51d1a70f01b754f5e050fce4d330c298b2a63e2b95216ddf

Observation 89c4f683-7652-4e9a-8d86-6fa52c307737 · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.552259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e895dfb9401f4626e4fe708eea319b60312a72d658ea8bf36273db7d88bf1c50

Observation 05b6f224-a575-464e-86de-ac1697d48362 · outbound

This paper cites doi: 10.1038/s41586-025-09422-z.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought doi: 10.1038/s41586-025-09422-z

Reference 23

Resolution
metadata mismatch
doi, observed 2026-05-12T04:36:21.137050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:87cf0ea73ac91780a24c638ffd1fea4248418433f824170172459a07c1cd9ee3

Observation 6f68a65a-daa6-4338-bfa5-7e2272e4520b · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.505077Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:c133495caf398d19cdcc20884379627599e6967b6b6d523fd539482df6fc06a2

Observation 942b2bbd-fe5f-487a-9dfb-7bd0c487b2ae · outbound

This paper cites 2023 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2023 , eprint=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.514731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:2c58924adfffd5acf6ca7313b9a2eb4d5d09792a8c6bf10959e2c878c3c640b2

Observation cb1cce58-5526-40ba-8235-24234745a489 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.511484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:01ff5943d1fd5133390322b9e5590af4d2e6641884e0f695ccc064af8be08840

Observation e0687a99-8e46-4064-b446-db7046b61490 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.567392Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e5a2f50c501924feb8bdc59a1b896ae9b239d3e0e623748ddf2674e9f910a910

Observation a3036e61-6463-461f-8834-8d9e36ea5f3e · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.498839Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:c001074f66cb1d7ebb68446a950a65dd2f347c243b49b2593c6c27a38d8db198

Observation e7cd9466-f6f1-416d-8bff-e7b30c04b9ea · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.555757Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:72b1b7125ce9ec5de351b0e817d7e79eaee6275fd0ff0071fe3f8d075293e6cb

Observation 50742264-1049-4bee-baf7-72571cc144ff · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.518160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:7a8ca0051228adc041e087c12a8fbf7016faf5f256e3163c6abf05bba1f7dbe5

Observation 0b1dccc7-0075-4aab-a7f6-69cfd88af1b1 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.521061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:8fa14ae220783dd08a6c98eee7476e02e4672303ab4a1b3934427ccf39a9d4e9

Observation b4bcb8fd-eb69-4731-892d-9a706369e735 · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:36:37.890631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:103d426c07f94310562810a71f3b11034e8931a396b1717a2a610eda9b326158

Observation 1d3b8829-05dd-4bc3-9a3b-51e3be4003e6 · outbound

This paper cites Spurious Correlations and Beyond: Understanding and Mitigating Shortcut Learning in SDOH Extraction with Large Language Models.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Spurious Correlations and Beyond: Understanding and Mitigating Shortcut Learning in SDOH Extraction with Large Language Models

Reference 33

Resolution
verified exact
doi, observed 2026-05-12T04:36:21.145754Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:2fe1d1e63b4158254aa9649624bd5638793c052ce6f52f1e0aab3ab5076429a9

Observation a47b40e1-4c02-49a6-9872-ba4c2399f30e · outbound

This paper cites an unresolved cited work.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-05-12T14:31:39.593131Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:5724ac55722149270ab646a466c4c44fd01a61053aeb8c282cc89a49cc16b71d

Observation 804fab10-fdfc-4cc8-8b9c-22dff6067d60 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.588621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e0aee8422c97b24728f024b6808d9f356dbadf0d718aa5363ec9833772b0d574

Observation f0b1ea68-cb16-4503-bd36-01a2903d33b0 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.534652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:e9b31d9f059c1d7d200da5d035010d9fe29db44302be58ff7aba4f486e486af7

Observation 2f420385-78fd-4acb-88dc-93418a9f5cd8 · outbound

This paper cites 2023 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2023 , eprint=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.542067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:cb5aae5c06624bba1e490b5096c3c7d91cf4d2a6e478310b25cf9b2943c77eaf

Observation 9177615a-d806-4d8e-a6ac-fbe1f4500c81 · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 38

Resolution
metadata mismatch
local_arxiv, observed 2026-05-12T06:11:22.896828Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:9c3d702b0cd44759eecda62eb3d1f740f15f197ada07ddbb599b578c92999a02

Observation e7e3579f-1e70-407b-be99-2ea12c9a229a · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.584931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:fcd0ec16197c4751ef0b7878a9ce04401c05cc43c49a282b39fa1d9a41fce30a

Observation 2822b90d-8486-4abf-8833-edf950f8e5ba · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.581582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:58389876dfb838a0a76abd03cedda1c45616dc11187cdcd486dc4c49653c4eeb

Observation c9cd4231-90d9-4315-a4b3-142477646a18 · outbound

This paper cites 2024 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2024 , eprint=

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:36:37.898575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:5dcd5483840a6a169b7005f1e1f4f3a0ab91a5bda5b80aef2b455a752238318d

Observation 11e68aba-61ac-4cfd-a653-fa6ba4269b21 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.574484Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:5ab1e92f3be119ae08fe3d95be173c05f0b46b7627b6afedd2ec7f11e099dfa4

Observation edae379c-0136-4572-9937-f373bbe0f882 · outbound

This paper cites 2025 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2025 , eprint=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.578188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:f685204393dfc30fbc07a7ab2a22982c7badea5c36179caed3d0081eddb51455

Observation c24e24b2-531e-498f-b247-6691d5386904 · outbound

This paper cites 2017 , eprint=.

Separate First, Fuse Later: Mitigating Cross-Modal Interference in Audio-Visual LLMs Reasoning with Modality-Specific Chain-of-Thought 2017 , eprint=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-12T14:31:39.571006Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=arxiv_source observed=2026-05-12T04:33:07.477217Z digest=sha256:7005e20541f4e5170e15ccca4240531b7af7f5f81a1f96b828c9757621beccde

Pith citing papers

No inbound Pith citation observations are available.