Pith. sign in

Paper Citation Record · LEDGER

AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2501.02135.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.02135 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:49:53.557961Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-21T10:19:59.955529Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 6c576a3d-7bb2-4124-9f5b-6226267d6620 · inbound

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey cites this paper.

Multimodal Chain-of-Thought Reasoning: A Comprehensive Survey AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 172

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:18:53.285323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-15T17:18:52.996467Z digest=sha256:928f15bbcc314243dd5a900aec4326617dd9109b2be0c79c96c439f95bcc17e0

Observation 710a6990-5b09-451f-beab-49c6ecc256c6 · inbound

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks cites this paper.

MAGNET: A Multi-agent Framework for Finding Audio-Visual Needles by Reasoning over Multi-Video Haystacks AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 109

Resolution
unresolved
no resolver link, observed 2026-08-07T05:49:53.557961Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:49:53.557961Z digest=sha256:70eb1839ea25e4852de133952cb82934a6e659623a6734359aba15c31642d116

Observation 3d4c15d4-49e6-4aea-9529-69d0ef60c2b1 · inbound

EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception cites this paper.

EgoAdapt: Adaptive Multisensory Distillation and Policy Learning for Efficient Egocentric Perception AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T22:41:05.113373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:41:05.113373Z digest=sha256:d65dab006d00bd17d133f222a7ac8d77ce10005dd3e61d294f1d7d102c60bad9

Observation f94943c4-2e0a-4714-975b-c1967b14ffec · inbound

FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs cites this paper.

FutureOmni: Evaluating Future Forecasting from Omni-Modal Context for Multimodal LLMs AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-03T09:29:58.979284Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-03T09:29:58.979284Z digest=sha256:8eee74df5263b67319b2fe67ac63478d6be2e4c2b1d90742b96cf72dbcb624ec

Observation 7e822f70-1aa7-4fdd-9ad2-ececa9b5e6df · inbound

Do Audio-Visual Large Language Models Really See and Hear? cites this paper.

Do Audio-Visual Large Language Models Really See and Hear? AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-05-13T20:58:15.817416Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-13T20:56:19.815569Z digest=sha256:c97a0cb2f3e2d1c8c039c363d5ab2013de2316567596a6ad80256ecd65a53de9

Observation 57e8c60e-a833-4a50-80ed-d173abf87baa · inbound

OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments cites this paper.

OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments AVTrustBench: Assessing and Enhancing Reliability and Robustness in Audio-Visual LLMs

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-05-21T10:19:59.958876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-21T10:16:42.095330Z digest=sha256:db1cf31c82bae5dcdcf3cdb4e350b3a72bb1745448d4ac28382d0e11a4ae8385