Pith. sign in

Paper Citation Record · LEDGER

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training

As of 8 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2507.22781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22781 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:22:11.094770Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation effd9037-0481-4087-af1f-a8d58e086193 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.411665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.411665Z digest=sha256:8d92acb249e01f98332541af01df5076e441ff3cd982ca374c7aee429a3a4139

Observation 0747430c-e929-4d4b-bce4-1dd373c9ef16 · outbound

This paper cites Qwen2.5-VL Technical Report.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.477912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.477912Z digest=sha256:ccadcb630a6956c9a5d81c183fb0a7daeb12c22c4d6dce82ffa1b99ce7ef6b30

Observation ad31c558-72e1-46b5-bcc5-adb8ac9b5e32 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.573195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.573195Z digest=sha256:f7a6dfbf6d9d406890de54dea786d55edff39a787e999079e6677818a06d5f0f

Observation f2d51abc-683e-4a18-ba05-5ca05c79365d · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.477087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:05.641843Z digest=sha256:d335c20c546c317c6d5a88d7fd1a275f84fb6e6bab38ed6f3592aefc7274413b

Observation ab54ee9f-25ef-417d-8411-9d7fae8c0eb0 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.467168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:05.745938Z digest=sha256:70346e31c13ec5c459eb60ceab834c60ffced356245786827014011b8454ea61

Observation c85ce3f1-f6e2-463a-8fcd-c8df423c9f91 · outbound

This paper cites AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.863936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.863936Z digest=sha256:24bb41c736920fdd75092480424bf345ca15307dd77224c1dd63e77bb8993720

Observation 6e5fbcac-e2cb-42b6-9d64-27246ed549d7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:05.896128Z digest=sha256:4275e26965b53b4e2aa1736d67c1b30e9e14294c597b1cd6c0f7a859bb389407

Observation a5491269-5b38-4986-91db-026a580beefb · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.445155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:05.947834Z digest=sha256:3fb1bf4b5dd455ae22c472512c7a2d92ac26770e31e5d8dcb6da8c49b722521e

Observation be9b7e5e-a7f2-4308-a71a-347824d9cc69 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.433466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.036307Z digest=sha256:4de58ab1c409efeae1366394716fcd99341309e57d85ccc522716d279f777b57

Observation e1a4e95e-0f2c-42cc-bc94-d6fcf5b3ce6d · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:06.232168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:06.232168Z digest=sha256:c815ffa2bc08cc97c4cd970456de529db1005beb78c55e926e71cb0f4bef8537

Observation b3ee7575-72e5-4b64-9a61-379af7054532 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.413407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.325533Z digest=sha256:3c9dd380f6feb0addb596015794513553aacbcd74105502170dc0507f7b6cb6d

Observation 951a740e-9443-47d1-b60e-ca3cce531441 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.402949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.418755Z digest=sha256:d45c671467a7e11a6617a6a660a6781e5c108ffb9aaa501ca131d6089625aec2

Observation 83199327-1f1c-43ce-b2e5-83f959260212 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.391508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.477914Z digest=sha256:813877fcd363eca954e9bf0830e9afdce2c55cfa1531a7298296387f13224601

Observation bc871ed8-fa1d-4bea-8be2-fe007dfce378 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.370167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.662146Z digest=sha256:26970c4b368bc59abd85435b26e795eee48b1d485344076a40bbda13c2eaf160

Observation 1ab97031-acef-4f9e-ad8d-c56aaaabeb64 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.358812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.770868Z digest=sha256:752d50ce49161ffc4ffcaad0ccf15f37c1006a2026dd517f4b7d987b4f194f11

Observation 2b7a913f-8324-4b44-9b1e-1b9e3f5b7820 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.347515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.990439Z digest=sha256:01251473cd03ee5fe9a84ade9bff959df64149aa5b4e27e52e530c2901333d1a

Observation cd284ad0-8c6d-44a8-8089-e4fe71e5290b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.335767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:07.137170Z digest=sha256:ee49fa72ea7faf95facc9ad6a6b8c9ad30cce43cb8acfb7a5d419f2668cf82c7

Observation 9611d3b2-7769-4130-9ec5-359f6276a384 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.192945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.192945Z digest=sha256:bfe24caf8007220034bc79173e2fa74ec20459d9cdd62abadaed06269ecb86bc

Observation ef17a6f6-5b3b-4eed-a1fd-b7145d65b926 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.239818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.239818Z digest=sha256:e51ebe3eb4b2f0b9af12fe327dc491c3fa739bafd951469cac9c89b0c0111d18

Observation 4200f296-1d42-43ed-b819-ecd860d9a56a · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.367992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.367992Z digest=sha256:ae820f6d3c7b2898dd6ac3c47633a23ec7436d40a0fa6d6edf6dabf516ddfb36

Observation 70d7e93e-960b-4341-aa21-9ae7c69c9739 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.450675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.450675Z digest=sha256:c77ad4fe706025bb5b33bd18513d33f9694cca57a25649dedd9d89e820499c1b

Observation 9e81d721-8450-46f8-8ee6-f1a21ce056e0 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training LLaVA-OneVision: Easy Visual Task Transfer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.629431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.629431Z digest=sha256:a873768daa63f5e243e97c8a12fb531b72879f5bb85507117a3f761bbd8e7cac

Observation 9c729d69-5cf2-4e11-a5d6-6a17c1349d08 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.305056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:07.745701Z digest=sha256:31384b1bf6b681255daba3fe87686e02114c50929d882bfd04fe20f1ef0c8ae1

Observation 83e6d246-856a-459c-a113-a803222f8168 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.293319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:07.841615Z digest=sha256:0d887e906a10e44c4dbfce55d07cc55cba7c234b5d7b330c6fe953e48730e0f2

Observation 967b6b02-e4b3-42d3-82d9-8529ca70ce21 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.958818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.958818Z digest=sha256:f0bc8674d4afa06c6d74e0de2d4ec853d36b9ba628fa35fae03ea4949db8fe0a

Observation 7a6f3d12-8ea7-4a62-a844-66cdb26152bb · outbound

This paper cites Decoupled Weight Decay Regularization.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Decoupled Weight Decay Regularization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.150028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.150028Z digest=sha256:f6e9361f93b3d98122b6d5413d6bd66c5e32215259ab1df8b177de3e708542c3

Observation e30bebd7-02f5-4476-921a-def0d6d8376b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.786336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:08.282021Z digest=sha256:084ba99cb709437887d69ad0eb5a087878656e75c1dcc3e2efa7233983a27744

Observation cb39c92b-bd58-4f6f-ad10-092f0a1893ef · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.369036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.369036Z digest=sha256:b68f9dd82421b8c02daff792712256d01760fd9b7625983c2679fba44b2217e0

Observation 37c58238-42ff-4bf9-953b-4bed40a2e803 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.500385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.500385Z digest=sha256:defecf5ad6ff548acc4c470b62212e260ee16b0e136faa478c86a84744aae7b4

Observation 1b3d9346-d2f6-4cd7-b211-1a9415148ae3 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.277858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:08.066644Z digest=sha256:c3f94facb161f8bd4a4f20ba31439ffd2e4f09014c96c732f742bd9ef342e700

Observation a1931f7f-fb2e-4e03-a7c1-f53aa8c2a5ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.257452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:08.744453Z digest=sha256:7c5a05e61d12af1b508f5d266b5f523ac6b5f85d88dfb6ad498de3faf742349b

Observation c7509f94-6061-400e-a3ab-c688ac59ffe7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.246457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:08.879157Z digest=sha256:d7993f1e24d209fa4cd0d5e9512e5551e6dc0b63740f79fc8d40af48dd575a87

Observation e50fe659-5e26-497c-971d-a732af370509 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.036940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.036940Z digest=sha256:badc58779711b0719453fd334404a686ee491dc3d8ba85f99ec422c8ee047435

Observation aed89379-dbe6-4ca4-84f2-23e2fba0b7cc · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.229217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.171661Z digest=sha256:8ce4a923cb2b3bfc5cd22209f5ac9783fcfb62857b1e866c78c924d5bd39a32e

Observation 950af699-33d2-4395-a2a1-eb8434b832ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.268243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:08.634729Z digest=sha256:b3c5787a898ff8369f19fcff4c6190f48f64806c9c087ff1fe3c6fa1cbf3e8f0

Observation c6c5c237-94e7-4014-9c08-2d56396cf76f · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.213411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.397609Z digest=sha256:c354d0a30189ea65c0f9b9c84bab43d45919227ddc69fa8b2df9386aca3367ce

Observation a0647a3f-c126-49dc-baa4-efa1791696ea · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.204296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.470703Z digest=sha256:c92cdcc14db1f6233ee2c2b684ca31f9ac686713cd47ba32d46d3e5773efd564

Observation d28d1515-0ce5-4446-99ca-2500ef15e566 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.106579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.527545Z digest=sha256:e6568dbc28e7c6e9a2aa21c576dd3bc9b2275e7a6a23d6ecbfc6f96c9a9519fa

Observation eff0526c-e9e6-4ba3-991b-9c7ef0626f98 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 39

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.616681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.564187Z digest=sha256:8e4424362aa9f1f9b29b499bdfecc9941700acda363b8afac18bb1b41d53191f

Observation 21685015-2d51-4f6e-827e-06f166f1b562 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.313558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.313558Z digest=sha256:e63b5090bb3af34f78b0b19c30619b07190466b1da97ef0054769eb994e4deaa

Observation f6e54fc4-3801-41d8-90dd-2a439f743eb6 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.183491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.723169Z digest=sha256:3a80e79f09c1ff8bc485d5f269521c27b8b9cef71102f13f775338c05668426e

Observation 69ecbad6-15f3-4c3e-b905-19afa07a3021 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.174336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.753390Z digest=sha256:bbb819a2f79b0da00b4176d2f7bda1310d1e5c2744d317943225e3f13c71cfdf

Observation 550cf1ac-bb23-4821-b51c-c87adf33cdb1 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.164644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.841458Z digest=sha256:7a1f12b86d703ccbd8b37a2c2aecf576c96671c9a9e1d0f45be375a3b5edd41b

Observation 67a6ee80-0852-4adb-9752-93e8c9d8ef55 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 44

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.459447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.925918Z digest=sha256:79581589d2e543331c6e371350337710aedef6515eeba838791e21fc2f881c84

Observation 22cc9f4b-dd66-4b1d-9dc2-f81a0c8f42b5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.193431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:09.652357Z digest=sha256:d6213a0a2daa5ab12f2a63397155c0f04902e407c33568c785b8421eb2b6e345

Observation aa3a5a36-b428-4ce8-b9ae-39384141ac51 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.154808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.171917Z digest=sha256:0e0beb824ff26d27dc0bbd1901aca221ac1b191ce42c5de93f0726d418f35bf4

Observation 4ac61163-a401-4056-a125-e57f5f5d7156 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.252632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.252632Z digest=sha256:69f98a5c02be9e081131c01059b33bf52cf02eee3150173b7ccb63ee7ab96791

Observation 0de8c958-c5e2-4b57-982b-ef280b189337 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.334837Z digest=sha256:e51947a9352cc7b705ce7c2e96e7c605340975ab99b9c5c7ac4a96db993f8a90

Observation 8768b146-5d26-4721-b7fb-5b8611c43ba7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.134274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.411765Z digest=sha256:87ff3b82926d7695f3df3a6d7d9963bb0ae364765f0021cfefa0ba843518fd6c

Observation 1171e4f6-1351-435f-ae4e-ac3f5ac56389 · outbound

This paper cites CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.032757Z digest=sha256:666d59a5cf306a28622d8071bd0096e99b8dfe3116ded38bf3cf25f1ad85fbca

Observation f4053ef7-be47-4908-be0b-255e067d5932 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.113630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.600580Z digest=sha256:338238f9b79d36390be41083157008ab24fb221d2928c97f7a6f3b52d165a151

Observation 0fa8b4b3-86be-4b42-b828-3ba73d142268 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 52

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.289869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.642109Z digest=sha256:9c61d0b61ac4c8889c33355b0f1d7374be1340f808b1ecb8f652aa69d6ca6bf6

Observation 06a8dfcf-bfd4-46a4-8af0-f84c57d4259c · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.727452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.727452Z digest=sha256:c184ceee5b6ab2fc593b67c55af196b28c9c99345a56edf40bbd538941cfe88b

Observation b942b722-472b-4e67-bf01-648e65d2b7a8 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.095957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.821908Z digest=sha256:24d808a743a687b5ace9f39fe16b186648190a4f928c1815412e6e4105843d4d

Observation e0b9dcc3-b21a-44a0-a7ba-60578e7b11f5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.124121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.511071Z digest=sha256:46fe821e662cfc751077b1ba3a3e56618acb2f0ce2862543db9e39667e5e6ec5

Observation 64682249-b972-4a9f-89aa-cf6e4eb7dae2 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:11.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:11.003082Z digest=sha256:ea54727f7211c961a492ac786fda181c8b113efd1f4a69449e2d8009b9b48ad8

Observation bf2b3cec-c6b8-4b10-b4dd-de27c27bcd7b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:12.967277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:11.094770Z digest=sha256:c3fb7b6d1c87e74fda9891924e97d060359c0f060e61423c0a8039a06c859886

Observation d2f5310c-e3d0-4be9-801b-ded63e567e65 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.083669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:10.908903Z digest=sha256:15c499231efa32a2697da04c678e937c1780203fd7547dcb6fb49ec0f064be0d

Observation 8a192271-f78a-483e-b291-d47b09b5462c · outbound

This paper cites Pattern Analysis and Applications 25, 4 (2022), 981–992.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Pattern Analysis and Applications 25, 4 (2022), 981–992

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.380191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.573227Z digest=sha256:35ee13ce89ec8c8bd3ef958612a0b7dbfca35e2f6cff9d5ccea0f5569d4bd9fd

Observation 413f7d4f-1e0c-4756-a4dd-e455d85535a7 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.423278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.137860Z digest=sha256:31a90efc12e7745618fb6febd5cc273f7bf77a19fcfa715a03b762df2f166884

Observation 1b457127-83af-4786-8a86-36ed0d4483e1 · outbound

This paper cites In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT).

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT)

Reference 2024

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.712624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:06.880564Z digest=sha256:64cc43a5ce576a8a018e522a0e42a932532e2ec870289cecbed397bd7852e4b2

Observation e5348dba-3fe9-4ea4-a7fc-2ebc60caace8 · outbound

This paper cites StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:22:12.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-06T11:22:07.063747Z digest=sha256:b2902a58da30ee05c4012a261b25b38a9488bc8fbf14b73a3f60bad2eb08cbc6

Pith citing papers

No inbound Pith citation observations are available.