Pith. sign in

Paper Citation Record · LEDGER

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training

As of 8 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2507.22781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22781 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:22:11.094770Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation effd9037-0481-4087-af1f-a8d58e086193 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.411665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.411665Z digest=sha256:8d92acb249e01f98332541af01df5076e441ff3cd982ca374c7aee429a3a4139

Observation 0747430c-e929-4d4b-bce4-1dd373c9ef16 · outbound

This paper cites Qwen2.5-VL Technical Report.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.477912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.477912Z digest=sha256:ccadcb630a6956c9a5d81c183fb0a7daeb12c22c4d6dce82ffa1b99ce7ef6b30

Observation ad31c558-72e1-46b5-bcc5-adb8ac9b5e32 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.573195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.573195Z digest=sha256:f7a6dfbf6d9d406890de54dea786d55edff39a787e999079e6677818a06d5f0f

Observation f2d51abc-683e-4a18-ba05-5ca05c79365d · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.477087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:05.641843Z digest=sha256:9bf80fa0c7c9a8ed5ce0be20da77d61207772dc1ee09faab87ee9c96b980b8cb

Observation ab54ee9f-25ef-417d-8411-9d7fae8c0eb0 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.467168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:05.745938Z digest=sha256:72214c04a3dbeffbe2e441a199f295df9904b3b51170364d9b7153f11185701d

Observation c85ce3f1-f6e2-463a-8fcd-c8df423c9f91 · outbound

This paper cites AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.863936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.863936Z digest=sha256:24bb41c736920fdd75092480424bf345ca15307dd77224c1dd63e77bb8993720

Observation 6e5fbcac-e2cb-42b6-9d64-27246ed549d7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:05.896128Z digest=sha256:b0e938f0354e64489f9373e5c9a98cb4cd277608138863ab93cacdb75d6e9852

Observation a5491269-5b38-4986-91db-026a580beefb · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.445155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:05.947834Z digest=sha256:ca120e267db39cd241c176481b48f9a110d65854d4db67490f0df1703e399e45

Observation be9b7e5e-a7f2-4308-a71a-347824d9cc69 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.433466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.036307Z digest=sha256:02aa39d400bd465bfa8c92580028cd86ab345bda7722f7d1a9ac93f077811612

Observation e1a4e95e-0f2c-42cc-bc94-d6fcf5b3ce6d · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:06.232168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:06.232168Z digest=sha256:c815ffa2bc08cc97c4cd970456de529db1005beb78c55e926e71cb0f4bef8537

Observation b3ee7575-72e5-4b64-9a61-379af7054532 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.413407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.325533Z digest=sha256:0f59fd3beb942301cff6085008b1fad93531922c7d1086493c9537e291a44384

Observation 951a740e-9443-47d1-b60e-ca3cce531441 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.402949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.418755Z digest=sha256:5163fe9a073ebf699383c5d9ce6a94c4515f9552580789eac2148450b7f5019f

Observation 83199327-1f1c-43ce-b2e5-83f959260212 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.391508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.477914Z digest=sha256:b15ccd4485f83853d429ef6b2e681a3e0b105bfba1cadb313c590a67637165c8

Observation bc871ed8-fa1d-4bea-8be2-fe007dfce378 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.370167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.662146Z digest=sha256:902c41ec1b3f6ba30483ab9e36a057e256aaa3488dd7642d75b0b76194d362f3

Observation 1ab97031-acef-4f9e-ad8d-c56aaaabeb64 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.358812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.770868Z digest=sha256:4a1acb2510fac26f72c029d2d6e153753e0eb8b5f057df5a529dcfda9a72ef8c

Observation 2b7a913f-8324-4b44-9b1e-1b9e3f5b7820 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.347515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.990439Z digest=sha256:38c46e003ce5c08793a0ee047e2d8b554f8ac241bf01b050d112a69a1014081c

Observation cd284ad0-8c6d-44a8-8089-e4fe71e5290b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.335767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:07.137170Z digest=sha256:7b3bc6e44b05d032fbdabebb9280bb17481ed54fdc8c62f8075835ead33a13d8

Observation 9611d3b2-7769-4130-9ec5-359f6276a384 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.192945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.192945Z digest=sha256:bfe24caf8007220034bc79173e2fa74ec20459d9cdd62abadaed06269ecb86bc

Observation ef17a6f6-5b3b-4eed-a1fd-b7145d65b926 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.239818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.239818Z digest=sha256:e51ebe3eb4b2f0b9af12fe327dc491c3fa739bafd951469cac9c89b0c0111d18

Observation 4200f296-1d42-43ed-b819-ecd860d9a56a · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.367992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.367992Z digest=sha256:ae820f6d3c7b2898dd6ac3c47633a23ec7436d40a0fa6d6edf6dabf516ddfb36

Observation 70d7e93e-960b-4341-aa21-9ae7c69c9739 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.450675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.450675Z digest=sha256:c77ad4fe706025bb5b33bd18513d33f9694cca57a25649dedd9d89e820499c1b

Observation 9e81d721-8450-46f8-8ee6-f1a21ce056e0 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training LLaVA-OneVision: Easy Visual Task Transfer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.629431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.629431Z digest=sha256:a873768daa63f5e243e97c8a12fb531b72879f5bb85507117a3f761bbd8e7cac

Observation 9c729d69-5cf2-4e11-a5d6-6a17c1349d08 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.305056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:07.745701Z digest=sha256:9cf5dfa0e21990cf2eb2ce904c3f1c3b71a6cab44125e183fc6370b491587042

Observation 83e6d246-856a-459c-a113-a803222f8168 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.293319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:07.841615Z digest=sha256:728430cde1dded9c4243a31f5dc87a5814c3030aead9a75e7e84aa4f6240f8ef

Observation 967b6b02-e4b3-42d3-82d9-8529ca70ce21 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.958818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.958818Z digest=sha256:f0bc8674d4afa06c6d74e0de2d4ec853d36b9ba628fa35fae03ea4949db8fe0a

Observation 7a6f3d12-8ea7-4a62-a844-66cdb26152bb · outbound

This paper cites Decoupled Weight Decay Regularization.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Decoupled Weight Decay Regularization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.150028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.150028Z digest=sha256:f6e9361f93b3d98122b6d5413d6bd66c5e32215259ab1df8b177de3e708542c3

Observation e30bebd7-02f5-4476-921a-def0d6d8376b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.786336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:08.282021Z digest=sha256:d01a061dbd3c41a0d2ac0567503684ae4ca31975459add067679016399b0b0e2

Observation cb39c92b-bd58-4f6f-ad10-092f0a1893ef · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.369036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.369036Z digest=sha256:b68f9dd82421b8c02daff792712256d01760fd9b7625983c2679fba44b2217e0

Observation 37c58238-42ff-4bf9-953b-4bed40a2e803 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.500385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.500385Z digest=sha256:defecf5ad6ff548acc4c470b62212e260ee16b0e136faa478c86a84744aae7b4

Observation 1b3d9346-d2f6-4cd7-b211-1a9415148ae3 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.277858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:08.066644Z digest=sha256:55054f24bfeae557c95aadf7a09a7d9537b93760300a0253db31fff2d6e1ffc7

Observation a1931f7f-fb2e-4e03-a7c1-f53aa8c2a5ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.257452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:08.744453Z digest=sha256:7845aa6c15aa8f2ad192ebc55ff691b959f9388677d95964e1c495e3e91c4bcf

Observation c7509f94-6061-400e-a3ab-c688ac59ffe7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.246457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:08.879157Z digest=sha256:01fd41fcf3ebeff4bc9b01160b747cb43e34781f21f8923ddba188b3ca92f057

Observation e50fe659-5e26-497c-971d-a732af370509 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.036940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.036940Z digest=sha256:badc58779711b0719453fd334404a686ee491dc3d8ba85f99ec422c8ee047435

Observation aed89379-dbe6-4ca4-84f2-23e2fba0b7cc · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.229217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.171661Z digest=sha256:726ae00016b672d398370df27162f102d8faa4ebd19cb46f1d5822fafe070191

Observation 950af699-33d2-4395-a2a1-eb8434b832ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.268243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:08.634729Z digest=sha256:190c434607ee9aea61b73e81a371a5ee382b36c7242e74470a9ee07dd5cd75ef

Observation c6c5c237-94e7-4014-9c08-2d56396cf76f · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.213411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.397609Z digest=sha256:018db67944fe36af912981f0c4085f561ce156b3376a1a99175382ebec14c842

Observation a0647a3f-c126-49dc-baa4-efa1791696ea · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.204296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.470703Z digest=sha256:14fd0db7a2fd900f4979ba858263d24341e5c5a3cb965998fcaf1f71f454fef7

Observation d28d1515-0ce5-4446-99ca-2500ef15e566 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.106579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.527545Z digest=sha256:0d33c5360af9e040ac35a8403556b35136bd47a9002a52f7ac373b1d2a2f707e

Observation eff0526c-e9e6-4ba3-991b-9c7ef0626f98 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 39

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.616681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.564187Z digest=sha256:5293ed9b607cec85775dfc391a59fdd526fc398925c7306220e03617f425332b

Observation 21685015-2d51-4f6e-827e-06f166f1b562 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.313558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.313558Z digest=sha256:e63b5090bb3af34f78b0b19c30619b07190466b1da97ef0054769eb994e4deaa

Observation f6e54fc4-3801-41d8-90dd-2a439f743eb6 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.183491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.723169Z digest=sha256:73b46afc6058eb0e4f94fb32f59b481bc110852c44833b4152b98d77aeec54c0

Observation 69ecbad6-15f3-4c3e-b905-19afa07a3021 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.174336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.753390Z digest=sha256:0abacea4189dba621c3e77481babb0f3328a77034fde6359900af36c8f446dce

Observation 550cf1ac-bb23-4821-b51c-c87adf33cdb1 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.164644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.841458Z digest=sha256:92658c77e5c56a88518d6fc8c94021e33ffdd114eab11218b4c03c1c5ed9db0b

Observation 67a6ee80-0852-4adb-9752-93e8c9d8ef55 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 44

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.459447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.925918Z digest=sha256:29c78e3116aaab8b34805ef3465b5b3ebe9722d72b56aa5deabc0c59851a2d88

Observation 22cc9f4b-dd66-4b1d-9dc2-f81a0c8f42b5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.193431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:09.652357Z digest=sha256:4a5794503062e0cd282c98f75d7398a3ebe3dbf9d3ee0ccf548a004fafa7cc2f

Observation aa3a5a36-b428-4ce8-b9ae-39384141ac51 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.154808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.171917Z digest=sha256:d50bd02bc98c8cf34d91c854a0deb2420d3227612dd1078da6fbe16362459b3f

Observation 4ac61163-a401-4056-a125-e57f5f5d7156 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.252632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.252632Z digest=sha256:69f98a5c02be9e081131c01059b33bf52cf02eee3150173b7ccb63ee7ab96791

Observation 0de8c958-c5e2-4b57-982b-ef280b189337 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.334837Z digest=sha256:a54547df1cb03f6a04d20831a6656db37f2f2a941a15979596d8e9c2fe4de94f

Observation 8768b146-5d26-4721-b7fb-5b8611c43ba7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.134274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.411765Z digest=sha256:e64717825ceeb38d2408bb7ac080d8b02147bb5d0856f14ffe36c73d2915e2b3

Observation 1171e4f6-1351-435f-ae4e-ac3f5ac56389 · outbound

This paper cites CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.032757Z digest=sha256:666d59a5cf306a28622d8071bd0096e99b8dfe3116ded38bf3cf25f1ad85fbca

Observation f4053ef7-be47-4908-be0b-255e067d5932 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.113630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.600580Z digest=sha256:2f8320d2f5391d8adf990c9657da287201f3211feadf4eed9bbab399c9f37559

Observation 0fa8b4b3-86be-4b42-b828-3ba73d142268 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 52

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.289869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.642109Z digest=sha256:dc228fa672307d1aec1f76c08d980e3837a68d32dbb1f19b5a7000d94dbe2466

Observation 06a8dfcf-bfd4-46a4-8af0-f84c57d4259c · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.727452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.727452Z digest=sha256:c184ceee5b6ab2fc593b67c55af196b28c9c99345a56edf40bbd538941cfe88b

Observation b942b722-472b-4e67-bf01-648e65d2b7a8 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.095957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.821908Z digest=sha256:d3463060197cb52702c45648a2b020314ef978e8a24d2beb044f190616a8115b

Observation e0b9dcc3-b21a-44a0-a7ba-60578e7b11f5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.124121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.511071Z digest=sha256:0533811fff6d2f3c1c48ab933e80decc8202575e96210e91cff351e6928e89f5

Observation 64682249-b972-4a9f-89aa-cf6e4eb7dae2 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:11.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:11.003082Z digest=sha256:ea54727f7211c961a492ac786fda181c8b113efd1f4a69449e2d8009b9b48ad8

Observation bf2b3cec-c6b8-4b10-b4dd-de27c27bcd7b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:12.967277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:11.094770Z digest=sha256:64ff06c709da1d93159ffb769e586638627c1a62e9b235efe9e1a937eb80d94d

Observation d2f5310c-e3d0-4be9-801b-ded63e567e65 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.083669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:10.908903Z digest=sha256:e8a00d576ea27317dc14f10434240b13591e99562ea9db1970d3e730e93591bd

Observation 8a192271-f78a-483e-b291-d47b09b5462c · outbound

This paper cites Pattern Analysis and Applications 25, 4 (2022), 981–992.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Pattern Analysis and Applications 25, 4 (2022), 981–992

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.380191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.573227Z digest=sha256:2ab3811327bb8801709bbf42625bff270838cef8412a4c1fc3a956a0f09f9f71

Observation 413f7d4f-1e0c-4756-a4dd-e455d85535a7 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.423278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.137860Z digest=sha256:afdb96856b06b6c682afddec8587378b2cac2ea9fad72e69544fddcd9c359052

Observation 1b457127-83af-4786-8a86-36ed0d4483e1 · outbound

This paper cites In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT).

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT)

Reference 2024

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.712624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:06.880564Z digest=sha256:fe8190f228aa74a7a4badc45ef0761fc148eb87a347a59aae71f9ef38ff441ea

Observation e5348dba-3fe9-4ea4-a7fc-2ebc60caace8 · outbound

This paper cites StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:22:12.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-06T11:22:07.063747Z digest=sha256:5edb8b38ecf0a2d7f9489d25435a126df5bc6ba8412d71d5315a79826eb48b05

Pith citing papers

No inbound Pith citation observations are available.