Pith. sign in

Paper Citation Record · LEDGER

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training

As of 22 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2507.22781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22781 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:22:11.094770Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation effd9037-0481-4087-af1f-a8d58e086193 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.411665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.411665Z digest=sha256:6d467702f788d3f4e21481f4035d7db09e9aa892b75b733a6fdc2240a0479963

Observation 0747430c-e929-4d4b-bce4-1dd373c9ef16 · outbound

This paper cites Qwen2.5-VL Technical Report.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.477912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.477912Z digest=sha256:5e928cd06baeb5b16823f0e59e37ca19207524d9c0acc8c6bb82ed48766f585e

Observation ad31c558-72e1-46b5-bcc5-adb8ac9b5e32 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.573195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.573195Z digest=sha256:f684d614cce185b88e910d72efbddb80ce8431961879099119f5e9649be887f5

Observation f2d51abc-683e-4a18-ba05-5ca05c79365d · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.477087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:05.641843Z digest=sha256:8a0830b76bb2eb715c11bee3047917477f5fa0bb2d8f106545416692266a1126

Observation ab54ee9f-25ef-417d-8411-9d7fae8c0eb0 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.467168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:05.745938Z digest=sha256:cd909adc352c8f11bc75c8329182f0ab1fff227e0b5635f06a2a30b243c5c42f

Observation c85ce3f1-f6e2-463a-8fcd-c8df423c9f91 · outbound

This paper cites AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.863936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.863936Z digest=sha256:39e07e2f372df299e6c38cf07766f7623fe2dab89161331ba4ffd5baaca53d64

Observation 6e5fbcac-e2cb-42b6-9d64-27246ed549d7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:05.896128Z digest=sha256:2562302107a8d9e5dcdd3bce8aa0d25322bf77d9b3e745441ac17e7573dbcc1e

Observation a5491269-5b38-4986-91db-026a580beefb · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.445155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:05.947834Z digest=sha256:8562f9c5cba1f73fb8082fed925c6532a9f7d7610677241bc1e079b43cf631ae

Observation be9b7e5e-a7f2-4308-a71a-347824d9cc69 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.433466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.036307Z digest=sha256:4b5219583b0d893fe62d5f277f91735dcf1f455ef99a356267ebf8f21e28901e

Observation e1a4e95e-0f2c-42cc-bc94-d6fcf5b3ce6d · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:06.232168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:06.232168Z digest=sha256:f105e334e0e889c742beb26b55531074b010367773de53f0a086744f132d6068

Observation b3ee7575-72e5-4b64-9a61-379af7054532 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.413407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.325533Z digest=sha256:8669a507f03b506a7fe586c6dc6dcc11e09d8fb3165f6c993c224f82d0e8413f

Observation 951a740e-9443-47d1-b60e-ca3cce531441 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.402949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.418755Z digest=sha256:3d65f3c1a156e33edae36bab3bccda8d8ac48532f8c786e1a1adfe79b09bc502

Observation 83199327-1f1c-43ce-b2e5-83f959260212 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.391508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.477914Z digest=sha256:8ea360aca602900f8e9e615d513aa5a6b6096e0e47078526a37df5cd38f692b4

Observation bc871ed8-fa1d-4bea-8be2-fe007dfce378 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.370167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.662146Z digest=sha256:9034c2e1c5d8b4707eba39aee77b8937f1b0d45b98eafa02d648013a86ea249e

Observation 1ab97031-acef-4f9e-ad8d-c56aaaabeb64 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.358812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.770868Z digest=sha256:1876636351fd3f113de5d1219b98fe74d9a5ebca84a91ad76fe8d5423f0835e2

Observation 2b7a913f-8324-4b44-9b1e-1b9e3f5b7820 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.347515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.990439Z digest=sha256:155dccd57d53e9c558fb6476f8ca7424c9da2f6c148cf46672d2853d565a1265

Observation cd284ad0-8c6d-44a8-8089-e4fe71e5290b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.335767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:07.137170Z digest=sha256:11fa1098385a8d2bd15df9905360afe399aa2ec1433e9663f4bc84f0a9c8c47a

Observation 9611d3b2-7769-4130-9ec5-359f6276a384 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.192945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.192945Z digest=sha256:4a3ccac6516665ff345d973e64a61b1fb74c65234bfabffe9205b7569e551d63

Observation ef17a6f6-5b3b-4eed-a1fd-b7145d65b926 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.239818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.239818Z digest=sha256:8d5f6b18e2ceb2d8736e5b9a6e7edca81d3f58dfb14dfca5958c4284f5bec671

Observation 4200f296-1d42-43ed-b819-ecd860d9a56a · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.367992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.367992Z digest=sha256:655c4e9970799d643a6ffe39c15967c2c4539e926acd18c71c6d0c79f29fdf63

Observation 70d7e93e-960b-4341-aa21-9ae7c69c9739 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.450675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.450675Z digest=sha256:1cb081f3fec683f7d2e68742e05afeefce763a0b281f9be94eda88f8544ecc45

Observation 9e81d721-8450-46f8-8ee6-f1a21ce056e0 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training LLaVA-OneVision: Easy Visual Task Transfer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.629431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.629431Z digest=sha256:473db43634b4c9bffb4cc7fa7e3f741fddf3c08a99eacd52774dfec868710e01

Observation 9c729d69-5cf2-4e11-a5d6-6a17c1349d08 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.305056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:07.745701Z digest=sha256:eb28ed050bceccd24d47408a1f63c0af7d2abd8204d33604b519624c4ac3507a

Observation 83e6d246-856a-459c-a113-a803222f8168 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.293319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:07.841615Z digest=sha256:e56330152c29780ae3608b3c67ca163ec6b055eb148c59b9261241c04ac5ba49

Observation 967b6b02-e4b3-42d3-82d9-8529ca70ce21 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.958818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.958818Z digest=sha256:b6e5948cbd4ebf969cda420a835d9543d58da802bebc00dc2b9393f2559900e9

Observation 7a6f3d12-8ea7-4a62-a844-66cdb26152bb · outbound

This paper cites Decoupled Weight Decay Regularization.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Decoupled Weight Decay Regularization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.150028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.150028Z digest=sha256:c3d1bb37ada3e1c4abf0a7aea4cc7db77b19092095a6a616628243c0a69fd066

Observation e30bebd7-02f5-4476-921a-def0d6d8376b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.786336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:08.282021Z digest=sha256:dc80d2b19bda8fe4423a3e04a7444aaeee75d26057585fa0152b6a3bc06efd87

Observation cb39c92b-bd58-4f6f-ad10-092f0a1893ef · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.369036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.369036Z digest=sha256:089a2a64ae05966780a1622f22e26b24100f70d44b78b8e34da966071ae90610

Observation 37c58238-42ff-4bf9-953b-4bed40a2e803 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.500385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.500385Z digest=sha256:897c5e90ef4fe4216fd8fee01f1101339a973469a2c579ba176d03228f66438c

Observation 1b3d9346-d2f6-4cd7-b211-1a9415148ae3 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.277858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:08.066644Z digest=sha256:9d827c395c1fe326c1e3fbb667fbc36a4e37ef4b2c3ad29eadef9e4707d5a900

Observation a1931f7f-fb2e-4e03-a7c1-f53aa8c2a5ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.257452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:08.744453Z digest=sha256:0c521f932a3fdca1116e35bc54e2f1294f36b713c4518a1cade727df34bf7a27

Observation c7509f94-6061-400e-a3ab-c688ac59ffe7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.246457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:08.879157Z digest=sha256:d46a02659685d1c5ebfccaa2267dd855c98c68528c57653209acfbe2c6e1afc6

Observation e50fe659-5e26-497c-971d-a732af370509 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.036940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.036940Z digest=sha256:d2af07bf62ce8f27868eeb9c1d3c56201741556d36acff46a4f7c013ae6cb470

Observation aed89379-dbe6-4ca4-84f2-23e2fba0b7cc · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.229217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.171661Z digest=sha256:107446e91522af72918dad2dd5c8ab9549ca47cae87d4101ae6e648197fe1757

Observation 950af699-33d2-4395-a2a1-eb8434b832ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.268243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:08.634729Z digest=sha256:8ec8d340979f1d50fdaaee46a6878b7daae79528aaab9692b676c8b44c20e9ad

Observation c6c5c237-94e7-4014-9c08-2d56396cf76f · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.213411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.397609Z digest=sha256:19838ef87c65bcb49e7671d6927c1c46e4a8a3b74a3ca020734dbb117894d659

Observation a0647a3f-c126-49dc-baa4-efa1791696ea · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.204296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.470703Z digest=sha256:2f375286c9d5bd10fa9f9c754f780e5f52cf455e97b5fa882061248c79a87e99

Observation d28d1515-0ce5-4446-99ca-2500ef15e566 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.106579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.527545Z digest=sha256:81e10a7be2b893490c5d5d2d6b4912e1034b7fcc54779ced06fe5c043ab24995

Observation eff0526c-e9e6-4ba3-991b-9c7ef0626f98 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 39

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.616681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.564187Z digest=sha256:ed3f9a29135b88c0b2abd8527f36c04609a752b294cfa5c076b6f507986c1016

Observation 21685015-2d51-4f6e-827e-06f166f1b562 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.313558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.313558Z digest=sha256:e2c71e95330025189f71b02c37450aa839845908d9916a54780322d83f477dc8

Observation f6e54fc4-3801-41d8-90dd-2a439f743eb6 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.183491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.723169Z digest=sha256:a8f54d6b8e19a7172c4f729a472061c1a0e1aacafd34b7c3cebb9bd7b3204af4

Observation 69ecbad6-15f3-4c3e-b905-19afa07a3021 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.174336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.753390Z digest=sha256:cf078b96a6f66020bb22eafb8615fa8c7a08f261e321e951e9777466bf96d858

Observation 550cf1ac-bb23-4821-b51c-c87adf33cdb1 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.164644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.841458Z digest=sha256:afa8505fa70f1cccff792da7b00ac23cca8d7e416b89fd23a0691ab193f73e8a

Observation 67a6ee80-0852-4adb-9752-93e8c9d8ef55 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 44

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.459447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.925918Z digest=sha256:d36b1ca47d5a7b9c24cc07de2e0e68090a5e39fb84ee49d6ae63b2dd27d0b82c

Observation 22cc9f4b-dd66-4b1d-9dc2-f81a0c8f42b5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.193431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:09.652357Z digest=sha256:ab69356acb4f811536f0477757f18c0920dec3389f12c948ed25a98a95b8b6db

Observation aa3a5a36-b428-4ce8-b9ae-39384141ac51 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.154808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.171917Z digest=sha256:3086caec3d2bea2afdc534dbdb31ecf0db8d97e4e2991eb6dbb40d9e053eb0b7

Observation 4ac61163-a401-4056-a125-e57f5f5d7156 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.252632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.252632Z digest=sha256:3ab4645307bc8bb04c851f976464dbea5ba80a85ce2331bd74ece32f5fc8b33f

Observation 0de8c958-c5e2-4b57-982b-ef280b189337 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.334837Z digest=sha256:138d1c89c426b41f8e1e541d88291338930dc10a153a8228cd4f26cb8988f252

Observation 8768b146-5d26-4721-b7fb-5b8611c43ba7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.134274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.411765Z digest=sha256:fba92326e61fdbd3579cda9004e535bcb1e67d9f4762eaf85f120b6f8375ada4

Observation 1171e4f6-1351-435f-ae4e-ac3f5ac56389 · outbound

This paper cites CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.032757Z digest=sha256:75ac5a7c304697d739402877711dba909cb82c0c04d515634dec9e29e4702289

Observation f4053ef7-be47-4908-be0b-255e067d5932 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.113630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.600580Z digest=sha256:b1af67f5f8d05a9de1817da3a648281fc57bf5ecb9709b29b378ca3d6f76713b

Observation 0fa8b4b3-86be-4b42-b828-3ba73d142268 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 52

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.289869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.642109Z digest=sha256:8d38ce0a0cf66bad9dfcb1042bc02914034bd087dd6c5f9219a7851f9dcb3e11

Observation 06a8dfcf-bfd4-46a4-8af0-f84c57d4259c · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.727452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.727452Z digest=sha256:8f9a2caf70bb6f6fb3954b5b049832f38e0ed6dbc5601c4664c86812c58dd11c

Observation b942b722-472b-4e67-bf01-648e65d2b7a8 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.095957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.821908Z digest=sha256:f4b86db357841350d38d64d11c6437f9b1c9625cf82b33590b7aa2edccadc50e

Observation e0b9dcc3-b21a-44a0-a7ba-60578e7b11f5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.124121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.511071Z digest=sha256:9daa348689345142f9c2eef00dec72ddc1b331ef63d572a20f37fc7ccb1548d8

Observation 64682249-b972-4a9f-89aa-cf6e4eb7dae2 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:11.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:11.003082Z digest=sha256:26ad454d96bff2adcc0c636ce16f72bbc74e732279e19c858b570eaa9a541bbe

Observation bf2b3cec-c6b8-4b10-b4dd-de27c27bcd7b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:12.967277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:11.094770Z digest=sha256:8a8b27bc4358ea331b2418641af347ef32a1a986f6b942e52f4b7c50ada9158c

Observation d2f5310c-e3d0-4be9-801b-ded63e567e65 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.083669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:10.908903Z digest=sha256:83c3bf9a65243dfb7358b6aeced2f40bd253f450fcb6e14914f22932179d9357

Observation 8a192271-f78a-483e-b291-d47b09b5462c · outbound

This paper cites Pattern Analysis and Applications 25, 4 (2022), 981–992.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Pattern Analysis and Applications 25, 4 (2022), 981–992

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.380191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.573227Z digest=sha256:8fea9aabf9fa7ac996b3c8f12372a9db2833ce5690b80c6bed6d5339e1db042c

Observation 413f7d4f-1e0c-4756-a4dd-e455d85535a7 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.423278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.137860Z digest=sha256:1ccce26be9553535426071c6a3029e9d8f253ef08b5bab46217e676a783496ab

Observation 1b457127-83af-4786-8a86-36ed0d4483e1 · outbound

This paper cites In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT).

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT)

Reference 2024

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.712624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:06.880564Z digest=sha256:7b2847dda40035488739f04693e6e2a0c7c14db3032f6f18519ed46df7854089

Observation e5348dba-3fe9-4ea4-a7fc-2ebc60caace8 · outbound

This paper cites StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:22:12.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-06T11:22:07.063747Z digest=sha256:5c91ca2146f4e17afeea523266ecc6a59f551568fbd44a8be6b518c0a07caf62

Pith citing papers

No inbound Pith citation observations are available.