Pith. sign in

Paper Citation Record · LEDGER

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration

As of 8 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2607.04472.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.04472 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T18:55:21.867598Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact10
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier9
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 33d20f8d-bbea-4610-80a3-5d3d38e313a6 · outbound

This paper cites 2025.Drifting A way from Truth: GenAI-Driven News Diversity Challenges LVLM- Based Misinformation Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Drifting A way from Truth: GenAI-Driven News Diversity Challenges LVLM- Based Misinformation Detection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:5489bc2671c9568a297388056360c5967c657f7f368e598748c7c1f89f991820

Observation c3e63054-ec9b-422c-912f-4caf3af4b213 · outbound

This paper cites 2025.Zooming In on Fakes: A Novel Dataset for Localized AI- Generated Image Detection with Forgery Amplification Approach.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Zooming In on Fakes: A Novel Dataset for Localized AI- Generated Image Detection with Forgery Amplification Approach

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:7d25955f90b28782808211d510117631378c2f5b87941a63d46dd6e3d81b5892

Observation bb644b91-cb7c-41b8-b9e1-be37bca4f3f0 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.230423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:1a809737a2a6e7314719b115d6930c68d9fea0bcabd82e47b0ef8b11f79bc4e7

Observation 87ab28e0-b3e6-4d3d-99db-88a588ca0534 · outbound

This paper cites 2015.Going deeper with convolutions.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2015.Going deeper with convolutions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0ef4930ead1fc778190ffe2282014d9518cb9f5576c9c98b32d265067319e075

Observation 3adb4922-09c0-479b-987d-2ce47a1f995b · outbound

This paper cites 2018.Cascade R-CNN: Delving into High Quality Object Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.Cascade R-CNN: Delving into High Quality Object Detection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:ef9a6a5eabd1005618d66837323094eae6e62e27a23b61717856cd788f0458de

Observation f70e452c-64b6-4813-9ab2-fc3f2be61ab3 · outbound

This paper cites 2022.Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:8e6561fc19aeecbe9a7f41460ececcb4c1f323e086d1942decf6ead31aec4712

Observation e9c224e1-a57e-4d82-ad7c-75e6859257c0 · outbound

This paper cites 2016.Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2016.Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

Reference 7

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e977da0efaf442ea702042c2c85ef1de2ff294b4f79e76fca5527802ec86cd31

Observation c46fd08f-133a-4261-9db6-5a34feb485fc · outbound

This paper cites Rojas, Ali Thabet, Bernard Ghanem.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Rojas, Ali Thabet, Bernard Ghanem

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:153f2be623c6b7683e000a0ef80f3f15c688a0e572be5e316ecb993fca6e28e0

Observation d68ba9a9-4716-4ce8-9efe-1c29659c2906 · outbound

This paper cites Exposing DeepFake Videos By Detecting Face Warping Artifacts.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Exposing DeepFake Videos By Detecting Face Warping Artifacts

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:32f76917b0919a29f746602e04bc2af026ecd04da2e86cc2c008d045b82c45eb

Observation be15f5d4-b99e-47c9-86aa-fa498afe596b · outbound

This paper cites Detecting Photoshopped Faces by Scripting Photoshop.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Detecting Photoshopped Faces by Scripting Photoshop

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:32bfaed69dedeb6d7586bfb17b9e782478c9f40295e63ef32c4c25277175e43e

Observation 12cf2aef-8320-4950-983a-254336ecfe52 · outbound

This paper cites 2020.DeepFake Detection via Facial Landmark Analysis.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.DeepFake Detection via Facial Landmark Analysis

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.332483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:01b5c970d02609740fd31f8a3b7deecd53930956856c1d65df6ca44051d8da05

Observation 6930d8a9-9a23-4c8f-aa20-1ddab8330d75 · outbound

This paper cites 2020.FakeCatcher: Detection of Syn- thetic Portrait Videos using Biological Signals.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.FakeCatcher: Detection of Syn- thetic Portrait Videos using Biological Signals

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:38551888e39ef990f17e61af0ce7aebf4418a10a2d2439922d62379e65da0c49

Observation dddb188e-5706-4a78-b528-084fc20b811d · outbound

This paper cites 2019.FaceForensics++: Learning to Detect Manipulated Facial Images.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2019.FaceForensics++: Learning to Detect Manipulated Facial Images

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:730012d1b519275875ed4a3b0158ccc6703cdf2376678bc2a00258cd80a09967

Observation b0c6b79d-de08-4416-baee-46078164b3be · outbound

This paper cites 2018.MesoNet: a Compact Facial Video Forgery Detection Network.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.MesoNet: a Compact Facial Video Forgery Detection Network

Reference 14

Resolution
malformed identifier
doi_truncated, observed 2026-07-11T18:58:11.340052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:164700edf5cfe1cde09b37a825bf06572051150d40aa861c80dd4333daf02848

Observation 6d57ef51-6bcc-4c49-a051-13981d8aa660 · outbound

This paper cites 2023.Dis- criminative Feature Mining Based on Frequency Information and Metric Learning for Face Forgery Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Dis- criminative Feature Mining Based on Frequency Information and Metric Learning for Face Forgery Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.216298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:1b9777013bb72dbd007fde9ff0d37e15fbd047e274d41d1049984c19131dc046

Observation b519d600-af5f-4508-a642-853e578bd6ca · outbound

This paper cites 2022.Adaptive Face Forgery Detection in Cross Domain.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Adaptive Face Forgery Detection in Cross Domain

Reference 17

Resolution
verified exact
doi, observed 2026-07-11T18:58:11.330101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:54ad98e0345babb9570930d8c37ef8d60dc273a36c8c37eef03d901852456484

Observation 88f8aafe-f030-4286-9b2e-30edb6810dd5 · outbound

This paper cites 2021.Learning Self-Consistency for Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.Learning Self-Consistency for Deepfake Detection

Reference 18

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:ee128a61ea60687d42f665ffb073f87d61340276c7c1c3ee872f7bf86406d9de

Observation 3bcae84b-aa1a-44ba-99bd-1c0009ce260d · outbound

This paper cites 2023.Deep Learning-Based Action Detection in Untrimmed Videos: A Survey.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Deep Learning-Based Action Detection in Untrimmed Videos: A Survey

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.282948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:704bf12433ec2f8edcded54e9faa25a5a9e1d2a5fadf5c08b293b90f87e5940e

Observation 788e5cd4-4d47-4a4b-8f99-8d9bbc2cb28f · outbound

This paper cites 2023.PivoTAL: Prior-Driven Supervision for Weakly- Supervised Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.PivoTAL: Prior-Driven Supervision for Weakly- Supervised Temporal Action Localization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:a99ad21f43845b947ea67ec49e54b97272c943bc866382832560f937b28f5b7a

Observation 6e231910-19e1-4ec5-bb16-7f17d12630d4 · outbound

This paper cites 2024.Blind and Low Vision Individuals’ Detec- tion of Audio Deepfakes.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.Blind and Low Vision Individuals’ Detec- tion of Audio Deepfakes

Reference 22

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0d5fa867aab7ab16935260721b2570a0ae7c657b3add614f734c81cadd36b6dd

Observation 021dbd0e-3cd1-48c3-aba5-5d65fb18fd58 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:392e908319c03228ea65a5fbc86909797510bfe1910be265a52824390825f533

Observation 39ae174d-d740-40ec-924a-d91c2e544339 · outbound

This paper cites 2021.An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:653cf50a2161da365d1abc38a0fa29bb2e37d6ec3facb895d786c69ec07cd846

Observation 0d3165fc-7253-4019-9478-8af76b40d921 · outbound

This paper cites 2024.End-to- End Temporal Action Detection with 1B Parameters Across 1000 Frames.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.End-to- End Temporal Action Detection with 1B Parameters Across 1000 Frames

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:f5e70c52c6a08c78e998ccfb236e946c1be9791871879e0e0f00d08c6fb5223f

Observation 35667870-73d3-48a3-b0f0-1da29570bdf0 · outbound

This paper cites 2021.BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:3d57cf4e866f9c62ef2a4df8ede23f1c977d79300f719d8ac50fe6cca5a53ab5

Observation 9dd1b3d6-814c-4489-b11e-7c4b468d898a · outbound

This paper cites Weinberger.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Weinberger

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0e3648ff58e7e50700b5cb503ed1113ecf99c5a53140595addb9ae4b9a40e837

Observation 88c41a1e-67dc-4cba-bcf8-936df28417e6 · outbound

This paper cites 2021.Cross- Attentional Audio-Visual Fusion for Weakly-Supervised Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.Cross- Attentional Audio-Visual Fusion for Weakly-Supervised Action Localization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:29df0508c5a2c591eddb41226c046854be6242eb13703bfead30aead8649c100

Observation ebfa1847-6807-4a3a-9bae-08b53016159a · outbound

This paper cites 2024.MLCA-A VSR: Multi-Layer Cross Attention Fusion Based Audio-Visual Speech Recognition.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.MLCA-A VSR: Multi-Layer Cross Attention Fusion Based Audio-Visual Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:ea3e0867e2f7e5dcf85afde32e52a03f8312ae91c17076d46001b685071ef842

Observation 15bf82fb-ac2c-4377-acf5-ef1a6ac7d2dc · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.261026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e17a03d2d60db6ce8a1d53b11c18091b945e7c8a0bea0d12d11cdac9d4c21416

Observation a93729f1-c023-4fb0-af58-a81d251471e6 · outbound

This paper cites Activity Graph Transformer for Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Activity Graph Transformer for Temporal Action Localization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:07692fbc26876b95d6d5dcad5bd794f0982bc848ef451eeeeb9b970bf42a49eb

Observation a4690baf-f505-454d-9227-1485eb4a8dbb · outbound

This paper cites 2019.BMN: Boundary- Matching Network for Temporal Action Proposal Generation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2019.BMN: Boundary- Matching Network for Temporal Action Proposal Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:5e85e1021f7082e864bab2ea96b56df16ae32e3ba446e74e1995c6212461c489

Observation e41b3593-35a3-4b72-a444-6de7b14357d7 · outbound

This paper cites Hear Me Out: Fusional Approaches for Audio Augmented Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Hear Me Out: Fusional Approaches for Audio Augmented Temporal Action Localization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:51ce1390b0875345aa34a273068934b9154bb0ab3fd9a0902a4252b51620af63

Observation 7ab3e9d9-8097-49a8-8939-c78f6e2d82fd · outbound

This paper cites 2021.SOFT: Softmax-free Transformer with Linear Complexity.https://proceedings.neurips.cc/paper_files/paper/2021/file/ b1d10e7bafa4421218a51b1e1f1b0ba2-Paper.pdf.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.SOFT: Softmax-free Transformer with Linear Complexity.https://proceedings.neurips.cc/paper_files/paper/2021/file/ b1d10e7bafa4421218a51b1e1f1b0ba2-Paper.pdf

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:d03fefd34290b4467205943974aed4ff3f4a22cec3966b54403e3ed0ff0d79d8

Observation 82796cb5-d708-49da-9a60-7d4d7b1fb03f · outbound

This paper cites 2022.ActionFormer: Localizing Moments of Actions with Transformers.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.ActionFormer: Localizing Moments of Actions with Transformers

Reference 35

Resolution
verified exact
doi, observed 2026-07-11T18:58:11.235015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c411bb20c5777bbdb08fb90205310c78cf87efef26ab9fe39c9f11620a6c8822

Observation 9473a0fc-d8a2-45aa-a4e9-49fc6af0a029 · outbound

This paper cites 2023.Ummaformer: A Universal Multimodal-Adaptive Transformer Framework for Temporal Forgery Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Ummaformer: A Universal Multimodal-Adaptive Transformer Framework for Temporal Forgery Localization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:4e5d2726af2af3cff37e86c3dddc54e039b288f03fc1b645e6b7edb802352b9b

Observation 37442a88-3105-480f-9aa7-a64e3f0ae0ed · outbound

This paper cites DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:8fbd46c24888fa0c977d8abe356573bf9ddb8951e729faa784f367616261ac69

Observation 2e2f0860-e15e-4a3e-a80d-c0b22a3616fd · outbound

This paper cites 2018.Mesonet: a compact facial video forgery detection network.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.Mesonet: a compact facial video forgery detection network

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:a8d26ef4b0f665da19e6755868f5cae85f4caec506f5853cf9bd263cb52e7684

Observation ed8c428b-127a-4ea4-bc89-1141ab7f9c61 · outbound

This paper cites 2023.Glitch in the matrix: A large scale benchmark for content driven audio-visual forgery detection and localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Glitch in the matrix: A large scale benchmark for content driven audio-visual forgery detection and localization

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.243035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:a34b466b2b131ac0a5803ed57e3614f07cf22db4ca7e910c58533485210a595e

Observation 6e393692-de00-4337-af65-61fa7b6bfaff · outbound

This paper cites 2023.Tridet: Temporal Action Detection with Relative Boundary Modeling.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Tridet: Temporal Action Detection with Relative Boundary Modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:f3b5de4ba517dabc40180f466cc8fdcae806534b24d61a68d9195d181ab7664d

Observation ebbebc9d-117b-4eb9-a4c6-c760c820041b · outbound

This paper cites Zhang et al.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Zhang et al

Reference 41

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:b3b648cd255775f26187fd84d18bb7fbda912f7ec29ffd94e988aa58af26b258

Observation fff92487-6a1c-45d4-a746-5c07a06c878e · outbound

This paper cites 2017.Attention Is All You Need.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2017.Attention Is All You Need

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:316187a64617817ac85119550e9eb40b340453026a8a7efbb9ca7f24b624074b

Observation 76151fdd-2049-495f-98a8-c3618ab35f7d · outbound

This paper cites 2024.MetaFormer Baselines for Vision.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.MetaFormer Baselines for Vision

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:9e169d3f3698ee638fde28af12048df2d2f8fdeee79b0e5e046e299f1d578dce

Observation 3f5ba93c-7d89-4714-97c0-9c13fc1f9bbf · outbound

This paper cites 2020.Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:90e93a7d446380f7e26e731170b443def914228bbc11e1fc594b5064a990bca3

Observation 5c75c3d5-4813-4835-835c-87eaff1c070a · outbound

This paper cites 2017.Focal Loss for Dense Object Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2017.Focal Loss for Dense Object Detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:13e6f409e2c47d31a9277e282210b9d1367cb51a91d14e212fcac7c074898a8b

Observation 8189c48c-bac5-432c-9517-41d9982a4231 · outbound

This paper cites 2025.Face Forgery Video Detection via Temporal Forgery Cue Unraveling.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Face Forgery Video Detection via Temporal Forgery Cue Unraveling

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:fa3e2e5bd07405be14549ec92fd6cfd432adc5ec44ddd2af7c2a762604b1e606

Observation 6075bbb9-e5f9-460c-b486-5a13bf6c055f · outbound

This paper cites 2024.A VFF: Audio-Visual Feature Fusion for Video Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.A VFF: Audio-Visual Feature Fusion for Video Deepfake Detection

Reference 47

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:5d6d5f7bc2e1180e682bb4384bb941f9d6d4fac85f407fe668e81a89c638d412

Observation 9664f815-11c6-4108-9e80-1fb5160e6ec3 · outbound

This paper cites 2024.Delocate: Detection and Localization for Deepfake Videos with Randomly-Located Tampered Traces.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.Delocate: Detection and Localization for Deepfake Videos with Randomly-Located Tampered Traces

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c852032c0f1879f9cd5c51b79ee591bd6d77135a96801ea88622f2a3f5fd2111

Observation 9b92cbf2-f075-4e26-8eb2-d496aa908954 · outbound

This paper cites 2025.Trusted Video Inpainting Localization via Deep Attentive Noise Learning.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Trusted Video Inpainting Localization via Deep Attentive Noise Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.254250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e161b0776c0b7f8af951a292ab0bb4cb8c768c49b1a808a4f6a64638deaa526c

Observation ec8d9b4d-aea8-44e9-8c94-9bd53255974f · outbound

This paper cites 2025.Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:eddb3bc344664341a22219560a3330f31b1668fd016de7c9536d9a47dc3941a9

Observation d9cef25f-7d7f-4a6c-b001-ba53856cd170 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:ffa666fa10331ac8576175c792fb13c49efff2d65954a8d60df51f86a60ea704

Observation a3afda54-d727-4e6f-a610-81d334a4b627 · outbound

This paper cites 2024.SafeEar: Content Privacy-Preserving Audio Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.SafeEar: Content Privacy-Preserving Audio Deepfake Detection

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0fadd5000ab59636cd8d01758790f7446366b2f7bf5e8d9b97b4d519b5942f14

Observation b2adb45a-2029-4eec-9dfa-dc971d4f2625 · outbound

This paper cites 2022.Localizing Fake Segments in Speech.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Localizing Fake Segments in Speech

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:9bc85731421c8137b710968f67f59d823c11be5f1e4ccf1966f712e72f8fc065

Observation 80e034b7-fa8b-4b7c-8c1f-d8781d5a88bb · outbound

This paper cites 2022.Proposal-Free Temporal Action Detection via Global Segmentation Mask Learning.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Proposal-Free Temporal Action Detection via Global Segmentation Mask Learning

Reference 54

Resolution
malformed identifier
doi_truncated, observed 2026-07-11T18:58:11.224232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:1057f514a941290485bd037abb93f063e62b6af980e8da6c39fb8bde1aec2768

Observation bef61c61-5bde-41ab-b8ab-445df0dde0c6 · outbound

This paper cites 2022.DCAN: Improving Temporal Action Detection via Dual Context Aggregation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.DCAN: Improving Temporal Action Detection via Dual Context Aggregation

Reference 55

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:5527d4f059f3f47123cfb0ed6aa470a027d9c31019bd55e83d638d0a5859afc7

Observation d0979a33-24b1-430b-a11f-7b8ed3c479ce · outbound

This paper cites 2024.A V-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.A V-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.236002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:7110d46d3e8d363f824d622f7961fa0ccd82cbe5cea908edbc84b3a376a88343

Pith citing papers

No inbound Pith citation observations are available.