Pith. sign in

Paper Citation Record · LEDGER

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration

As of 9 August 2026, this Paper Citation Record lists 54 of 54 outbound references and 0 inbound Pith citation observations for arXiv:2607.04472.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2607.04472 v1

Coverage vector

measured 54 of 54 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-07-11T18:55:21.867598Z

measured 54 of 54 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

54 of 54 outbound references displayed

  • verified exact10
  • verified fuzzy0
  • unresolved35
  • parse uncertain0
  • malformed identifier9
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 33d20f8d-bbea-4610-80a3-5d3d38e313a6 · outbound

This paper cites 2025.Drifting A way from Truth: GenAI-Driven News Diversity Challenges LVLM- Based Misinformation Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Drifting A way from Truth: GenAI-Driven News Diversity Challenges LVLM- Based Misinformation Detection

Reference 1

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0a1c9c70757381db638674e5047517c8af4083e77b1833d8e5737c313baa278d

Observation c3e63054-ec9b-422c-912f-4caf3af4b213 · outbound

This paper cites 2025.Zooming In on Fakes: A Novel Dataset for Localized AI- Generated Image Detection with Forgery Amplification Approach.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Zooming In on Fakes: A Novel Dataset for Localized AI- Generated Image Detection with Forgery Amplification Approach

Reference 2

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:6606609922f6ca583c030d42c7ffa42773a83accdd9402e28fa04b73d0916891

Observation bb644b91-cb7c-41b8-b9e1-be37bca4f3f0 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 3

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.230423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:4e8360432f82c61b46485b0ee3f53eba27dcd87d7fa4d421ceb2e92006603779

Observation 87ab28e0-b3e6-4d3d-99db-88a588ca0534 · outbound

This paper cites 2015.Going deeper with convolutions.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2015.Going deeper with convolutions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:d8fe8bcacd881f7e3334d45962c81e241b5e792f0f0385317d76ff1e50fbf13a

Observation 3adb4922-09c0-479b-987d-2ce47a1f995b · outbound

This paper cites 2018.Cascade R-CNN: Delving into High Quality Object Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.Cascade R-CNN: Delving into High Quality Object Detection

Reference 5

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e633bfb1162a54bf308938f8950db8790a84e2be487cb94c0281c8fa1e3f7695

Observation f70e452c-64b6-4813-9ab2-fc3f2be61ab3 · outbound

This paper cites 2022.Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Do You Really Mean That? Content Driven Audio-Visual Deepfake Dataset and Multimodal Method for Temporal Forgery Localization

Reference 6

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:ed737d614530d75813079fb4d5cd825070151286c86da6bc5a2ba6733a3154e3

Observation e9c224e1-a57e-4d82-ad7c-75e6859257c0 · outbound

This paper cites 2016.Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2016.Faster R-CNN: Towards Real-Time Object Detection with Region Proposal Networks

Reference 7

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:3bb5d43a737659c2df4d908e8e63c5fe476e1f8f4ef885a80c7785687c5a0200

Observation c46fd08f-133a-4261-9db6-5a34feb485fc · outbound

This paper cites Rojas, Ali Thabet, Bernard Ghanem.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Rojas, Ali Thabet, Bernard Ghanem

Reference 8

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:7a4e475a1c52c0c8bedd5d75d47373fbc54ba86e636f859b1714dd6f9fa1252e

Observation d68ba9a9-4716-4ce8-9efe-1c29659c2906 · outbound

This paper cites Exposing DeepFake Videos By Detecting Face Warping Artifacts.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Exposing DeepFake Videos By Detecting Face Warping Artifacts

Reference 9

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:6be6fa3618b0320f95d571e5e24719822415daab76acd3628fb898184e5c6e2b

Observation be15f5d4-b99e-47c9-86aa-fa498afe596b · outbound

This paper cites Detecting Photoshopped Faces by Scripting Photoshop.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Detecting Photoshopped Faces by Scripting Photoshop

Reference 10

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:185926d0d696e2efb3c2ded24a136ecfe5aafcb1b26156c8b61ab3e162917146

Observation 12cf2aef-8320-4950-983a-254336ecfe52 · outbound

This paper cites 2020.DeepFake Detection via Facial Landmark Analysis.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.DeepFake Detection via Facial Landmark Analysis

Reference 11

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.332483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:1d193af820915ce94825c0c4304d7a7db064042f535f1b839df2f6554768269a

Observation 6930d8a9-9a23-4c8f-aa20-1ddab8330d75 · outbound

This paper cites 2020.FakeCatcher: Detection of Syn- thetic Portrait Videos using Biological Signals.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.FakeCatcher: Detection of Syn- thetic Portrait Videos using Biological Signals

Reference 12

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:aa314033fe67a920af2f384abecf6ff12300c855fcccfbe3200e12188cc0f62d

Observation dddb188e-5706-4a78-b528-084fc20b811d · outbound

This paper cites 2019.FaceForensics++: Learning to Detect Manipulated Facial Images.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2019.FaceForensics++: Learning to Detect Manipulated Facial Images

Reference 13

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c2c16f1b5578319fed753c8c17c44c40a9c70c7dd211327df9bda4f2ab7b800e

Observation b0c6b79d-de08-4416-baee-46078164b3be · outbound

This paper cites 2018.MesoNet: a Compact Facial Video Forgery Detection Network.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.MesoNet: a Compact Facial Video Forgery Detection Network

Reference 14

Resolution
malformed identifier
doi_truncated, observed 2026-07-11T18:58:11.340052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c35477277c79d178e238747f7e61a9bb99a16d68db241e217d958631b654dc22

Observation 6d57ef51-6bcc-4c49-a051-13981d8aa660 · outbound

This paper cites 2023.Dis- criminative Feature Mining Based on Frequency Information and Metric Learning for Face Forgery Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Dis- criminative Feature Mining Based on Frequency Information and Metric Learning for Face Forgery Detection

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.216298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:40ecdceeb36af39f97b0cacfd6cd0ae2c2eec5cff7102386cb39d0d0586ec929

Observation b519d600-af5f-4508-a642-853e578bd6ca · outbound

This paper cites 2022.Adaptive Face Forgery Detection in Cross Domain.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Adaptive Face Forgery Detection in Cross Domain

Reference 17

Resolution
verified exact
doi, observed 2026-07-11T18:58:11.330101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:2d75e6e3cef5efde5788cc40ca957b64cc8e7a77023a44a5c201da2ee8ef1889

Observation 88f8aafe-f030-4286-9b2e-30edb6810dd5 · outbound

This paper cites 2021.Learning Self-Consistency for Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.Learning Self-Consistency for Deepfake Detection

Reference 18

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:cdfaf0452d040b18759cad200b5508422af4af5cfcd5581c0bc7a6165e082829

Observation 3bcae84b-aa1a-44ba-99bd-1c0009ce260d · outbound

This paper cites 2023.Deep Learning-Based Action Detection in Untrimmed Videos: A Survey.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Deep Learning-Based Action Detection in Untrimmed Videos: A Survey

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.282948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:73561d406e4cc70938cbaf01e9a2fa016acba4ff2d15e740eb5c711ec2c232e9

Observation 788e5cd4-4d47-4a4b-8f99-8d9bbc2cb28f · outbound

This paper cites 2023.PivoTAL: Prior-Driven Supervision for Weakly- Supervised Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.PivoTAL: Prior-Driven Supervision for Weakly- Supervised Temporal Action Localization

Reference 21

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:512754cb13ab522e128aa0e82f730d4b7e68aeabe3082ac9d9f857265f133b59

Observation 6e231910-19e1-4ec5-bb16-7f17d12630d4 · outbound

This paper cites 2024.Blind and Low Vision Individuals’ Detec- tion of Audio Deepfakes.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.Blind and Low Vision Individuals’ Detec- tion of Audio Deepfakes

Reference 22

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:24e9714f00eb83ba52901021aa8e3a1ac1c8ad2975f833c14a2da05aabf6ff1b

Observation 021dbd0e-3cd1-48c3-aba5-5d65fb18fd58 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 23

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:3d74e0cb9e39f31075cb1879922cb88aba6e62a8cb0e19319df07dcf6a80d9f0

Observation 39ae174d-d740-40ec-924a-d91c2e544339 · outbound

This paper cites 2021.An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.An Image Is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 24

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c4cfc17d6cab118bfc48d1ac0a9fc0937e2e5ae6782bd690f1266cbe6f57d6c7

Observation 0d3165fc-7253-4019-9478-8af76b40d921 · outbound

This paper cites 2024.End-to- End Temporal Action Detection with 1B Parameters Across 1000 Frames.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.End-to- End Temporal Action Detection with 1B Parameters Across 1000 Frames

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:068abd4716b2ad378207aac38b2e782cfd81ce89e4cdf57f0d7b0d5118ec5dac

Observation 35667870-73d3-48a3-b0f0-1da29570bdf0 · outbound

This paper cites 2021.BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.BYOL for Audio: Self-Supervised Learning for General-Purpose Audio Representation

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:2ee8645168e1fa35286fb798090a3976e46ec387f3bd0bc9d1ceed9e470f7ca1

Observation 9dd1b3d6-814c-4489-b11e-7c4b468d898a · outbound

This paper cites Weinberger.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Weinberger

Reference 27

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e0e2a8f6358c9f502d980edd17db4c2da22d11bcfa78585f8a2ed9a01023177b

Observation 88c41a1e-67dc-4cba-bcf8-936df28417e6 · outbound

This paper cites 2021.Cross- Attentional Audio-Visual Fusion for Weakly-Supervised Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.Cross- Attentional Audio-Visual Fusion for Weakly-Supervised Action Localization

Reference 28

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:f3977e35faa5b7d82ddae7cb247946fbfafcb9f708d58a00e92c5f6862cd1d8d

Observation ebfa1847-6807-4a3a-9bae-08b53016159a · outbound

This paper cites 2024.MLCA-A VSR: Multi-Layer Cross Attention Fusion Based Audio-Visual Speech Recognition.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.MLCA-A VSR: Multi-Layer Cross Attention Fusion Based Audio-Visual Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:cf75d90c78bfbf217f945035d6e2859873e5f84e898d978339aba243e8ed145a

Observation 15bf82fb-ac2c-4377-acf5-ef1a6ac7d2dc · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 30

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.261026Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:dfb37f71f655addfd28888c66d1390c4840bb93b510e67682f52978f2b846272

Observation a93729f1-c023-4fb0-af58-a81d251471e6 · outbound

This paper cites Activity Graph Transformer for Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Activity Graph Transformer for Temporal Action Localization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:1a88199a9adda1269e525c85304c2972f7d4d2faebe0fabc2df434702dc73a14

Observation a4690baf-f505-454d-9227-1485eb4a8dbb · outbound

This paper cites 2019.BMN: Boundary- Matching Network for Temporal Action Proposal Generation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2019.BMN: Boundary- Matching Network for Temporal Action Proposal Generation

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:41850bda12238b7fb1614063350a54500be59bc9d345feffb059f0ea19069b83

Observation e41b3593-35a3-4b72-a444-6de7b14357d7 · outbound

This paper cites Hear Me Out: Fusional Approaches for Audio Augmented Temporal Action Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Hear Me Out: Fusional Approaches for Audio Augmented Temporal Action Localization

Reference 33

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c45afc4229dcaf3512ce837322aa325299b015916c36181639ca1a6e9ffc98aa

Observation 7ab3e9d9-8097-49a8-8939-c78f6e2d82fd · outbound

This paper cites 2021.SOFT: Softmax-free Transformer with Linear Complexity.https://proceedings.neurips.cc/paper_files/paper/2021/file/ b1d10e7bafa4421218a51b1e1f1b0ba2-Paper.pdf.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2021.SOFT: Softmax-free Transformer with Linear Complexity.https://proceedings.neurips.cc/paper_files/paper/2021/file/ b1d10e7bafa4421218a51b1e1f1b0ba2-Paper.pdf

Reference 34

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e81776378dac8704fbaf5baa798ea32648b65fe5fa5e7ff92956f32966d3a41f

Observation 82796cb5-d708-49da-9a60-7d4d7b1fb03f · outbound

This paper cites 2022.ActionFormer: Localizing Moments of Actions with Transformers.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.ActionFormer: Localizing Moments of Actions with Transformers

Reference 35

Resolution
verified exact
doi, observed 2026-07-11T18:58:11.235015Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:d13255ccf086a16bc08cb446f26e6f68a10341ed3d2f33f9ace383ad8709cb2b

Observation 9473a0fc-d8a2-45aa-a4e9-49fc6af0a029 · outbound

This paper cites 2023.Ummaformer: A Universal Multimodal-Adaptive Transformer Framework for Temporal Forgery Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Ummaformer: A Universal Multimodal-Adaptive Transformer Framework for Temporal Forgery Localization

Reference 36

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:54150cc6d8cad8e70f7c9a4c04f0b89f11419ae3d57530cc33f2ab34b0833d71

Observation 37442a88-3105-480f-9aa7-a64e3f0ae0ed · outbound

This paper cites DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration DiMoDif: Discourse Modality-information Differentiation for Audio-visual Deepfake Detection and Localization

Reference 37

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:4358383b2679564c0d2fd956c8f486b81800b812c26156c97099f5db3f466821

Observation 2e2f0860-e15e-4a3e-a80d-c0b22a3616fd · outbound

This paper cites 2018.Mesonet: a compact facial video forgery detection network.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2018.Mesonet: a compact facial video forgery detection network

Reference 38

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:86492c238e65445ce13ac0c891763e514e7bebdad74041b087fc509a0ff43e67

Observation ed8c428b-127a-4ea4-bc89-1141ab7f9c61 · outbound

This paper cites 2023.Glitch in the matrix: A large scale benchmark for content driven audio-visual forgery detection and localization.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Glitch in the matrix: A large scale benchmark for content driven audio-visual forgery detection and localization

Reference 39

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.243035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:b4e4937da6b331d071be83dca809e4b4f7d3f89316882e28f26e7dc830fa4acd

Observation 6e393692-de00-4337-af65-61fa7b6bfaff · outbound

This paper cites 2023.Tridet: Temporal Action Detection with Relative Boundary Modeling.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2023.Tridet: Temporal Action Detection with Relative Boundary Modeling

Reference 40

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:5a68597a79304250a5df29217659a864736d17d059365f0d9f8bdae529f606d2

Observation ebbebc9d-117b-4eb9-a4c6-c760c820041b · outbound

This paper cites Zhang et al.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Zhang et al

Reference 41

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:19cd9221663e8b2bbafb7e9aab77ecbf53d9c5c1ce60c7f6c02c4223900dd606

Observation fff92487-6a1c-45d4-a746-5c07a06c878e · outbound

This paper cites 2017.Attention Is All You Need.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2017.Attention Is All You Need

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:4b96f6d24f0cfdb15816f5a129ddc3aa6e080152ba0dbcf891ff26d4ba3d8edc

Observation 76151fdd-2049-495f-98a8-c3618ab35f7d · outbound

This paper cites 2024.MetaFormer Baselines for Vision.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.MetaFormer Baselines for Vision

Reference 43

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:94a6a220c63888595b476eb4bf04e404e41fd956cee318553799e259621d407a

Observation 3f5ba93c-7d89-4714-97c0-9c13fc1f9bbf · outbound

This paper cites 2020.Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2020.Distance-IoU Loss: Faster and Better Learning for Bounding Box Regression

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:7845107a48e8810fa4570731ab073924d393bcda1dfcc0bd5e687e3cb380697a

Observation 5c75c3d5-4813-4835-835c-87eaff1c070a · outbound

This paper cites 2017.Focal Loss for Dense Object Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2017.Focal Loss for Dense Object Detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:af500af6717ef5e8a664efad13481f7431db02bf0195ec04b03dfaa8d44fd6a4

Observation 8189c48c-bac5-432c-9517-41d9982a4231 · outbound

This paper cites 2025.Face Forgery Video Detection via Temporal Forgery Cue Unraveling.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Face Forgery Video Detection via Temporal Forgery Cue Unraveling

Reference 46

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:d8bc2fa17fe846cea6ec78e1b6bf4341f66426a08362b0431e019af585c866b2

Observation 6075bbb9-e5f9-460c-b486-5a13bf6c055f · outbound

This paper cites 2024.A VFF: Audio-Visual Feature Fusion for Video Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.A VFF: Audio-Visual Feature Fusion for Video Deepfake Detection

Reference 47

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:810357c0d334f7ea9d6b0a73ddfeb950d73b8ee4734d7f1b3fbedd87741579bf

Observation 9664f815-11c6-4108-9e80-1fb5160e6ec3 · outbound

This paper cites 2024.Delocate: Detection and Localization for Deepfake Videos with Randomly-Located Tampered Traces.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.Delocate: Detection and Localization for Deepfake Videos with Randomly-Located Tampered Traces

Reference 48

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:76ad7b2dcdd0daeadaeb208c0ddf20a2a0fae61413862aefcb249cdaa897c14c

Observation 9b92cbf2-f075-4e26-8eb2-d496aa908954 · outbound

This paper cites 2025.Trusted Video Inpainting Localization via Deep Attentive Noise Learning.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Trusted Video Inpainting Localization via Deep Attentive Noise Learning

Reference 49

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.254250Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:0ddab76ab28653f29b22cededfea311279c6bccc43a362b7b69283286b5c32de

Observation ec8d9b4d-aea8-44e9-8c94-9bd53255974f · outbound

This paper cites 2025.Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2025.Bridge the Gap: From Weak to Full Supervision for Temporal Action Localization with PseudoFormer

Reference 50

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:e7c79c60810d8e8fe63a665e996ef1b841a09416ba0cb0f6db95d5e61ba2643c

Observation d9cef25f-7d7f-4a6c-b001-ba53856cd170 · outbound

This paper cites an unresolved cited work.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration Unresolved cited work

Reference 51

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:a008f9679570a506652f1767a91cf13944e5e9b2694af1cbf44b0db3d4c4c01a

Observation a3afda54-d727-4e6f-a610-81d334a4b627 · outbound

This paper cites 2024.SafeEar: Content Privacy-Preserving Audio Deepfake Detection.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.SafeEar: Content Privacy-Preserving Audio Deepfake Detection

Reference 52

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:cfb271548ec6fde185ddf905d203ee5ecf063e7846a143fc51d46b54ea8f1c35

Observation b2adb45a-2029-4eec-9dfa-dc971d4f2625 · outbound

This paper cites 2022.Localizing Fake Segments in Speech.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Localizing Fake Segments in Speech

Reference 53

Resolution
unresolved
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:c97e7e4ae9ecb6786f8568e7845b744ca5bb740bfc766433de29d6a362a0bf4a

Observation 80e034b7-fa8b-4b7c-8c1f-d8781d5a88bb · outbound

This paper cites 2022.Proposal-Free Temporal Action Detection via Global Segmentation Mask Learning.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.Proposal-Free Temporal Action Detection via Global Segmentation Mask Learning

Reference 54

Resolution
malformed identifier
doi_truncated, observed 2026-07-11T18:58:11.224232Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:8122b06881ee9cfa4cd58500d1c31e2fe5ec70a262f7c05f2ca9a10c317ad0da

Observation bef61c61-5bde-41ab-b8ab-445df0dde0c6 · outbound

This paper cites 2022.DCAN: Improving Temporal Action Detection via Dual Context Aggregation.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2022.DCAN: Improving Temporal Action Detection via Dual Context Aggregation

Reference 55

Resolution
malformed identifier
no resolver link, observed 2026-07-11T18:55:21.867598Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:287b326c1fe262ce35ee1b314ce43c74f15bb411cce9eb3f011bafd5e435364e

Observation d0979a33-24b1-430b-a11f-7b8ed3c479ce · outbound

This paper cites 2024.A V-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset.

EVAS: Efficient Multimodal Temporal Forgery Localization via Audio-Visual Synergy and Steered Boundary Calibration 2024.A V-Deepfake1M: A Large-Scale LLM-Driven Audio-Visual Deepfake Dataset

Reference 56

Resolution
verified exact
arxiv_id, observed 2026-07-11T18:58:11.236002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-07-11T18:55:21.867598Z digest=sha256:d72964a9c10133c7b7528ba98eddacfda5d8c27867706311b9c4ca4474c383e3

Pith citing papers

No inbound Pith citation observations are available.