Pith. sign in

Paper Citation Record · LEDGER

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training

As of 15 August 2026, this Paper Citation Record lists 62 of 62 outbound references and 0 inbound Pith citation observations for arXiv:2507.22781.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.22781 v1

Coverage vector

measured 62 of 62 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T11:22:11.094770Z

measured 62 of 62 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

62 of 62 outbound references displayed

  • verified exact6
  • verified fuzzy3
  • unresolved52
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation effd9037-0481-4087-af1f-a8d58e086193 · outbound

This paper cites Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Phi-3 Technical Report: A Highly Capable Language Model Locally on Your Phone

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.411665Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.411665Z digest=sha256:05ff69789307fe5a723be07d1018e4ba0194f959c1beeb5b820de7e6cc584d9d

Observation 0747430c-e929-4d4b-bce4-1dd373c9ef16 · outbound

This paper cites Qwen2.5-VL Technical Report.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Qwen2.5-VL Technical Report

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.477912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.477912Z digest=sha256:45613b64d7cfcdd47d8204e7d649e35c669d3aefd4e57f71c6ac2937e9eeb8f0

Observation ad31c558-72e1-46b5-bcc5-adb8ac9b5e32 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.573195Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.573195Z digest=sha256:24f03cdbaaeb9ea7dfcea27dcfa49618767488dea6833e81cc3a72a14df654cd

Observation f2d51abc-683e-4a18-ba05-5ca05c79365d · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.477087Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:05.641843Z digest=sha256:8e77c47b2d7c448e93a057802767babab7c42c7e12dc88fdde620373ad99d701

Observation ab54ee9f-25ef-417d-8411-9d7fae8c0eb0 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.467168Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:05.745938Z digest=sha256:cb9ee2131cedc4e8d8ba06e3bf47ad31f220e1d1e002e49e0af38d6d6d5877d6

Observation c85ce3f1-f6e2-463a-8fcd-c8df423c9f91 · outbound

This paper cites AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training AV-Deepfake1M++: A Large-Scale Audio-Visual Deepfake Benchmark with Real-World Perturbations

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:05.863936Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:05.863936Z digest=sha256:d8815de526cec554a70bea2eab46d5480759e76434e8a0df6c7726516209fa25

Observation 6e5fbcac-e2cb-42b6-9d64-27246ed549d7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.456470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:05.896128Z digest=sha256:b7212eaf12e6ebf7e2c7df29666aca5ccd16222c7d489404a6ee23dc2b3835c9

Observation a5491269-5b38-4986-91db-026a580beefb · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.445155Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:05.947834Z digest=sha256:979b78a889844e3d3b9bad96e9d18a7334f9768722d4c0426d6c0e34640a3b5e

Observation be9b7e5e-a7f2-4308-a71a-347824d9cc69 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.433466Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.036307Z digest=sha256:5d6dce8c3abfeab535a3fe0f3dc6ec170f60d79d1172dddbf4e26ae8249b55cf

Observation e1a4e95e-0f2c-42cc-bc94-d6fcf5b3ce6d · outbound

This paper cites ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training ECAPA-TDNN: Emphasized Channel Attention, Propagation and Aggregation in TDNN Based Speaker Verification

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:06.232168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:06.232168Z digest=sha256:eaebd08d3608664265f3c1380f731641aa611fe60b4926879e81ceb874ef6da3

Observation b3ee7575-72e5-4b64-9a61-379af7054532 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.413407Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.325533Z digest=sha256:d8b6765e8c0c6b068ef15eab0c898865cad0e368d07789135bf282a97b2bd4e0

Observation 951a740e-9443-47d1-b60e-ca3cce531441 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.402949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.418755Z digest=sha256:0d6d6093c3de0b58d90f89227e386458f88b89c58d592a4f005f2aa391180ce8

Observation 83199327-1f1c-43ce-b2e5-83f959260212 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.391508Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.477914Z digest=sha256:83e72e120d3609f7e7b5bf6ae8dd80939b3bbd6b0ac8f59bac2c738419f9fba3

Observation bc871ed8-fa1d-4bea-8be2-fe007dfce378 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.370167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.662146Z digest=sha256:70e57ab46dc70cc8190f4d2fe67b498ba348b7a1f02cecad75fb7eecb3902e14

Observation 1ab97031-acef-4f9e-ad8d-c56aaaabeb64 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.358812Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.770868Z digest=sha256:463ec11658f9a564ea4a3bf0229a29848a9bd0b99533400429b40ee48421fbb9

Observation 2b7a913f-8324-4b44-9b1e-1b9e3f5b7820 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.347515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.990439Z digest=sha256:08308396f5b568ca39f8e928fc35e06746782e14aeda1d5f3f0ef04b2ab7304f

Observation cd284ad0-8c6d-44a8-8089-e4fe71e5290b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.335767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:07.137170Z digest=sha256:7bf8df1fdad74e51eeae7786a2a70ccce6e32a664469cc0b28e120702904f2ae

Observation 9611d3b2-7769-4130-9ec5-359f6276a384 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.192945Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.192945Z digest=sha256:1c550f22992c207e2df2f4983ec915177309652bb4cc57259f5cd5fdee75b80f

Observation ef17a6f6-5b3b-4eed-a1fd-b7145d65b926 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.239818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.239818Z digest=sha256:0d8f1e4f3897bdfe636dbc3ec6fa6072fa0e503b552f45de1a4327e0b9f2bfff

Observation 4200f296-1d42-43ed-b819-ecd860d9a56a · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.367992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.367992Z digest=sha256:c50782462d3ebe85ba5513c78a0c8684dfca6e7c05c5bd535150d427f5138a66

Observation 70d7e93e-960b-4341-aa21-9ae7c69c9739 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.450675Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.450675Z digest=sha256:a794a04268d1fcef111e759705d915d80f1229cec94fc5d8b70df78fb40e71cd

Observation 9e81d721-8450-46f8-8ee6-f1a21ce056e0 · outbound

This paper cites LLaVA-OneVision: Easy Visual Task Transfer.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training LLaVA-OneVision: Easy Visual Task Transfer

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.629431Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.629431Z digest=sha256:c706a6cf04e58d7fe4f1097b9137c949a92c8f321f21b8202f5c973da8bc484b

Observation 9c729d69-5cf2-4e11-a5d6-6a17c1349d08 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.305056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:07.745701Z digest=sha256:82ae3762d3cc7a72424f81ab56ab501b69697d5ebef4646d356132c3a6c96b8e

Observation 83e6d246-856a-459c-a113-a803222f8168 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.293319Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:07.841615Z digest=sha256:9a8856ce1d51524732b67edb05c10c43d8802fdfe3cc1076cdc0bc6def3df458

Observation 967b6b02-e4b3-42d3-82d9-8529ca70ce21 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:07.958818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:07.958818Z digest=sha256:49d9c3592bb460c7918610756adca039505de7609a5250c098921dc6fbed6a07

Observation 7a6f3d12-8ea7-4a62-a844-66cdb26152bb · outbound

This paper cites Decoupled Weight Decay Regularization.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Decoupled Weight Decay Regularization

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.150028Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.150028Z digest=sha256:e2b1eb97d44f06d67e5a992644877a9c876c47d2680e2f2ec0a6dc288deaab0f

Observation e30bebd7-02f5-4476-921a-def0d6d8376b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 27

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.786336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:08.282021Z digest=sha256:bfef0ac2df2dc7045aea2d297b831a30e85eab71e4aac892f06a18eaa2d8d12e

Observation cb39c92b-bd58-4f6f-ad10-092f0a1893ef · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.369036Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.369036Z digest=sha256:342c52f639f6254e5ae0dbcbe84ff1f4c1cc3fe400480e42bdfa57ed8a7c0252

Observation 37c58238-42ff-4bf9-953b-4bed40a2e803 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:08.500385Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:08.500385Z digest=sha256:a2cb27db66a5ef1d788e8c1dee20751d3bd3759c777a0ad2efc2d0052794d693

Observation 1b3d9346-d2f6-4cd7-b211-1a9415148ae3 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.277858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:08.066644Z digest=sha256:c8a45be2789f23b8c22a02cda554b37238d1e46da515fc2adb2287c57b8c13ad

Observation a1931f7f-fb2e-4e03-a7c1-f53aa8c2a5ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.257452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:08.744453Z digest=sha256:37d710512b1b3766335442c0c7f687be382e998dc13d606a1ae69cdd71422814

Observation c7509f94-6061-400e-a3ab-c688ac59ffe7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.246457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:08.879157Z digest=sha256:2107b3d5f6c8307bd26007476c039591ea3ddd2a9d7461cd9b535d0af66e52d9

Observation e50fe659-5e26-497c-971d-a732af370509 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.036940Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.036940Z digest=sha256:8b94498082a8479b20214cfe19956addd64b2de21564e28331b416d07e1b8456

Observation aed89379-dbe6-4ca4-84f2-23e2fba0b7cc · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.229217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.171661Z digest=sha256:8cce977c70df349adcd48394f85b12ea01778b4fa05003998a78557e7bc180ad

Observation 950af699-33d2-4395-a2a1-eb8434b832ca · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 35

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.268243Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:08.634729Z digest=sha256:72fa2c54cd4cba622a10c693dc7458849ad24592ad69780658e660b9207d53f5

Observation c6c5c237-94e7-4014-9c08-2d56396cf76f · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.213411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.397609Z digest=sha256:8265ca8fb3ad000c7d9c5c9a0aa0e2e5da7978856564f34de08721159ca4b89d

Observation a0647a3f-c126-49dc-baa4-efa1791696ea · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 37

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.204296Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.470703Z digest=sha256:fe36ca1bdf0884c139d49a3abdbef608b0c25cb4efd108b93e8d6d473284ac06

Observation d28d1515-0ce5-4446-99ca-2500ef15e566 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 38

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.106579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.527545Z digest=sha256:52bc5c3b72eadddd18275e18557603ec1dd408fe7c847ee8f02d355c575dafdc

Observation eff0526c-e9e6-4ba3-991b-9c7ef0626f98 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 39

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.616681Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.564187Z digest=sha256:0413aefa18136f44547896b3970965f9cc87147909ccda8b62fcc9151c63b79a

Observation 21685015-2d51-4f6e-827e-06f166f1b562 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:09.313558Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:09.313558Z digest=sha256:d83e65c4685d3ae5a3ed2c71e4a7ad81a19fd4b0e3d2b67645d523d856e41096

Observation f6e54fc4-3801-41d8-90dd-2a439f743eb6 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.183491Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.723169Z digest=sha256:03c8ad3758ad7892197c75ecc6e2ae5177182714b9b8faa4de658575306dffdb

Observation 69ecbad6-15f3-4c3e-b905-19afa07a3021 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.174336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.753390Z digest=sha256:0807e5e9d3629635f4710d6181939f160b9f35d612690503e2b02ae898d6f685

Observation 550cf1ac-bb23-4821-b51c-c87adf33cdb1 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.164644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.841458Z digest=sha256:bb5dd1490784a8f337a0d2b6fdaae0e71f4aad40cd1f57d5d23ea41ee55190c8

Observation 67a6ee80-0852-4adb-9752-93e8c9d8ef55 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 44

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.459447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.925918Z digest=sha256:be602a8467c2a60573eaebc44e9aef82d4fc31444daf8447badf26ab33101593

Observation 22cc9f4b-dd66-4b1d-9dc2-f81a0c8f42b5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 45

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.193431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:09.652357Z digest=sha256:6fdacb75924a0d8d07db8f349ec40db1a15d45b5d236e95b943bae2e7d5f4552

Observation aa3a5a36-b428-4ce8-b9ae-39384141ac51 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.154808Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.171917Z digest=sha256:23cccc57f8a2837e52ea3ace4bd60a0c2f1fae15c35648ae61502db3559ad9fe

Observation 4ac61163-a401-4056-a125-e57f5f5d7156 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.252632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.252632Z digest=sha256:f913805a10059b54d9e420f44c01941c95eac1e164d35e2f4626c564d9de3ab6

Observation 0de8c958-c5e2-4b57-982b-ef280b189337 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 48

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.144085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.334837Z digest=sha256:c920d1764a6fdc79b310a2e6586e6deddaff520f64b78010470945dcfaec1294

Observation 8768b146-5d26-4721-b7fb-5b8611c43ba7 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 49

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.134274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.411765Z digest=sha256:db54eac9af37e55f4372882c23eeaebdbbb95971b39a66d12c7fbc9b7c6a8ea1

Observation 1171e4f6-1351-435f-ae4e-ac3f5ac56389 · outbound

This paper cites CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training CAM++: A Fast and Efficient Network for Speaker Verification Using Context-Aware Masking

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.032757Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.032757Z digest=sha256:f9f81342834207fe124676895c127d4c9eab4e30a41429124e043883e464e4ed

Observation f4053ef7-be47-4908-be0b-255e067d5932 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 51

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.113630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.600580Z digest=sha256:f42793daf59f7a5cd01d09ed66df7c5abb7b3330f819deeec13fd82a0281fc8f

Observation 0fa8b4b3-86be-4b42-b828-3ba73d142268 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 52

Resolution
verified exact
doi, observed 2026-08-06T11:22:11.289869Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.642109Z digest=sha256:400f587cf641750d783ef8b69ae675d4573a445c475a2be18a84bfe613911d8b

Observation 06a8dfcf-bfd4-46a4-8af0-f84c57d4259c · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:10.727452Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:10.727452Z digest=sha256:8d6078c566d556cf77468585c77ec75e473a573a6a9286bf3ef33faf8ecde8f2

Observation b942b722-472b-4e67-bf01-648e65d2b7a8 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 54

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.095957Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.821908Z digest=sha256:e5bdd0f9412a575a8d5374fa24af716398be19e152db3b94e8fdd1b4125c5421

Observation e0b9dcc3-b21a-44a0-a7ba-60578e7b11f5 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 55

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.124121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.511071Z digest=sha256:3d4362172f620d6304f9d077f13a9786282cf61146acfcb3929e32e0a8ef6d5c

Observation 64682249-b972-4a9f-89aa-cf6e4eb7dae2 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 56

Resolution
unresolved
no resolver link, observed 2026-08-06T11:22:11.003082Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:22:11.003082Z digest=sha256:074e4afa83d4fc6dbc0c6a317f7b72030881ab5679fa9e21ba137257059f8cb8

Observation bf2b3cec-c6b8-4b10-b4dd-de27c27bcd7b · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 57

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:12.967277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:11.094770Z digest=sha256:a9fc8f7b0c91915f9f4d9894e4bb5a4c223d5dc8753fe3f251abdb99877f2050

Observation d2f5310c-e3d0-4be9-801b-ded63e567e65 · outbound

This paper cites an unresolved cited work.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Unresolved cited work

Reference 60

Resolution
unresolved
raw_fallback, observed 2026-08-06T11:22:13.083669Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:10.908903Z digest=sha256:d9f4319204ae179050f968fe12e65d6a2a83fd50eff5f77b0379bdfbed5da72c

Observation 8a192271-f78a-483e-b291-d47b09b5462c · outbound

This paper cites Pattern Analysis and Applications 25, 4 (2022), 981–992.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training Pattern Analysis and Applications 25, 4 (2022), 981–992

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.380191Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.573227Z digest=sha256:ba08c0c1a03c13eed619c9fc2232f95d9f51e254af2836e740e5e6a21352329a

Observation 413f7d4f-1e0c-4756-a4dd-e455d85535a7 · outbound

This paper cites In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In Proceedings of the IEEE/CVF conference on computer vision and pattern recognition

Reference 2023

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T11:22:13.423278Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.137860Z digest=sha256:0ea2b68bfe852d08606b763fe79ff47babbabff247ed115e016f3655c3810da9

Observation 1b457127-83af-4786-8a86-36ed0d4483e1 · outbound

This paper cites In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT).

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training In 2024 International Conference on Signal Processing, Computation, Electronics, Power and Telecommunication (IConSCEPT)

Reference 2024

Resolution
verified exact
raw_fallback, observed 2026-08-06T11:22:12.712624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:06.880564Z digest=sha256:11b15a84d2e8790f5b94e739d6d42a8dbaca5c53e11d6ea0bf19d393dfef758e

Observation e5348dba-3fe9-4ea4-a7fc-2ebc60caace8 · outbound

This paper cites StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model.

HOLA: Enhancing Audio-visual Deepfake Detection via Hierarchical Contextual Aggregations and Efficient Pre-training StreamUni: Achieving Streaming Speech Translation with a Unified Large Speech-Language Model

Reference 2025

Resolution
metadata mismatch
local_arxiv, observed 2026-08-06T11:22:12.420970Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-06T11:22:07.063747Z digest=sha256:2552165ad8cd8ef08ab01734123c478192528f2ebfd8c4fc59d6399f85c8cb6d

Pith citing papers

No inbound Pith citation observations are available.