Pith. sign in

Paper Citation Record · LEDGER

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework

As of 17 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 0 inbound Pith citation observations for arXiv:2506.07358.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.07358 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:42:04.276221Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

49 of 49 outbound references displayed

  • verified exact3
  • verified fuzzy3
  • unresolved43
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d932c7ed-4ca3-4368-8ab5-c4001ae4f8f8 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.213542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.029926Z digest=sha256:53776d2ee4300321c40511efe9200decd6d280763b9d9c348c3be3ad09f8347f

Observation 82f9062f-70ba-4552-a3a2-a6ccc8d4823f · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.192476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.035372Z digest=sha256:02f65d3724bfacb01aa2f151b83b953bfb134466ab5ea30a28af2a78c83935da

Observation 477f9f35-e939-4510-8703-2cef79e37f96 · outbound

This paper cites D.; Junior, A.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework D.; Junior, A

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:42:05.175277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.040660Z digest=sha256:f3e987fb883e19caf0f6e29300d8ff2a50a2c89c5ba771c4ea363ff65820bbb7

Observation 5179a6b0-ace7-42ff-a490-04386f60718d · outbound

This paper cites Voice-Face Homogeneity Tells Deepfake.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Voice-Face Homogeneity Tells Deepfake

Reference 4

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:42:04.494537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.045983Z digest=sha256:f40074da3a4a42b5d07f946a7dcb6e7245bef8062a89256075b5ac47f1a98a2f

Observation dfe00bdd-68d6-42cf-aaf0-9d32c5e636d6 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.156901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.051439Z digest=sha256:5ffea0a7f9804323bcb0622956f0626d758f10a2ae7a92f2bdc6fbba8934aa5d

Observation 1354ea75-9676-432c-8178-1f179cdbc2fd · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.138982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.056461Z digest=sha256:11111ae86179b03ed872fc42be8f94cb2b95a46c21e4b51d641c016c93b5630e

Observation 7858a586-c626-431c-a6c5-2f2033923550 · outbound

This paper cites VoxCeleb2: Deep Speaker Recognition.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework VoxCeleb2: Deep Speaker Recognition

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.062150Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.062150Z digest=sha256:e2fcefd9afb831ef8e73ea69d14aabf18a5b069b0b678f68f6282c41f081bc0e

Observation 350e9093-c8b6-4d6c-ae25-4fb36b3c1385 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.118325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.067527Z digest=sha256:e8ef99c639f2506378cfe1d5ddb475b5d7f2e1fa4d19d7422d682e0983499bdd

Observation adf32a93-4406-4062-ba40-040327f3a857 · outbound

This paper cites The DeepFake Detection Challenge (DFDC) Dataset.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework The DeepFake Detection Challenge (DFDC) Dataset

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.072119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.072119Z digest=sha256:31b01230a2d292401783ec9de80cc815f1fe9da6c04e3231dca6cc2fbbe55dd6

Observation d03c9ffe-7b19-40f5-b286-53097946da34 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.097516Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.077748Z digest=sha256:b334b07135d6daf85abb997f2f28d89a2e58fd174a6e20ec9fa8799781cb1ee3

Observation aa1e58d2-70cd-4c7a-b398-ad2a56443a3b · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 11

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.072596Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.082974Z digest=sha256:2e345b3669a378e94a31ad9354511eeb3d18a6f03d96846285ec6e07a91161fa

Observation b0c1d5a8-0354-4659-b093-397d46653bff · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 12

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.053385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.088532Z digest=sha256:ce8712c18a56dda078f67091d61da93f81867b8502c9723ea8374b76ef316d37

Observation 3e329a2f-7aeb-4940-b0f7-24471e8db1ab · outbound

This paper cites Visual Attention Network.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Visual Attention Network

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.093554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.093554Z digest=sha256:cc5795f6d413e84d986e2204dff4814ab85f706dbf9e5f5e1fb1ed653ef29a98

Observation 2e6e09ac-d64e-4d00-9906-05608cca957a · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 14

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.035225Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.099424Z digest=sha256:ffe4b636c6f4360c193f0a6eb7f0cd23cea042e389331e4fb27a32001d243799

Observation eb5c6911-9d0a-45eb-8d58-7dfec0a78c0d · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 15

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:05.016501Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.104079Z digest=sha256:da6cc420e3816dffca8b16b2ecbe82f1fa53fb4648d267051b309a08d0bfe3b3

Observation 97b184dc-c1ad-4ac6-aedc-bde5f357aa0e · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.996973Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.108993Z digest=sha256:9825f959d7a2fe49d41c6246c40ec6f9d9de3e87a0255e7e787dbdaa734bd057

Observation 8931119d-a7c6-4e17-86d4-049805251d5b · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 17

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.976518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.113762Z digest=sha256:82cb2eebcc8eaa55193c3ed4cb96ba3f656b1e2bdb7cf7a0f81d71812f0cd706

Observation 69583204-6771-499c-b77a-fe3710c19bde · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 18

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.960142Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.118779Z digest=sha256:dc72ec7dcb36ce09be10c9bcf8b7d8b531ed93e90dd3420f92c6e04fee2cc49a

Observation 09b72f0f-5205-4a4b-b6bf-38f28ddcbba4 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 19

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.942102Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.123674Z digest=sha256:8fbab60bede3fb282e92a1958035d5af4732ba91362fdedb35f4288655561c4e

Observation 51b657b0-6b5d-413e-b823-158be3e7b645 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 20

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.923622Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.130316Z digest=sha256:a388066fb230b95a1b5e8b24daeb866bca96c779d2e186ed3eb9f356e0cc9ab8

Observation 114e5795-2b01-417c-b5ed-be5f5f79031f · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 21

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.906680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.134920Z digest=sha256:bba47c1b2c7a27c6b5d0bcd40e1e1909885671a54b1429d579f056f8c4527ed5

Observation e0a363cf-e82a-4ba5-afd2-67d875d113bb · outbound

This paper cites FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework FakeAVCeleb: A Novel Audio-Video Multimodal Deepfake Dataset

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.139918Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.139918Z digest=sha256:723c3da0303636c23d1f2aa95ce260d4e1a45837c14791cee163050147fb5d5f

Observation 124a687b-c5fd-4620-85c6-ed4338df9ec1 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 23

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.887298Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.145973Z digest=sha256:8f10434ce68d37416e83086b9eded156c5cb5515f0b81e85c785aef5c700490b

Observation 16bc51e1-b55a-4c9b-9f7d-0a85edcc0805 · outbound

This paper cites DeepFakes: a New Threat to Face Recognition? Assessment and Detection.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework DeepFakes: a New Threat to Face Recognition? Assessment and Detection

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.151101Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.151101Z digest=sha256:957193cfd4c842b0e36c54e1f418f6f0d5b4c5712a1c85ea46720ee9a9b2f251

Observation 075f088a-0ce5-4626-b19d-845f7ca9d732 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 25

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.870124Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.156090Z digest=sha256:f9e943c820d88acdc81787f0f00581a9600740f3f4533fed177bbf5bb3f8eea5

Observation 6fbc019d-22d1-4515-b423-d5826f8e5cc6 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 26

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.853030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.161234Z digest=sha256:b963e68bbadaa159aee1956486dc6af533bbba7621c3d58917c1ff5c3db5e01b

Observation dba3cc92-5a90-4e22-98b7-52f15aa7496e · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 27

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.836199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.165938Z digest=sha256:0952d45e28ffb3708054f1104b2094242303021bca1b78d42fe9f8cfbbcd3192

Observation e279f52c-3a1c-44e3-a4f6-c9fc4d7aa873 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 28

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.819749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.170906Z digest=sha256:e9bb2692d1cd18304beae9e2db870b3e915f689e78fb115fe2650aa787e6dbff

Observation 9036c8f4-c8aa-42ec-ae27-f01edbeca5ed · outbound

This paper cites Decoupled Weight Decay Regularization.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Decoupled Weight Decay Regularization

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.176670Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.176670Z digest=sha256:d34b5049b90472c61db18bc04a418e06aa7ece87bacc66d47b3e28fd22afcfe9

Observation 15c30617-855d-4a27-8777-5f634c6a02bf · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 30

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.802411Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.182189Z digest=sha256:df1406a0abb20b17835889b7d7ec44fe25d80cd1e93ace7891386e269ea15483

Observation bcf4e421-3825-4861-805d-4c8fb5e67a58 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 31

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.782273Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.187427Z digest=sha256:be8a33295b6510a617c35201653acac600092540f10771e9e69def70dd6c7fe6

Observation d6faeb58-a3c4-45b8-a913-ad177519e942 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 32

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.764403Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.192604Z digest=sha256:ec06de58061a6ad5c6f06a71c84929530b84dd23675e75dc465723c8a32edf45

Observation 4f560643-f977-4c47-9d24-8fbb25ec6b83 · outbound

This paper cites A.; and Malik, K.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework A.; and Malik, K

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:42:04.747238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.197467Z digest=sha256:05881aee03be7f46f755f1bdbf0c8a31203c8a4d0a477eae5112d36889e32c41

Observation f5424935-9ffb-498d-b29d-984dbbbbf514 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 34

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.731007Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.202159Z digest=sha256:077b4199fdcd0bd6423af3d066f36f626e99ab374d48901c1fbcb3266d3be307

Observation e594b4d7-a089-4136-b846-96af3ffb29b5 · outbound

This paper cites R.; Cogswell, M.; Das, A.; Vedantam, R.; Parikh, D.; and Batra, D.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework R.; Cogswell, M.; Das, A.; Vedantam, R.; Parikh, D.; and Batra, D

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:42:04.712739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.206748Z digest=sha256:3547d69d0939b85340751c54beb2bc51521000482432ba53c8da9a89c1d0540c

Observation 7a8c4977-ffcc-4bc7-8c4f-21634d48fdee · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 36

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.689876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.212124Z digest=sha256:335fff06799ac9e92aee815371e83d05c477edd8edba12c38f5a55e389b4b4a0

Observation fbf7b5a5-2694-496c-8f7b-412c85292972 · outbound

This paper cites N.; Kaiser, .; and Polosukhin, I.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework N.; Kaiser, .; and Polosukhin, I

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.216996Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.216996Z digest=sha256:fd060c0e3d9c42089f9a2201b44f6526a1992f3c786028bb8977147d126d301b

Observation 1a2bdf86-1ac2-4163-82c1-223489d87b37 · outbound

This paper cites Attention-Based Lip Audio-Visual Synthesis for Talking Face Generation in the Wild.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Attention-Based Lip Audio-Visual Synthesis for Talking Face Generation in the Wild

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:42:04.366185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.221624Z digest=sha256:95fdb971cfd01932c7ef1b5b56bad20731af8389756ed0857e299501a9bd8b36

Observation 51521d00-ae19-43ae-a675-c6a3df7618ad · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 39

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.658394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.226704Z digest=sha256:f6af61f6d8170565f5a375ad8d45ad19538646ffd622a3d847d80f9abd0fc287

Observation 784490bb-7a32-41bf-83f2-e747ac6d1da3 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 40

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.641703Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.231527Z digest=sha256:53f8a763497e070280a80367aced1e282e3721c445951e0bd99652dcc0afd2b8

Observation c425bffc-f7cc-4f6c-be41-a3022939a282 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 41

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.615851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.236318Z digest=sha256:0637b772ee54348692bc7cd5e8ff020146f620fca631b04a77d0588b7ed380a8

Observation 301da320-ab60-4373-a39c-0fc550e59ec1 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 42

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.596632Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.241476Z digest=sha256:fa578cc1ddb8fe662c8ced3825625bba966db14381d1a2171a88789f25f671f6

Observation d3921aea-112f-424d-b687-0dd2c06866d1 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 43

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.579304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.246518Z digest=sha256:2705a0d675326a56122cbeed95990295af7cd431c55692f42f4691794699d4f5

Observation cf4de127-1219-4142-86cc-51c55b0f22ae · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 44

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.562993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.251250Z digest=sha256:1047746d66ac9b41f10d8d83d0e22c3d67bbbaf409e19916411c945b1a4d09af

Observation 5778d8ee-344e-4019-b8b7-83b4e9e9b7bb · outbound

This paper cites Self-supervised Transformer for Deepfake Detection.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Self-supervised Transformer for Deepfake Detection

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.255746Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.255746Z digest=sha256:bdd2bc17f461fc60e6e77f3a21c6fe8a745d34aae7c2077f18f200397ed58b9a

Observation 5dba9f14-9890-4697-a8b9-c56734a51e47 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 46

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.546299Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.261343Z digest=sha256:4a4540ae6bb41fe6c3b1cd076d408abacff46a7dcce6a035aa3940a397139efe

Observation 5df90ccb-32ae-4fe1-9a1c-3b9944f0a3a1 · outbound

This paper cites an unresolved cited work.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Unresolved cited work

Reference 47

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:42:04.527667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.266216Z digest=sha256:ba1a82f04de390f42772a7d70523e80be2f35261f0e08866b1d9c1885d8b88b8

Observation a13fa297-47ff-435d-a5aa-11a9d6feaa33 · outbound

This paper cites Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection

Reference 48

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:42:04.324836Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=arxiv_source observed=2026-08-07T05:42:04.271018Z digest=sha256:e25122c7df0068da58de1619c60b913743841e50359ea0d069a6a6eab4a660cf

Observation 4ded6f32-1db2-4e76-bf15-f0db686ffdc8 · outbound

This paper cites write newline.

Lightweight Joint Audio-Visual Deepfake Detection via Single-Stream Multi-Modal Learning Framework write newline

Reference 49

Resolution
unresolved
no resolver link, observed 2026-08-07T05:42:04.276221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T05:42:04.276221Z digest=sha256:09da1992dd6117ba7bbd8dd1c8bc89bb31ba97471723d1ce0c5adc9596cfe772

Pith citing papers

No inbound Pith citation observations are available.