Pith. sign in

Paper Citation Record · LEDGER

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries

As of 20 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2507.12723.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.12723 v1

Coverage vector

measured 41 of 41 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:45:54.536804Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T16:45:49.818821Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-06T16:45:54.946643Z

Reference resolution

41 of 41 outbound references displayed

  • verified exact2
  • verified fuzzy26
  • unresolved13
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a25f8065-3a17-4cef-a89b-32277d43078f · outbound

This paper cites an unresolved cited work.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:46:00.234428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:49.757320Z digest=sha256:5074e63b5b85fb2441fdd59bfb3200a445b8b4d397324776ac827869dcf5dcbb

Observation 25c12320-9ae1-42b1-82a1-810bb793b25c · outbound

This paper cites Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:45:55.025578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:49.818821Z digest=sha256:9221643317b79ccb9ac971a39917899bc46d4288ff86eaab076297347ee3eba4

Observation d6619807-2e32-46e3-baed-cd826fddd4c1 · outbound

This paper cites an unresolved cited work.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:46:00.104538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:49.931600Z digest=sha256:d323615b817ac1f7a0f2c612ceda4d4671dbc70cbb7419a0b0d20fb5b55bc06e

Observation a3f5f2fb-ceaf-4fe4-9090-1bafd1e75d36 · outbound

This paper cites Semantic Feature Contrastive LossTo ensure robust tam- per localization, we compare the tampered and recovered audio streams in a semantic feature space.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Semantic Feature Contrastive LossTo ensure robust tam- per localization, we compare the tampered and recovered audio streams in a semantic feature space

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:59.913070Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:49.996219Z digest=sha256:0bbe4f52a4e44a1495c2383938f857b975a727b7beab391092c1368f1c9e0a2d

Observation ec97d494-547a-4834-8835-68a045094ef1 · outbound

This paper cites No Mask” refers to training without a mask. Our masking strategies outperform “No Mask.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries No Mask” refers to training without a mask. Our masking strategies outperform “No Mask

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:59.766837Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.073833Z digest=sha256:cbcd308d0646e50e7afe2435635e76a03fd1fc70aafa13eb42eff4c6b17d10b4

Observation f0e59333-d118-47a3-a28f-f69472846bb2 · outbound

This paper cites To achieve this, we propose cross-modal watermarking method not only localizing tampered regions but also recover- ing authentic audio.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries To achieve this, we propose cross-modal watermarking method not only localizing tampered regions but also recover- ing authentic audio

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:59.598713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.133984Z digest=sha256:c5c8bd16da50fd88ad122eef41f7d68d79b7a6a98b959b1dc48f41a8a23078f3

Observation 73223795-f2e0-4ca2-9208-39d41d1c7d46 · outbound

This paper cites an unresolved cited work.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T16:45:59.401149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.216060Z digest=sha256:dad3af651a13d6d075c96c664dccdd5b3bc41ca9a3ea931ce9d9c9e69b035688

Observation cc92fce9-8588-4000-81ec-b2f8b062ca31 · outbound

This paper cites V oice- box: Text-guided multilingual universal speech generation at scale,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries V oice- box: Text-guided multilingual universal speech generation at scale,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:59.214657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.326051Z digest=sha256:0caf75c7b96d1c316bc9265a894221d6f0dc88e57b5f84f394b9407bdfbeba1f

Observation 71f584fe-cfd8-42c9-965f-2992f64d5566 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:59.049190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.431336Z digest=sha256:d5ddefc2c3bd819a8ddf54b751b474198597b599858b8c249f9da0e7e33e3d6d

Observation 0b4fa284-f4ef-4059-a7e3-cfe7be6d70e6 · outbound

This paper cites Prompttts++: Controlling speaker identity in prompt-based text-to-speech using natural lan- guage descriptions,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Prompttts++: Controlling speaker identity in prompt-based text-to-speech using natural lan- guage descriptions,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:58.870054Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.498661Z digest=sha256:11bfb3327423ca8b35b6e214575f9cae7c4df27349e21dff4e6d737d175d3e7e

Observation b89aa9a3-7056-4595-8637-97608f5fe06a · outbound

This paper cites OpenVoice: Versatile Instant Voice Cloning.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries OpenVoice: Versatile Instant Voice Cloning

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:50.581929Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:50.581929Z digest=sha256:3de90bd5e6c6de881cdf33b0aa3540a82d1a25dcfe9fe06d264d4a227d8bed41

Observation 81502a34-be0a-47c3-9da7-091f66d9bc2e · outbound

This paper cites Paddlespeech: An easy-to-use all-in-one speech toolkit,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Paddlespeech: An easy-to-use all-in-one speech toolkit,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:58.712099Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.667815Z digest=sha256:042af66d4a7892d4e36fe6986413735d9a629f6959d55b747c15aa5673904de8

Observation 3aba5e75-4eea-4ad4-9ecf-3adae756963e · outbound

This paper cites Voice Cloning: a Multi-Speaker Text-to-Speech Synthesis Approach based on Transfer Learning.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Voice Cloning: a Multi-Speaker Text-to-Speech Synthesis Approach based on Transfer Learning

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:50.765524Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:50.765524Z digest=sha256:f75856459145bfc8d604edf7487b44525213fb694efb3fa0e2a9156c3ba62294

Observation 229c49e3-fa14-4f9c-bb0f-38942db4d6a1 · outbound

This paper cites A lip sync expert is all you need for speech to lip generation in the wild,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries A lip sync expert is all you need for speech to lip generation in the wild,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:50.866421Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:50.866421Z digest=sha256:aee558bbd110c0c2e670ab74c80a58337bf796a6879659f643d9cc3ae86b3ae2

Observation f5d8182c-2d28-4ac4-b4b0-136a5d140498 · outbound

This paper cites Diff2lip: Audio conditioned diffusion models for lip- synchronization,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Diff2lip: Audio conditioned diffusion models for lip- synchronization,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:58.518164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:50.990832Z digest=sha256:0925f2e765bb884820e5798155dbd6978729970de6f42fa3268d69d50c07c79d

Observation 53301581-388c-4586-a1e1-8555189d3b96 · outbound

This paper cites Pose-controllable talking face generation by implicitly modular- ized audio-visual representation,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Pose-controllable talking face generation by implicitly modular- ized audio-visual representation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:58.303254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:51.130244Z digest=sha256:e655f1db69a18ff40d721aad202f38da5d8bccce662f5333ccf4a0bfb49349d2

Observation 98f9c78d-94b4-4d71-bb5e-9f17137ba254 · outbound

This paper cites TransFace: Unit-Based Audio-Visual Speech Synthesizer for Talking Head Translation.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries TransFace: Unit-Based Audio-Visual Speech Synthesizer for Talking Head Translation

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:45:54.731731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:51.295199Z digest=sha256:3521612527a287f18b7a7fcaf67050b564a8fc254f772e4f22d0fda703ffa329

Observation b077b6a2-bf19-4acd-922f-2bf55b3c6242 · outbound

This paper cites Synctalklip: Highly synchronized lip-readable speaker generation with multi-task learning,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Synctalklip: Highly synchronized lip-readable speaker generation with multi-task learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:58.113236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:51.472799Z digest=sha256:8c32ba5997b49567143856f6df99c4e3c7f91df007e44adcb3fb668093c3ffcb

Observation b5bf171c-b948-419b-8bfc-8ca4c4a494a6 · outbound

This paper cites Editguard: Versatile image watermarking for tamper localization and copy- right protection,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Editguard: Versatile image watermarking for tamper localization and copy- right protection,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.865630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:51.591498Z digest=sha256:fcc70474bb314516326166f166bd567510dbcbc47894d56779c8238a8e05badd

Observation 72669cce-d637-4d0e-b732-dca031e82883 · outbound

This paper cites Proactive detection of voice cloning with localized watermarking,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Proactive detection of voice cloning with localized watermarking,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.715993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:51.723042Z digest=sha256:ab7198f2950fba232e2491373ef68d409be84b31714a02ed1288da51e9d6a025

Observation 1f2ba057-3157-4c10-ab2e-eab0e2af9c8d · outbound

This paper cites Enhancing Partially Spoofed Audio Localization with Boundary-aware Attention Mechanism.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Enhancing Partially Spoofed Audio Localization with Boundary-aware Attention Mechanism

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:51.888779Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:51.888779Z digest=sha256:38b4c6f2d7336667f355cb61d63bb703e8c0cb190c44376e60b8e969a02a68c1

Observation 0080dac6-140f-43b4-b90a-2f0c27ad71ea · outbound

This paper cites Coarse- to-fine proposal refinement framework for audio temporal forgery detection and localization,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Coarse- to-fine proposal refinement framework for audio temporal forgery detection and localization,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.559506Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.038159Z digest=sha256:4725e4ad0fd0aaf3e57050474f4d7574a6c87c699796483242da9bcf564c642b

Observation c0570cfc-711b-4ea6-8770-abba5c953757 · outbound

This paper cites Wavmark: Watermarking for audio generation,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Wavmark: Watermarking for audio generation,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.343824Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.153256Z digest=sha256:8c4ba8b8780d53e7f7553c79c985300858af8050cc0e0ac672c5bc55d5506781

Observation 615ad3a1-de66-4b32-943f-8f0fe7785935 · outbound

This paper cites Hiding data in images by simple lsb substitution,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Hiding data in images by simple lsb substitution,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.191686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.320247Z digest=sha256:a6e2a86d608d0d394fdc771c9f784507851c3c41c6f10b9d39df833d4d39b6fa

Observation c8440b55-d5a8-4687-a597-fddd9084fa02 · outbound

This paper cites Complete video quality-preserving data hiding,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Complete video quality-preserving data hiding,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:57.038920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.445262Z digest=sha256:d2d92ef9911a6a80c8bf95f4cf69eb639e040cc2e4da8cc96273d75770db7963

Observation b970d1cc-97a7-4998-b259-4e1814ea24cd · outbound

This paper cites SteganoGAN: High Capacity Image Steganography with GANs.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries SteganoGAN: High Capacity Image Steganography with GANs

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:52.573245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:52.573245Z digest=sha256:9818e1f0e9a54894a4c30d30c84dedaef78c27e120e3ea879e7c79a9401420d2

Observation 67a01fa2-a2a6-4d05-9fcf-565b0113052e · outbound

This paper cites NICE: Non-linear Independent Components Estimation.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries NICE: Non-linear Independent Components Estimation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:52.709737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:52.709737Z digest=sha256:9368ac200281eae9b09815252b7721520df34e53e71d7ffa6845c683e45eb6c1

Observation f153a800-45aa-4f98-a867-49995a956c3a · outbound

This paper cites Hinet: Deep image hiding by invertible network,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Hinet: Deep image hiding by invertible network,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:56.844075Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.852091Z digest=sha256:3c8685d610b8f1ef560e6f0e2eceab75ce44aac26726cb07f6d892b05f913631

Observation 5fb90bcc-4121-42f4-abb8-0613fea70645 · outbound

This paper cites Large-capacity and flexible video steganography via invertible neural network,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Large-capacity and flexible video steganography via invertible neural network,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:56.713613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:52.960294Z digest=sha256:02b825f050be525e3b948c348a5b249d3d3151c9e069aa7e7072796b9b0b78e7

Observation 11ac612b-a30f-41e9-91da-3642659a8a73 · outbound

This paper cites Thinimg: Cross-modal steganography for presenting talking heads in images,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Thinimg: Cross-modal steganography for presenting talking heads in images,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:56.529900Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:53.117346Z digest=sha256:f9cf0c8bfbecc3cc141e854145207c573afc66032c4044d64a432894fb0a1e20

Observation 953805b2-14eb-4e8f-abaa-34a962a09c8e · outbound

This paper cites Densely connected convolutional networks,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Densely connected convolutional networks,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:56.342670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:53.257898Z digest=sha256:8b352548c717b6d800035311ec11b6b99b74a8f57d0b43ece3d3dd46653c1b61

Observation afe04c73-1ffd-4c35-ac6d-8985cd75e0ca · outbound

This paper cites Cnn architectures for large-scale audio classification,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Cnn architectures for large-scale audio classification,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:56.157926Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:53.380957Z digest=sha256:b5a5bf9ffca38dd56d131dbfa893545bdfc0878b257902dc91060660a94524aa

Observation 6d9353e5-e375-432a-b6f9-ec3483484898 · outbound

This paper cites Representation Learning with Contrastive Predictive Coding.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Representation Learning with Contrastive Predictive Coding

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:53.537636Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:53.537636Z digest=sha256:f15087c41c6d3ea6dbf679153c2f4ee75e41a200068304069f37178e145e4413

Observation ad9db548-882e-41dd-acb4-d7516b626039 · outbound

This paper cites Lightface: A hybrid deep face recognition framework,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Lightface: A hybrid deep face recognition framework,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:55.988215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:53.679594Z digest=sha256:5d75b1a6d6e4dc5fc6108492d51b7aa747ef605f9e56351e3c64846e2e39ba48

Observation 2a2bd7e8-bdc5-4d37-beb6-66f7fe7deca1 · outbound

This paper cites Flow-guided one-shot talk- ing face generation with a high-resolution audio-visual dataset,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Flow-guided one-shot talk- ing face generation with a high-resolution audio-visual dataset,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:55.759725Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:53.799916Z digest=sha256:3142d09985f3abdd64380bdb59e8ec08bd235abacf56590a539753b6ccfe08b8

Observation 80ba801b-e27a-466a-b7de-f8c384c73ac8 · outbound

This paper cites LRS3-TED: a large-scale dataset for visual speech recognition.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries LRS3-TED: a large-scale dataset for visual speech recognition

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:53.924279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:53.924279Z digest=sha256:712327e36ecf60b229c8ec5abcb640dd2be7e12ccc0b31b4d43befdcdd392fa0

Observation 097491e6-36c7-4393-9268-f2bb64e5a0a4 · outbound

This paper cites V2a- mark: Versatile deep visual-audio watermarking for manipulation localization and copyright protection,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries V2a- mark: Versatile deep visual-audio watermarking for manipulation localization and copyright protection,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:55.579282Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:54.102283Z digest=sha256:8bef9f016510b7c2f7ff0886d93b483e6b86802e196e3ef1a2209b2f3c034558

Observation 8bb5e626-134a-475d-9179-74e5dfaa2c73 · outbound

This paper cites Musetalk: Real-time high quality lip synchorization with latent space inpainting,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Musetalk: Real-time high quality lip synchorization with latent space inpainting,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:55.436801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:54.229478Z digest=sha256:57cdf1e1f48a14490f496231bf21cb24d864c279d58e2a408f6b1d1113435007

Observation c93bb14d-866d-450a-aaea-3dc1355e87b4 · outbound

This paper cites Video enhancement with task-oriented flow,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Video enhancement with task-oriented flow,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:54.316832Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:54.316832Z digest=sha256:3ab1f7554b87061cf5d282f64a5bd114dd84ee64cc58ba122390af337a1479fd

Observation 3dcdb2be-cec6-4c2c-b9ec-e1ce30e325c3 · outbound

This paper cites FMA: A dataset for music analysis,.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries FMA: A dataset for music analysis,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T16:45:55.220089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:54.433749Z digest=sha256:38f2bcd7b4a34d28ed7d6bcd4592ee51f0236f5eb307ad3d55bebd71dbd46dca

Observation 4fee4e4e-2041-4248-86d3-983a29494207 · outbound

This paper cites FMA: A Dataset For Music Analysis.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries FMA: A Dataset For Music Analysis

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:54.536804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:54.536804Z digest=sha256:200a901fb0a19ebd15a33b1c28e340e2c48677236b87b01310d82c2e4ee574fc

Pith citing papers

Observation 25c12320-9ae1-42b1-82a1-810bb793b25c · inbound

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries cites this paper.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-06T16:45:55.025578Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-06T16:45:49.818821Z digest=sha256:9221643317b79ccb9ac971a39917899bc46d4288ff86eaab076297347ee3eba4