Pith. sign in

Paper Citation Record · LEDGER

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

As of 9 August 2026, this Paper Citation Record lists 33 of 33 outbound references and 1 inbound Pith citation observation for arXiv:2505.16279.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16279 v1

Coverage vector

measured 33 of 33 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:10.491851Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:07:07.961535Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T15:07:10.886803Z

Reference resolution

33 of 33 outbound references displayed

  • verified exact2
  • verified fuzzy20
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c74c4cec-69f9-42aa-861a-f36edf271657 · outbound

This paper cites Exist- ing dubbing methods can be categorized into two groups, each focusing on learning different styles of key prior information to generate high-quality voices.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Exist- ing dubbing methods can be categorized into two groups, each focusing on learning different styles of key prior information to generate high-quality voices

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:13.014975Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:07.840682Z digest=sha256:a293b30f0fa58fbf5855fa5ea6dba0be77eaa2a06923eef92f9dbe3bcb06a2ad

Observation 74e62b26-d73e-4307-998b-7010f449b1a2 · outbound

This paper cites MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T15:07:10.927698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:07.961535Z digest=sha256:ae8dacb63eb246e68eb785f38ed0c5a84af74d7c2854f206d026a2a789a9b1ce

Observation 907a7eee-cea2-48df-a5a4-d728b32c66ec · outbound

This paper cites Datasets Emilia is a comprehensive multilingual speech generation dataset containing a total of 101,654 hours of speech data across six languages [21].

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Datasets Emilia is a comprehensive multilingual speech generation dataset containing a total of 101,654 hours of speech data across six languages [21]

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.775612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.031502Z digest=sha256:7c5f3f251f34822ea7042d3a2a78fc6c3a826313fdaa42498771574db681b9ff

Observation 375117c2-b77b-41b1-86d7-a80476fd686f · outbound

This paper cites To as- sess pronunciation accuracy, we use Word Error Rate (WER) with Whisper-V3[24] as the ASR model.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing To as- sess pronunciation accuracy, we use Word Error Rate (WER) with Whisper-V3[24] as the ASR model

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.637611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.114058Z digest=sha256:adc8cdd1a7085274e8275bfe67e757e3244b058f443b178166e152a2ba53cc0a

Observation 498d685b-d814-4216-89ac-e165952d658e · outbound

This paper cites Additionally, we have de- veloped a movie dubbing dataset with multi-type annotations to enhance movie understanding and improve dubbing quality.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Additionally, we have de- veloped a movie dubbing dataset with multi-type annotations to enhance movie understanding and improve dubbing quality

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.543048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.198474Z digest=sha256:0e2e0640b4d57ea6a18febce9f0e96f5d00a927027d39150b627a3c9421aa2a5

Observation 0aea5d8d-5aee-43db-9ce9-3407e0333e0f · outbound

This paper cites V2c: Vi- sual voice cloning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing V2c: Vi- sual voice cloning,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.439554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.292579Z digest=sha256:6335c624ff74fa58e7cc055ba21c677b6a6331cba4e8007bd0e270a59a763ea7

Observation fc319574-6ad6-4352-ad12-923d5cb6d196 · outbound

This paper cites More than words: In-the-wild visually-driven prosody for text-to-speech,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing More than words: In-the-wild visually-driven prosody for text-to-speech,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.336180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.360687Z digest=sha256:61e0c35910ace192dd07b09ad7f4123420a7e356687d244bd3463fa730fbb59c

Observation 575950f6-6b37-4f81-b5a9-6886c6669eb4 · outbound

This paper cites Generalized end-to-end loss for speaker verification,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Generalized end-to-end loss for speaker verification,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:08.454604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:08.454604Z digest=sha256:630556f3f3ace516e9cd21b9b642393173f3aa66554f9e4b3a2fc2cadef08f95

Observation d6475ee3-f567-450d-b6ad-8bbabb0b15de · outbound

This paper cites Learning to dub movies via hierarchical prosody models,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning to dub movies via hierarchical prosody models,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.220221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.557957Z digest=sha256:5f9e2a8005aa22a647350c7da5e769f635eaf113ea3def3d2aba621e2887ea34

Observation 52d242fa-daf6-4206-9957-d93433bf20b2 · outbound

This paper cites Neu- ral dubber: Dubbing for videos according to scripts,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Neu- ral dubber: Dubbing for videos according to scripts,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:12.046317Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.676630Z digest=sha256:ad7133b93ee13ec14373ea02d02f231bdcff87694a83b46b5063b0ad45ac320d

Observation bc5635ba-34aa-4c19-942f-ce1a908c9175 · outbound

This paper cites Imaginary voice: Face- styled diffusion model for text-to-speech,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Imaginary voice: Face- styled diffusion model for text-to-speech,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.944032Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.818516Z digest=sha256:dfdf7b464d5ccfcfe5a543bbdba0bb553dff6dc929e93e3f9071598418c5b448

Observation 0912692c-d95d-4188-8855-7853c0c477ef · outbound

This paper cites Mcdubber: Multimodal context-aware expressive video dubbing,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Mcdubber: Multimodal context-aware expressive video dubbing,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.867601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:08.943533Z digest=sha256:0c9cd383b57ee51c6658cd3eb451fb9ab2c579911a84d8c3cab1259f66e1a570

Observation 4c9175e7-1513-408d-8ac9-b6bb9cec84cd · outbound

This paper cites Audiopedia: Audio qa with knowledge,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Audiopedia: Audio qa with knowledge,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.766439Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:09.054916Z digest=sha256:6752eb929411049c2739f4a87df6c7f6e2ef056dac1875660535c34468b19fe7

Observation e501b4aa-ab4b-48e9-bb12-9dc7b57355a7 · outbound

This paper cites Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Insight-V: Exploring Long-Chain Visual Reasoning with Multimodal Large Language Models

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.148574Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.148574Z digest=sha256:1a33749ee970974e96b003b4ecdddf4377da2b11e7b92ac65ca554f403fc9164

Observation f8be3954-69be-438e-9ad5-4432531ef75d · outbound

This paper cites VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing VideoLLaMA 2: Advancing Spatial-Temporal Modeling and Audio Understanding in Video-LLMs

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.258423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.258423Z digest=sha256:e929d9827d4a1effb479bcc4b5a067fb5b6401d63864ee2a303cf06cdd9041ff

Observation 6f78254c-f256-495b-aee1-8f76d4b0dadb · outbound

This paper cites Learning to dub movies via hierarchical prosody models,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning to dub movies via hierarchical prosody models,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.669537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:09.359041Z digest=sha256:6aeb97c8b040e4cef37550cce223b964562d2883b2ee10b395cc67d48aec88ee

Observation 37ce9657-f384-4ed4-a705-5791653bab47 · outbound

This paper cites StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing StyleDubber: Towards Multi-Scale Style Learning for Movie Dubbing

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.453123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.453123Z digest=sha256:a21dec0eb6d230630f21d7b95e4e8b2210540423f08705d327c959cd1c05fcba

Observation 3a0d21ad-481e-4ab7-aa36-f256148f54d8 · outbound

This paper cites From speaker to dubber: Movie dubbing with prosody and duration consistency learning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing From speaker to dubber: Movie dubbing with prosody and duration consistency learning,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.589518Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:09.527245Z digest=sha256:6c38e16c8e44c44e59c8546c45681242ab9bdba16e18d079cd7e6af2b5923bb0

Observation a2ca5ff3-e452-482c-a4cc-9c0d122a62cd · outbound

This paper cites Visual instruction tuning,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Visual instruction tuning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.508937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:09.645485Z digest=sha256:ed942fbad5976ec8f129935dc1e4abc5ca8cbc40d40008e9e0d4d701cebaee8c

Observation fb820020-97f2-4de1-9232-86c7858f292d · outbound

This paper cites LLaVA-CoT: Let Vision Language Models Reason Step-by-Step.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing LLaVA-CoT: Let Vision Language Models Reason Step-by-Step

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.737133Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.737133Z digest=sha256:52fa1f8776a8276bce9541d8652300da3ba8715e5013d61907f1623c7e0b6c05

Observation b22e3ad7-cc45-48fe-b3c1-c8e53935a9a1 · outbound

This paper cites F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing F5-TTS: A Fairytaler that Fakes Fluent and Faithful Speech with Flow Matching

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:09.853775Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:09.853775Z digest=sha256:d66e604ea70e3cef6018cf057711266dad58195781a0c105e9ce6955b80dbb49

Observation 297f42d7-bd54-44b4-aed5-644aafba4cb8 · outbound

This paper cites Flow matching for generative modeling,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Flow matching for generative modeling,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.422727Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:09.981897Z digest=sha256:6df2a9fc9e940b220b109e7445f5cda511b2410192c4948b797387ab305bc954

Observation e275e094-90a4-4712-9e20-fe0c7af3ff27 · outbound

This paper cites An audio-visual corpus for speech perception and automatic speech recognition,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing An audio-visual corpus for speech perception and automatic speech recognition,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.333160Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.333160Z digest=sha256:42aae111ff7cd8139019a5e56e993dec60ff1a777c0f854eec328c54faeae5b4

Observation 3d6cc985-0e82-459d-850e-bb737c7a2e05 · outbound

This paper cites Exploring the limits of transfer learning with a unified text-to-text transformer,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Exploring the limits of transfer learning with a unified text-to-text transformer,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.331277Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.067509Z digest=sha256:9ff9ea78c0c4365bcadbd1069e3cd980745561fbdd6bde4a5f3a71d9e657f626

Observation 0cdbc5a6-b3fb-4b09-aebe-f4d579bae2a7 · outbound

This paper cites Learning transferable visual models from natural language supervision,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Learning transferable visual models from natural language supervision,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.240478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.120717Z digest=sha256:eb4c86db79e028c3db777ac8d11baaf1c34b76710aed702bcfffe0cb2dc3dcbe

Observation 710d353d-232f-42a6-875b-bfde1a713ea6 · outbound

This paper cites DiVE: Dit-based video generation with enhanced control,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing DiVE: Dit-based video generation with enhanced control,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.170141Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.174456Z digest=sha256:300dd4a3299003c1295722ef36fd276e94911609ed2bf2215225d0730aeb0948

Observation f3fe19ef-6c80-4bbd-9835-209e6f266201 · outbound

This paper cites Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Emilia: An Extensive, Multilingual, and Diverse Speech Dataset for Large-Scale Speech Generation

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.212017Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.212017Z digest=sha256:70e57d5f520af47d310c54a2632ee4df70a627e6748354b690420b8dfb996f8f

Observation 5e30b8f5-a9a6-4481-8b01-4d2924fa4b96 · outbound

This paper cites V2C: Visual Voice Cloning.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing V2C: Visual Voice Cloning

Reference 28

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:07:10.775143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.279629Z digest=sha256:a965f7b8acbb9ab2f282056ca30c7c7070a795df5261bb6d4b6a3a18f70822e1

Observation 092bb8b5-818c-403a-9762-39b6ad937b9a · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Robust Speech Recognition via Large-Scale Weak Supervision

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.363576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.363576Z digest=sha256:c27a76d8314b8874779455188acdc14d23fdacd751a49b804d77f65fd1e50ce7

Observation fd3fd5fe-5aa3-4ab2-b700-eb822cd43a8a · outbound

This paper cites Location-Relative Attention Mechanisms For Robust Long-Form Speech Synthesis.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Location-Relative Attention Mechanisms For Robust Long-Form Speech Synthesis

Reference 31

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:07:10.612708Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.404068Z digest=sha256:980260784273b6dc95d44f363099b333f3348d98562433eeffaca752f6cac7b6

Observation 9edd3850-dac4-4e3a-a2ff-8b7a03a45d5b · outbound

This paper cites Tem- poral modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Tem- poral modeling matters: A novel temporal emotional modeling approach for speech emotion recognition,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.100841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.445186Z digest=sha256:a810bda6182f5826ce311a084a8d4b7b6603b8958225aaab5b9977deee03a8f4

Observation 0c2c5a5c-a7cc-448a-aa58-3a770be52bba · outbound

This paper cites Out of time: automated lip sync in the wild,.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Out of time: automated lip sync in the wild,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:07:11.031819Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:10.491851Z digest=sha256:4a0217b33b082455a2b76ba749f0bfda691efdce0b7ba86911d4633375ef5b73

Observation afeebc99-ff03-4cc8-8e88-c14dd6001225 · outbound

This paper cites Available: https://openreview.net/forum?id= PqvMRDCJT9t.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing Available: https://openreview.net/forum?id= PqvMRDCJT9t

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:10.028412Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:10.028412Z digest=sha256:c0f7d9b06cae92379ede228835ccefbafbc93cddd227aaf2e3c1cc7752ce4e22

Pith citing papers

Observation 74e62b26-d73e-4307-998b-7010f449b1a2 · inbound

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing cites this paper.

MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing MM-MovieDubber: Towards Multi-Modal Learning for Multi-Modal Movie Dubbing

Reference 2

Resolution
malformed identifier
local_arxiv, observed 2026-08-07T15:07:10.927698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T15:07:07.961535Z digest=sha256:ae8dacb63eb246e68eb785f38ed0c5a84af74d7c2854f206d026a2a789a9b1ce