Pith. sign in

Paper Citation Record · LEDGER

Zero-shot Voice Conversion with Diffusion Transformers

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2411.09943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09943 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:58:57.726231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:29:41.883949Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e4e2bbb6-2e69-4f3a-8511-830be7742da8 · inbound

Kimi-Audio Technical Report cites this paper.

Kimi-Audio Technical Report Zero-shot Voice Conversion with Diffusion Transformers

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:27.252453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T19:21:26.933349Z digest=sha256:34d5c841b28f347b113569ee910f97ad2c6d951f307626222476387bee82156c

Observation 0b9c0f93-9ca5-4759-9b6e-400d1b5d827d · inbound

EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion cites this paper.

EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:58:57.726231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:58:57.726231Z digest=sha256:36902ef80ed1a9a0d3a68b103cf3d8ef25cc47e11c3727c3e6d7186b761542de

Observation d4926be1-b5a6-4c69-9098-2f81b5e375b4 · inbound

IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech cites this paper.

IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:21:56.314992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:21:56.314992Z digest=sha256:9b05c84c3695791a2d1fe310e271aee8774920894c963fb6f855cfc5479aaf2a

Observation 9f9872d1-24ce-41d6-8e04-0265e059ac42 · inbound

De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks cites this paper.

De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks Zero-shot Voice Conversion with Diffusion Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:21.074104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:29:21.074104Z digest=sha256:628fedaa918177cdcfbbf96cd9d770e1edebdf09be71f7817d1175e726f5745e

Observation 6fcb33aa-9703-4594-a6f5-c5c54c7c1d50 · inbound

The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents cites this paper.

The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents Zero-shot Voice Conversion with Diffusion Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T17:45:45.111222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:45:45.111222Z digest=sha256:39cc4448c715ab11a82a774d5ac43629d58ea0228748eb2c63eb266957e5ed0c

Observation fb5a0ba8-41c0-4959-a554-c5bd674b3d0c · inbound

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods cites this paper.

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods Zero-shot Voice Conversion with Diffusion Transformers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T12:49:19.452862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:49:19.452862Z digest=sha256:09147637ae3119cad77993ff406cc48018e8a22e49d94561384dba9b02ef4fa8

Observation b728c6e2-5022-4993-a1fa-bde3f00b2e67 · inbound

REF-VC: Robust, Expressive and Fast Zero-Shot Voice Conversion with Diffusion Transformers cites this paper.

REF-VC: Robust, Expressive and Fast Zero-Shot Voice Conversion with Diffusion Transformers Zero-shot Voice Conversion with Diffusion Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:41:51.071923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:41:51.071923Z digest=sha256:350143bbbcfe38484778c759b6f6d042e3be0b4e9dfc4d56b3de7087b966fa57

Observation ac23beba-29e1-4989-9b7f-6b0eebea3e0b · inbound

Semantic-Aware Ship Detection with Vision-Language Integration cites this paper.

Semantic-Aware Ship Detection with Vision-Language Integration Zero-shot Voice Conversion with Diffusion Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:41:57.979302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:41:57.979302Z digest=sha256:44a5f73404a622f63eea3a428ee763460d5384365b26a79f9808a96a0440e6c3

Observation 8c260918-7ea5-4d2c-8c26-e641f28fb380 · inbound

Entropy-based Coarse and Compressed Semantic Speech Representation Learning cites this paper.

Entropy-based Coarse and Compressed Semantic Speech Representation Learning Zero-shot Voice Conversion with Diffusion Transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:05.343480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:05.343480Z digest=sha256:5182742daac57e66c0a79fc19e1a260ebb047bf80aadf47bae3780c19dc352dc

Observation 0efb3374-c69e-4d13-9310-fc9df2c2a120 · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Zero-shot Voice Conversion with Diffusion Transformers

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.238340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:c543d9834ccb221334d12892a2a35be867c96209dcd9a9eac12990250326feb4

Observation a6cdb647-77bb-46ec-b961-3b52975bdb6c · inbound

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis cites this paper.

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis Zero-shot Voice Conversion with Diffusion Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:19:33.944804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:19:33.944804Z digest=sha256:e3a81d05c8818f60c528765c5d46d92756bc4934767e06dcc583af480e60894f

Observation c308a249-ce11-45e9-bd90-e015265ee38c · inbound

Universal Speech Content Factorization cites this paper.

Universal Speech Content Factorization Zero-shot Voice Conversion with Diffusion Transformers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-15T12:21:48.698333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:21:48.698333Z digest=sha256:bc1fbef851ea1213419c22e9ee6635def272c8d0bfe6e2b4cc945d42a9f8c143

Observation 7c8828e4-e183-4b8f-8777-4669fb22fe49 · inbound

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan cites this paper.

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan Zero-shot Voice Conversion with Diffusion Transformers

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:21:00.644713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:40:10.657590Z digest=sha256:b82669b7d4e387b563881097129e2ffba06c9f258cca505d30e2c822cf817c3f

Observation c463817e-4ee0-4514-900f-962ebbd3c4e3 · inbound

X-VC: Zero-shot Streaming Voice Conversion in Codec Space cites this paper.

X-VC: Zero-shot Streaming Voice Conversion in Codec Space Zero-shot Voice Conversion with Diffusion Transformers

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:20:29.712580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T14:20:01.885544Z digest=sha256:8b9e2d0683ea4130fb5514930abfaa2e6542ef48549ca4dbc9126ee7edc20c11

Observation 48646aca-6227-401e-acaa-d2dcad05656b · inbound

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction cites this paper.

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction Zero-shot Voice Conversion with Diffusion Transformers

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:59:50.728859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:59:22.416259Z digest=sha256:2437653b172c410dfb99d09bdc78a03172e9956ff0d7450c2d6153a6a0ce1000

Observation a5ef4e07-2da1-43d8-94f8-0df1aa567665 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? Zero-shot Voice Conversion with Diffusion Transformers

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:04.023382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:44b36e31e7745a657c292c69d8cc96674d7813ca72693409108e0737228535a3

Observation 2ff6df34-4f54-4a2a-b6f1-6ff99c44e7b2 · inbound

RTCFake: Speech Deepfake Detection in Real-Time Communication cites this paper.

RTCFake: Speech Deepfake Detection in Real-Time Communication Zero-shot Voice Conversion with Diffusion Transformers

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:26:17.872375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T05:29:45.895667Z digest=sha256:11c45be96604f336e8ca2fee1910b697d7db67bef099deedc42dde8e7f010ed6

Observation 854b3826-4e6a-4e10-8493-9af9f355a1ac · inbound

Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling cites this paper.

Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling Zero-shot Voice Conversion with Diffusion Transformers

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:12:13.878150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T04:09:48.737629Z digest=sha256:b873048f3603a47a03e3068dcd96476a2e649c016b3e27d27e5aff684bd67400

Observation 0ea1d867-762f-42d1-b092-b9539be5185b · inbound

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing cites this paper.

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing Zero-shot Voice Conversion with Diffusion Transformers

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T01:06:23.865289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T12:54:20.815371Z digest=sha256:051dcd9ed6f81cc0b798513e086c706af93e4a8e3b81a1147ceed9b71875eca6

Observation dfd36285-41fe-4150-8876-fa50210360b2 · inbound

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data cites this paper.

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data Zero-shot Voice Conversion with Diffusion Transformers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:28.663888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T17:39:36.966812Z digest=sha256:ddb98d2f555b6632d827e7b52f2a0675821f8c47b4dc3ff20fa3b36ef2c330bc

Observation 4c810e7a-48ce-47fc-9a6f-d83682abe6fe · inbound

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data cites this paper.

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data Zero-shot Voice Conversion with Diffusion Transformers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:04:37.810024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T10:57:44.041182Z digest=sha256:70038311e55a5b9f4d244210228885500d49ccf11b87e79a8ba072b72e4052fc

Observation 4ef1ae2b-b912-4f11-afe5-0354b55e4378 · inbound

MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion cites this paper.

MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.386457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T15:11:53.320505Z digest=sha256:6d467ef8ffbf5d982d501c053a44770e4551f563c42fafd479bb3d115f103347

Observation c2d6f366-a499-4e1a-b1db-2a481ae472c7 · inbound

Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control cites this paper.

Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control Zero-shot Voice Conversion with Diffusion Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:28:49.066595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T03:02:52.096117Z digest=sha256:24dca4ade6c1015130f56f2838ca977e5580036f334b0b5520e7e6b8254f19e1

Observation e19300d7-d149-422d-b35a-0fc20c528157 · inbound

Zero-VC: Zero-Lookahead Streaming Voice Conversion via Speaker Anonymization cites this paper.

Zero-VC: Zero-Lookahead Streaming Voice Conversion via Speaker Anonymization Zero-shot Voice Conversion with Diffusion Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:39:39.691633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T15:49:21.097255Z digest=sha256:eca43140452583025fc038937fa19fe559dad0c872adf2e4301b7612a65c63f7

Observation 99764790-3ccd-4382-9982-029011139817 · inbound

Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach cites this paper.

Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach Zero-shot Voice Conversion with Diffusion Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:38.658890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T13:27:42.302206Z digest=sha256:e99a05e133886ca3e842fd8fe69056fbea16a58119e5172dc0f58e4892499ca6

Observation 53d830a9-414e-4786-b458-eedce1f71606 · inbound

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion cites this paper.

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:44.602908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T11:49:04.308326Z digest=sha256:47659a88d616e3e3a5f0dc159127342f5e34c71fae7de638b38bc2b4a23aa14d

Observation bb720f14-37a2-42bd-b2fb-03139d65bf02 · inbound

AugCodec: A Low-Bitrate Disentangled Neural Speech Codec via Data Augmentation cites this paper.

AugCodec: A Low-Bitrate Disentangled Neural Speech Codec via Data Augmentation Zero-shot Voice Conversion with Diffusion Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:41.885749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T11:37:39.960212Z digest=sha256:3dfb48744cbb03d5b38c12b0069e20bd7e7219cb63853f9da558ca269d119c27

Observation 5a422359-a41d-4708-96e6-b045c0e9070f · inbound

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech cites this paper.

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T10:34:43.898772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T10:34:43.898772Z digest=sha256:3f53a5f948549abe1adde88892ecfe47ca2a7e4fb9da974ecfd14368f1954061

Observation bed75580-beb7-4272-b054-acaa075688fb · inbound

Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis cites this paper.

Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis Zero-shot Voice Conversion with Diffusion Transformers

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T06:56:44.305504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T06:36:42.174254Z digest=sha256:c317a5c07b75058c5f965ff38e9479ad2500990c82cb822b5dbff5da55f976b9

Observation 93f6fa2d-b120-4162-88d0-753d170d0aab · inbound

Beyond Words: Towards Effective Modeling of Non-Verbal Vocalizations in ASR cites this paper.

Beyond Words: Towards Effective Modeling of Non-Verbal Vocalizations in ASR Zero-shot Voice Conversion with Diffusion Transformers

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:29.598161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T00:29:39.365107Z digest=sha256:aca101a273436275697f2957a4f8812b909dcf53fba69bbfd8ee7742f6e36c71

Observation 298657dc-1185-40b7-9a76-4246648b1dcc · inbound

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech cites this paper.

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T08:16:37.711326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:16:37.711326Z digest=sha256:455667c3fe4f5cdf2b2323d89bac6eae77b3ae74906b5ac0190bc31043cf9af9

Observation 5f34df43-1869-47d7-b6a6-8d82c20fac19 · inbound

VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion cites this paper.

VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T03:43:46.836705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T03:43:46.836705Z digest=sha256:d437a20617017bed8e93c9932fbf325b463f1a73a2f4378fbc6c1a0ce5af3fea

Observation d2287fb6-a75e-473e-9877-3627d630a0ef · inbound

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors cites this paper.

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors Zero-shot Voice Conversion with Diffusion Transformers

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T22:37:43.616198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:37:43.616198Z digest=sha256:232dfe89053e8ab827f5278dfbf958bcb5c516a2233ee8d8827a8f5915283e22

Observation c19194d9-67a2-41f6-87b9-cd07aaaeab6b · inbound

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors cites this paper.

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors Zero-shot Voice Conversion with Diffusion Transformers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T01:43:19.292962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:43:19.292962Z digest=sha256:39f88238806fd2ccafdf39a63ceaa57ca2812677d6031d6c66814bd856489251