Pith. sign in

Paper Citation Record · LEDGER

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data

As of 8 August 2026, this Paper Citation Record lists 30 of 30 outbound references and 0 inbound Pith citation observations for arXiv:2506.01039.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.01039 v1

Coverage vector

measured 30 of 30 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:56:27.059279Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

30 of 30 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation ec237109-970f-47f6-b7dd-e23055e5d808 · outbound

This paper cites Dddm-vc: Decoupled denoising diffusion models with disentangled representation and prior mixup for verified robust voice conversion,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Dddm-vc: Decoupled denoising diffusion models with disentangled representation and prior mixup for verified robust voice conversion,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:24.465301Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:24.465301Z digest=sha256:212206d2dddd3d9097f35bd57172693eaea84c915ea9d40acaaabf6e13a9cde2

Observation aa799c4c-0577-429c-b2e1-c1f83bcb40d4 · outbound

This paper cites Freevc: Towards high-quality text-free one-shot voice conversion,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Freevc: Towards high-quality text-free one-shot voice conversion,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:24.555184Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:24.555184Z digest=sha256:f9736cc23216481516e9224259d43c737ac002a5e4521df0b5f81ea62e8d8b24

Observation 2cf0268e-ad31-4557-aaf9-845817045e3f · outbound

This paper cites Phoneme hallucinator: One-shot voice conversion via set expansion,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Phoneme hallucinator: One-shot voice conversion via set expansion,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:29.188342Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:24.694732Z digest=sha256:5a09b69936304d0cceddc299c8f57f87c22995580bb685810a6f36c08fb75fd9

Observation 21ed0170-26d5-4d1e-99ae-844419eef796 · outbound

This paper cites Voice Conversion With Just Nearest Neighbors.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Voice Conversion With Just Nearest Neighbors

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:24.801370Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:24.801370Z digest=sha256:d7b65d0726b7d0193a45045a36103dca8fa100e4a25209dee523110974668280

Observation a841bcd7-b621-497f-bb70-fedf9f440417 · outbound

This paper cites Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:24.880991Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:24.880991Z digest=sha256:374245c7a91b48fa3a23cd47bb6148b3b6d0d523f7e33f679403bc01a1080abd

Observation 6c9caed6-2b3e-4319-a1ad-1ae03efe49d1 · outbound

This paper cites Gr0: Self-supervised global representation learning for zero-shot voice conversion,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Gr0: Self-supervised global representation learning for zero-shot voice conversion,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:29.045835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:24.953021Z digest=sha256:1240f058c8fc0dbc2d5cc025d68130fd8d8e7767dcb37e882f7d3940865cb900

Observation e618572b-d283-4f48-b92b-599cf5aeef00 · outbound

This paper cites Parallel-Data-Free Voice Conversion Using Cycle-Consistent Adversarial Networks.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Parallel-Data-Free Voice Conversion Using Cycle-Consistent Adversarial Networks

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.030407Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.030407Z digest=sha256:aa731994dfd4648a7c988942966736175b95886d5cd1bcb6b73c66f522686ac3

Observation bb01ac4e-1e61-4f83-b227-f6c69dc99d1c · outbound

This paper cites Average modeling approach to voice conversion with non-parallel data.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Average modeling approach to voice conversion with non-parallel data

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.923320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:25.106397Z digest=sha256:99f3acec6d5a71d5678d421fe8d4c5adbcb22cbd5d4dd2b48ae5565c60a05890

Observation a74ab71f-29cb-4ccd-be8e-e7bead416de7 · outbound

This paper cites The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data The Voice Conversion Challenge 2018: Promoting Development of Parallel and Nonparallel Methods

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.228067Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.228067Z digest=sha256:a87e4f317a98f2b0fd1ffb660053b52338a571a5b2946ee379807bb3e4eaeb08

Observation 9fa78162-ec8f-4773-aadc-9d3c006be38f · outbound

This paper cites Unsupervised speech decomposition via triple information bottleneck,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Unsupervised speech decomposition via triple information bottleneck,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.301193Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.301193Z digest=sha256:04785d7e67aa7e873776cd1eea258ecdd4ccf83fb7ca55401d8dd045e1c2a37a

Observation 97e3dbdf-0d66-4d08-8000-2dc26cc546ef · outbound

This paper cites Speech- split2. 0: Unsupervised speech disentanglement for voice conversion without tuning autoencoder bottlenecks,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Speech- split2. 0: Unsupervised speech disentanglement for voice conversion without tuning autoencoder bottlenecks,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.377373Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.377373Z digest=sha256:ff4b90dfce0d95dec963b4e1187118962aff0c49d0412d76115c055d8d39476b

Observation a5368fe6-2c6d-4ebf-b3b4-33d07480cc49 · outbound

This paper cites Neural analysis and synthesis: Reconstructing speech from self-supervised representations,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Neural analysis and synthesis: Reconstructing speech from self-supervised representations,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.751386Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:25.467236Z digest=sha256:dca4e32ed55762af3b1b4b047a4cfd03612760573a48cc5462561ef303b8f0bc

Observation df2a8189-7adf-417e-b903-49b0e5261b38 · outbound

This paper cites Scheduled sampling for sequence prediction with recurrent neural networks,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Scheduled sampling for sequence prediction with recurrent neural networks,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.566949Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.566949Z digest=sha256:b65abaedbf91692164eab2db94aa8115cbdf3b4a7547818ed7c7a32ca7252338

Observation c9f1e1dc-efae-458e-9cf7-1d534460606f · outbound

This paper cites V ocal tract length perturbation (vtlp) improves speech recognition,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data V ocal tract length perturbation (vtlp) improves speech recognition,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.625637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:25.652713Z digest=sha256:1c4e0df53acc98e40be005d40ceadef7f2288599e41614e5a55ff66e87ea66c6

Observation b002e381-9818-43ce-b6d0-d9d18d099d60 · outbound

This paper cites Wavlm: Large-scale self-supervised pre- training for full stack speech processing,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Wavlm: Large-scale self-supervised pre- training for full stack speech processing,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.744123Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.744123Z digest=sha256:7d726e09223231f4faafb9c20f7a9e7bccdac8ca5295ad019de6bc1789659d77

Observation 62a9ee90-5550-4369-9c00-725b8d1d1534 · outbound

This paper cites Variational inference with normalizing flows,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Variational inference with normalizing flows,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:25.836802Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:25.836802Z digest=sha256:eb3a37c6ae6551a35da29c492393711680f9431f223721ee249b59f2481602c4

Observation 51ee43f1-27b0-473e-a4d8-c6a8dd2f530b · outbound

This paper cites Any-to-many voice conversion with location-relative sequence-to-sequence modeling,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Any-to-many voice conversion with location-relative sequence-to-sequence modeling,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.437664Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:25.901957Z digest=sha256:46915216eb6e15d8146601bc9f5b8618eae15d67aab5634f0e63dfd331cbbb84

Observation 82763542-df3f-45d8-9c7f-415dd85f4828 · outbound

This paper cites Waveglow: A flow-based generative network for speech synthesis,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Waveglow: A flow-based generative network for speech synthesis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.271208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:25.978705Z digest=sha256:9272d956b06ad1371f2181ef5e693bfaf0924938c75395434d56a2c4773672fb

Observation 018c4897-5729-4d45-9073-8c49d6e41fa2 · outbound

This paper cites Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Hifi-gan: Generative adversarial networks for efficient and high fidelity speech synthesis,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:26.046923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:26.046923Z digest=sha256:95b09bb40b300b0ef96785844aaa8964b30bedb58192278cf13875be119a26a0

Observation 42de39c4-66c9-45cd-84c8-42a95ff22337 · outbound

This paper cites Least squares generative adversarial networks,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Least squares generative adversarial networks,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:28.096705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.123139Z digest=sha256:e0504359bd3f6dbbf8f37fa14b1e30841050d787ea00edae2fd85958c738ce3a

Observation de7abcd3-82cd-4cfc-afd7-1e13eafb2ca4 · outbound

This paper cites Autoencoding beyond pixels using a learned similarity metric,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Autoencoding beyond pixels using a learned similarity metric,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:27.925838Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.218072Z digest=sha256:c036229ca72d3684a509cbd16843af14e0556d6bab154d272e102e4fba09f6ee

Observation fd45b1a7-e236-41df-b5c3-1a57c3102ecc · outbound

This paper cites Fixmatch: Simplifying semi- supervised learning with consistency and confidence,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Fixmatch: Simplifying semi- supervised learning with consistency and confidence,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:26.312898Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:26.312898Z digest=sha256:9edfe441838827e1a07e5814a427225362d58681c0f9f761c818bc1a9d919486

Observation 309723d9-36fc-4ac3-a246-1c2e29ef50bf · outbound

This paper cites Censer: Curriculum Semi-supervised Learning for Speech Recognition Based on Self-supervised Pre-training.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Censer: Curriculum Semi-supervised Learning for Speech Recognition Based on Self-supervised Pre-training

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:56:27.291844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.384058Z digest=sha256:f6c655749cb813024a0a093049d77b154bc9a4119a083b6b7ab1c64b06690592

Observation 554fc953-cfb4-4876-b665-384c0db07d8f · outbound

This paper cites Pseudo-labeling and confirmation bias in deep semi-supervised learn- ing,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Pseudo-labeling and confirmation bias in deep semi-supervised learn- ing,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:27.765446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.464619Z digest=sha256:c3e921a561c69f2bf5bfe321b7f506aa679b53958ccddda8dd8acf2e0c4aaab8

Observation 76ba54b3-3367-45a6-a649-d74185514532 · outbound

This paper cites Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Cstr vctk corpus: English multi-speaker corpus for cstr voice cloning toolkit,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:27.625854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.558249Z digest=sha256:c621e5454053023284bfcd351fc7dae93efb6ab15447ccc016bce5d8b90747e6

Observation 46b180ff-2452-4051-9bfd-b5c6ee81e47e · outbound

This paper cites LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data LibriTTS: A Corpus Derived from LibriSpeech for Text-to-Speech

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:26.654340Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:26.654340Z digest=sha256:0439b4706ec513f68466e96d0484df914908331737b528c6e3fe21f0d74b6c68

Observation 8cc886d8-07ab-4b03-b08c-7f4964580c39 · outbound

This paper cites Amphion: An Open-Source Audio, Music and Speech Generation Toolkit.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Amphion: An Open-Source Audio, Music and Speech Generation Toolkit

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:26.750676Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:26.750676Z digest=sha256:d78179c52a9b8cf69218e144a8f002410d288383bb656c62d3e7fda69fe76fc8

Observation a0fad8fe-4b1f-4c91-9637-8faa99945396 · outbound

This paper cites Robust speech recognition via large-scale weak supervi- sion,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Robust speech recognition via large-scale weak supervi- sion,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:26.847623Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:26.847623Z digest=sha256:72bf4db352b1c37772d52da1275ba8e4074916ac348b8eff9e84a8c312a9c06b

Observation 482ae3e6-2175-434e-9ced-f39c711f6f46 · outbound

This paper cites Generalized end-to-end loss for speaker verification,.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Generalized end-to-end loss for speaker verification,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:56:27.486257Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:56:26.968831Z digest=sha256:3f191172bbaaba82fa6d2ae7038d6017c8ffaa623388c6ea52462c5222151dc4

Observation fa5ba85e-aa72-48dd-a113-b57147d68185 · outbound

This paper cites Visualizing data using t-sne.

PseudoVC: Improving One-shot Voice Conversion with Pseudo Paired Data Visualizing data using t-sne

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T11:56:27.059279Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:56:27.059279Z digest=sha256:6623b9af66d12050ba34101899bd1e3d2f30766ffb557a2edc75cf8a3038bc4f

Pith citing papers

No inbound Pith citation observations are available.