Pith. sign in

Paper Citation Record · LEDGER

Training-Free Voice Conversion with Factorized Optimal Transport

As of 10 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2506.09709.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09709 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:46:58.250585Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:46:58.165183Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:46:58.296609Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 040479bc-4fb1-427b-bb1f-7fef9fdae6e1 · outbound

This paper cites Training-Free Voice Conversion with Factorized Optimal Transport.

Training-Free Voice Conversion with Factorized Optimal Transport Training-Free Voice Conversion with Factorized Optimal Transport

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:46:58.300939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.165183Z digest=sha256:66959d3d143b64b166886c1145446887813419c8e51abea2021cf18f6840d107

Observation 360c8022-259e-4105-8ed0-9cbdb0680580 · outbound

This paper cites analysis- mapping-reconstruction.

Training-Free Voice Conversion with Factorized Optimal Transport analysis- mapping-reconstruction

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.523235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.169685Z digest=sha256:471984c6c4f3c0d6432f1283aaac31bd376a1a9545c972fcbba9f561d7dcf495

Observation 73d6c292-445b-4c6a-a824-89eadf556b10 · outbound

This paper cites Optimal Transport Maps are Good Voice Converters.

Training-Free Voice Conversion with Factorized Optimal Transport Optimal Transport Maps are Good Voice Converters

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:46:58.284954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.173480Z digest=sha256:2974a33533d8f9090ef5cdc1943965144a6b1956b3bd6e8dd3b2160b28a6a0d4

Observation 8bcd653a-ddc9-48c6-8226-94c308820865 · outbound

This paper cites Datasets To evaluate any speaker to any speaker voice conversion we conduct our experiments on a LibriSpeech dataset [14], which consist of 40 speakers.

Training-Free Voice Conversion with Factorized Optimal Transport Datasets To evaluate any speaker to any speaker voice conversion we conduct our experiments on a LibriSpeech dataset [14], which consist of 40 speakers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.511920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.177606Z digest=sha256:2d34a40b87c0014be7d0ffc408d32ce3d8516ed30ee3e694e3c929386ed41faf

Observation f8abe61b-d45d-4d92-85e6-cdfe9069b44c · outbound

This paper cites MKL stands for Monge-Kantorovich Linear map- ping and is based on the exact solution of the quadratic opti- mal transport problem for Gaussian distributions.

Training-Free Voice Conversion with Factorized Optimal Transport MKL stands for Monge-Kantorovich Linear map- ping and is based on the exact solution of the quadratic opti- mal transport problem for Gaussian distributions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.500750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.181674Z digest=sha256:fb1cb3e7757bddf5aaf4a97c48036a9b5cdb74a045806ca1aa49b96bba7f62a0

Observation c3fee13e-6d95-498d-b30d-5cb176e681fb · outbound

This paper cites To prevent misuse, it is crucial to develop robust speech detection methods and avoid voice authentication in high-security applications.

Training-Free Voice Conversion with Factorized Optimal Transport To prevent misuse, it is crucial to develop robust speech detection methods and avoid voice authentication in high-security applications

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.489573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.185471Z digest=sha256:3aaf0e7857b1c5aed05aefee30e98d57f5d7a94d87545d2adcdb319df287a613

Observation 5cc02536-5603-45c2-8073-e49e2202bf7a · outbound

This paper cites An overview of voice conversion and its challenges: From statistical modeling to deep learning,.

Training-Free Voice Conversion with Factorized Optimal Transport An overview of voice conversion and its challenges: From statistical modeling to deep learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.478706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.189231Z digest=sha256:6ea38c7a9d018f514c504d228f66f536e860cdee84412c56c096f135ab4d0036

Observation 03efe944-c0f5-4f77-899c-913b8f637817 · outbound

This paper cites V oice conversion with just nearest neighbors,.

Training-Free Voice Conversion with Factorized Optimal Transport V oice conversion with just nearest neighbors,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.467170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.192739Z digest=sha256:e8e794da56a75e3fd3194aee668555d90dd198bbcf97dd21258690cf36ef1891

Observation 15d04ac7-e848-4a4a-b116-1f16a1d0e708 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Training-Free Voice Conversion with Factorized Optimal Transport Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.196130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.196130Z digest=sha256:3d3beff8b806c1b69a48110e28673da8ec073a776f912631a6f8dac01764a86f

Observation 8e17d4bb-7b97-44f3-8607-9b8ff4e6d14a · outbound

This paper cites Overview of voice conversion methods based on deep learning,.

Training-Free Voice Conversion with Factorized Optimal Transport Overview of voice conversion methods based on deep learning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.450108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.199652Z digest=sha256:abec5a2337d6f82a7609192920d19c1c33b2a1707da431481b32bb814851ee8e

Observation f5bfd163-ef5e-479c-b208-0f64fbd446b6 · outbound

This paper cites Naturalspeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,.

Training-Free Voice Conversion with Factorized Optimal Transport Naturalspeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.438863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.203015Z digest=sha256:51aa0c2d42136be8c49f1c7c71e092f79d4dd17fece2c3a47541d3491a8eee51

Observation 01d2ddd3-d4ae-4af4-8562-51d50b27117c · outbound

This paper cites Freevc: Towards high-quality text-free one-shot voice conversion,.

Training-Free Voice Conversion with Factorized Optimal Transport Freevc: Towards high-quality text-free one-shot voice conversion,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.427879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.206415Z digest=sha256:d4d9781df06e28a7e1c332fe559523b806fc32665e193cf3e5cfad57597a51a1

Observation cf068a00-5400-42b7-9bf0-71432c187458 · outbound

This paper cites Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,.

Training-Free Voice Conversion with Factorized Optimal Transport Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.209925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.209925Z digest=sha256:35dc9bf3bdc4da63f5fd9fe0e9e488f4ee5d0a9d97ac921b12c63aba841bf856

Observation bf7e626e-ec35-4540-b765-de1997883652 · outbound

This paper cites Diffusion-based voice conversion with fast maxi- mum likelihood sampling scheme,.

Training-Free Voice Conversion with Factorized Optimal Transport Diffusion-based voice conversion with fast maxi- mum likelihood sampling scheme,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.409817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.213250Z digest=sha256:bc35b9910ebefd36ad2d9c422a55dc1bbf2ddbd46f03e6278d9f4f63c4e40b6d

Observation 80a1c7c7-9caf-4141-a0aa-f18e494756f1 · outbound

This paper cites Vqmivc: Vector quantization and mutual information-based un- supervised speech representation disentanglement for one-shot voice conversion,.

Training-Free Voice Conversion with Factorized Optimal Transport Vqmivc: Vector quantization and mutual information-based un- supervised speech representation disentanglement for one-shot voice conversion,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.399176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.216996Z digest=sha256:04c078c0b82c19b88fd4c7bf02be48aedc55548d439fde76cf53f1f8bc6cafe3

Observation ed234397-fd74-426d-ac2f-34b63a62122d · outbound

This paper cites Synthesizing a choir in real-time using pitch synchronous over- lap add (psola).

Training-Free Voice Conversion with Factorized Optimal Transport Synthesizing a choir in real-time using pitch synchronous over- lap add (psola)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.388146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.220290Z digest=sha256:46d276d92a487a74e5e75172a6fa039ff1a9793104d4b256158a98a3bc4ae362

Observation fa222018-5511-4844-9a06-a4925f9b3950 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

Training-Free Voice Conversion with Factorized Optimal Transport Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.223576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.223576Z digest=sha256:eaa69cee7462cc1a4a5620613a2f8e1ad7ca1a1910d389ffc43123ba8b1b66c0

Observation ca7b0d54-dcbc-45ee-9077-35609d1e7e7d · outbound

This paper cites Sinkhorn distances: Lightspeed computation of op- timal transport,.

Training-Free Voice Conversion with Factorized Optimal Transport Sinkhorn distances: Lightspeed computation of op- timal transport,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.371130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.226853Z digest=sha256:403e230e360018d4ed589ac1d1f187ef360c48bc6ed9757cfb608c508cc9c13e

Observation 7a1909ea-d35a-46c2-bd93-8bca21ef8fa2 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Training-Free Voice Conversion with Factorized Optimal Transport Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.230241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.230241Z digest=sha256:aa23f3de900a4f6b2476ed588d222416b107d87c2d0abc6e604fa0fb28909c2f

Observation 516dea9b-38ca-45e5-b662-3994b5bd404f · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Training-Free Voice Conversion with Factorized Optimal Transport Lib- rispeech: an asr corpus based on public domain audio books,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.233596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.233596Z digest=sha256:98839701e7905f87ab66d8e6561eacdb3a1a299414852fca530a1506bdea9feb

Observation a5626f5b-bc48-49e9-982c-e8c3d65a2e77 · outbound

This paper cites The linear monge-kantorovitch linear colour mapping for example-based colour transfer,.

Training-Free Voice Conversion with Factorized Optimal Transport The linear monge-kantorovitch linear colour mapping for example-based colour transfer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.347565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.237365Z digest=sha256:42600da9edc9b094cb669fd8bb22dac6efe1b6cc1bf46e70120a517b5c97e679

Observation 136742ef-6804-4aec-b0ed-fb867687741d · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Training-Free Voice Conversion with Factorized Optimal Transport Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.240867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.240867Z digest=sha256:1d7bc10aad4e4e9f9597657622c622f0119394ea8cd74a61817e7deda236e2c7

Observation d08e3d2c-0d1d-4693-a9db-2cb833fbcf80 · outbound

This paper cites Spoken language recognition using x-vectors.

Training-Free Voice Conversion with Factorized Optimal Transport Spoken language recognition using x-vectors

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.329876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.244049Z digest=sha256:5e2c51f25eae43e1c0c9a17798a1e420acda66656d9d3f21c96cfc4e498c9242

Observation 0627924d-96b0-444e-ad3d-de298300a310 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Training-Free Voice Conversion with Factorized Optimal Transport Robust speech recognition via large-scale weak supervision,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.247356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.247356Z digest=sha256:5a17105e9441ecc13e37f0aaa4be2d3e4d763be773d4007f41e8fa4c1b746c9c

Observation 36b41260-1266-4bca-99de-4a7b79dfdc52 · outbound

This paper cites Moseley,Atlas of the World’s Languages in Danger.

Training-Free Voice Conversion with Factorized Optimal Transport Moseley,Atlas of the World’s Languages in Danger

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.311979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.250585Z digest=sha256:9248cd2cc42a308c18dfd099f970edbfe56c64a18b05dbdf6ef085e05cd5d8a9

Pith citing papers

Observation 040479bc-4fb1-427b-bb1f-7fef9fdae6e1 · inbound

Training-Free Voice Conversion with Factorized Optimal Transport cites this paper.

Training-Free Voice Conversion with Factorized Optimal Transport Training-Free Voice Conversion with Factorized Optimal Transport

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:46:58.300939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T04:46:58.165183Z digest=sha256:66959d3d143b64b166886c1145446887813419c8e51abea2021cf18f6840d107