Pith. sign in

Paper Citation Record · LEDGER

Training-Free Voice Conversion with Factorized Optimal Transport

As of 7 August 2026, this Paper Citation Record lists 25 of 25 outbound references and 1 inbound Pith citation observation for arXiv:2506.09709.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09709 v1

Coverage vector

measured 25 of 25 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:46:58.250585Z

measured 26 of 26 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:46:58.165183Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:46:58.296609Z

Reference resolution

25 of 25 outbound references displayed

  • verified exact1
  • verified fuzzy16
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 040479bc-4fb1-427b-bb1f-7fef9fdae6e1 · outbound

This paper cites Training-Free Voice Conversion with Factorized Optimal Transport.

Training-Free Voice Conversion with Factorized Optimal Transport Training-Free Voice Conversion with Factorized Optimal Transport

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:46:58.300939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.165183Z digest=sha256:f2500195ecbfd97c32a7dd1ff1c5b4192add9d747a88551e76fa5884fc36f7b9

Observation 360c8022-259e-4105-8ed0-9cbdb0680580 · outbound

This paper cites analysis- mapping-reconstruction.

Training-Free Voice Conversion with Factorized Optimal Transport analysis- mapping-reconstruction

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.523235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.169685Z digest=sha256:908cf036040f61b398be660e0113114edadb17aca923bcce73cad6c6df27df17

Observation 73d6c292-445b-4c6a-a824-89eadf556b10 · outbound

This paper cites Optimal Transport Maps are Good Voice Converters.

Training-Free Voice Conversion with Factorized Optimal Transport Optimal Transport Maps are Good Voice Converters

Reference 3

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:46:58.284954Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.173480Z digest=sha256:713b9060d11f650d8e41606e37a49e4a13db6411272a3b46963f2a953e394d2e

Observation 8bcd653a-ddc9-48c6-8226-94c308820865 · outbound

This paper cites Datasets To evaluate any speaker to any speaker voice conversion we conduct our experiments on a LibriSpeech dataset [14], which consist of 40 speakers.

Training-Free Voice Conversion with Factorized Optimal Transport Datasets To evaluate any speaker to any speaker voice conversion we conduct our experiments on a LibriSpeech dataset [14], which consist of 40 speakers

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.511920Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.177606Z digest=sha256:d690f4f627edfee9692578bc04e144e942c2b95952075ddf218ac64c7f8eeaf8

Observation f8abe61b-d45d-4d92-85e6-cdfe9069b44c · outbound

This paper cites MKL stands for Monge-Kantorovich Linear map- ping and is based on the exact solution of the quadratic opti- mal transport problem for Gaussian distributions.

Training-Free Voice Conversion with Factorized Optimal Transport MKL stands for Monge-Kantorovich Linear map- ping and is based on the exact solution of the quadratic opti- mal transport problem for Gaussian distributions

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.500750Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.181674Z digest=sha256:c86c8f77940d5405ae565a3e1a29d36c1d5df7d8ee3aa465612dae74007b83da

Observation c3fee13e-6d95-498d-b30d-5cb176e681fb · outbound

This paper cites To prevent misuse, it is crucial to develop robust speech detection methods and avoid voice authentication in high-security applications.

Training-Free Voice Conversion with Factorized Optimal Transport To prevent misuse, it is crucial to develop robust speech detection methods and avoid voice authentication in high-security applications

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.489573Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.185471Z digest=sha256:9acb7a025e6cd2087a897f7ee31200fccce5a56efec020d26bfc3b632f69624f

Observation 5cc02536-5603-45c2-8073-e49e2202bf7a · outbound

This paper cites An overview of voice conversion and its challenges: From statistical modeling to deep learning,.

Training-Free Voice Conversion with Factorized Optimal Transport An overview of voice conversion and its challenges: From statistical modeling to deep learning,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.478706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.189231Z digest=sha256:933a7e4cc26f065ab0a3ae85e88836fe0a768601a0942347367c05d0ada9bf1d

Observation 03efe944-c0f5-4f77-899c-913b8f637817 · outbound

This paper cites V oice conversion with just nearest neighbors,.

Training-Free Voice Conversion with Factorized Optimal Transport V oice conversion with just nearest neighbors,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.467170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.192739Z digest=sha256:2a5805d2b6b0a93e9778a4b2afc428affe097386544eb63e394f6dec2598d8ca

Observation 15d04ac7-e848-4a4a-b116-1f16a1d0e708 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Training-Free Voice Conversion with Factorized Optimal Transport Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.196130Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.196130Z digest=sha256:3d3beff8b806c1b69a48110e28673da8ec073a776f912631a6f8dac01764a86f

Observation 8e17d4bb-7b97-44f3-8607-9b8ff4e6d14a · outbound

This paper cites Overview of voice conversion methods based on deep learning,.

Training-Free Voice Conversion with Factorized Optimal Transport Overview of voice conversion methods based on deep learning,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.450108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.199652Z digest=sha256:deae8687dba3ea291bc60d6f936ecfdea1323da5ed3b8f43827815b5afce99e9

Observation f5bfd163-ef5e-479c-b208-0f64fbd446b6 · outbound

This paper cites Naturalspeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,.

Training-Free Voice Conversion with Factorized Optimal Transport Naturalspeech 3: Zero-shot speech syn- thesis with factorized codec and diffusion models,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.438863Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.203015Z digest=sha256:5db2368438069c2519dd89df7d9df6366903a1a4064426699885ca57c103fe13

Observation 01d2ddd3-d4ae-4af4-8562-51d50b27117c · outbound

This paper cites Freevc: Towards high-quality text-free one-shot voice conversion,.

Training-Free Voice Conversion with Factorized Optimal Transport Freevc: Towards high-quality text-free one-shot voice conversion,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.427879Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.206415Z digest=sha256:8c359367dfb7581c4404fcd484db764da345c1c31373f7655c12a8e468a21ac1

Observation cf068a00-5400-42b7-9bf0-71432c187458 · outbound

This paper cites Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,.

Training-Free Voice Conversion with Factorized Optimal Transport Yourtts: Towards zero-shot multi-speaker tts and zero-shot voice conversion for everyone,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.209925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.209925Z digest=sha256:35dc9bf3bdc4da63f5fd9fe0e9e488f4ee5d0a9d97ac921b12c63aba841bf856

Observation bf7e626e-ec35-4540-b765-de1997883652 · outbound

This paper cites Diffusion-based voice conversion with fast maxi- mum likelihood sampling scheme,.

Training-Free Voice Conversion with Factorized Optimal Transport Diffusion-based voice conversion with fast maxi- mum likelihood sampling scheme,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.409817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.213250Z digest=sha256:2b7159be87d80cd72fe8c16459ca5f8ecf744144d35c914e3cc11699b9644c1a

Observation 80a1c7c7-9caf-4141-a0aa-f18e494756f1 · outbound

This paper cites Vqmivc: Vector quantization and mutual information-based un- supervised speech representation disentanglement for one-shot voice conversion,.

Training-Free Voice Conversion with Factorized Optimal Transport Vqmivc: Vector quantization and mutual information-based un- supervised speech representation disentanglement for one-shot voice conversion,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.399176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.216996Z digest=sha256:6c8bd4c123b61b40dd3a444f452acf9e879d2a71dc76ba4f0fc78292055b45d1

Observation ed234397-fd74-426d-ac2f-34b63a62122d · outbound

This paper cites Synthesizing a choir in real-time using pitch synchronous over- lap add (psola).

Training-Free Voice Conversion with Factorized Optimal Transport Synthesizing a choir in real-time using pitch synchronous over- lap add (psola)

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.388146Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.220290Z digest=sha256:f7133ab8ec913b58204cbab57277b29bbe399d059d7217dd029e486ac73d1432

Observation fa222018-5511-4844-9a06-a4925f9b3950 · outbound

This paper cites Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,.

Training-Free Voice Conversion with Factorized Optimal Transport Hubert: Self-supervised speech represen- tation learning by masked prediction of hidden units,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.223576Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.223576Z digest=sha256:eaa69cee7462cc1a4a5620613a2f8e1ad7ca1a1910d389ffc43123ba8b1b66c0

Observation ca7b0d54-dcbc-45ee-9077-35609d1e7e7d · outbound

This paper cites Sinkhorn distances: Lightspeed computation of op- timal transport,.

Training-Free Voice Conversion with Factorized Optimal Transport Sinkhorn distances: Lightspeed computation of op- timal transport,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.371130Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.226853Z digest=sha256:b68af5489abf1a1135ded688221434e00844c3808e3d727324defcf15ac1b446

Observation 7a1909ea-d35a-46c2-bd93-8bca21ef8fa2 · outbound

This paper cites Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,.

Training-Free Voice Conversion with Factorized Optimal Transport Hifi-gan: Generative adversarial net- works for efficient and high fidelity speech synthesis,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.230241Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.230241Z digest=sha256:aa23f3de900a4f6b2476ed588d222416b107d87c2d0abc6e604fa0fb28909c2f

Observation 516dea9b-38ca-45e5-b662-3994b5bd404f · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Training-Free Voice Conversion with Factorized Optimal Transport Lib- rispeech: an asr corpus based on public domain audio books,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.233596Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.233596Z digest=sha256:98839701e7905f87ab66d8e6561eacdb3a1a299414852fca530a1506bdea9feb

Observation a5626f5b-bc48-49e9-982c-e8c3d65a2e77 · outbound

This paper cites The linear monge-kantorovitch linear colour mapping for example-based colour transfer,.

Training-Free Voice Conversion with Factorized Optimal Transport The linear monge-kantorovitch linear colour mapping for example-based colour transfer,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.347565Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.237365Z digest=sha256:d87387f1ae3b5f3eaa1d16038a203904ed8df776c8046463ddbfba2a30abd2b5

Observation 136742ef-6804-4aec-b0ed-fb867687741d · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

Training-Free Voice Conversion with Factorized Optimal Transport Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.240867Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.240867Z digest=sha256:1d7bc10aad4e4e9f9597657622c622f0119394ea8cd74a61817e7deda236e2c7

Observation d08e3d2c-0d1d-4693-a9db-2cb833fbcf80 · outbound

This paper cites Spoken language recognition using x-vectors.

Training-Free Voice Conversion with Factorized Optimal Transport Spoken language recognition using x-vectors

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.329876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.244049Z digest=sha256:e31972d54a6a4853a5f9e02a847bcd3e77d4406674f23d881f4566d36b3d50a7

Observation 0627924d-96b0-444e-ad3d-de298300a310 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Training-Free Voice Conversion with Factorized Optimal Transport Robust speech recognition via large-scale weak supervision,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T04:46:58.247356Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:46:58.247356Z digest=sha256:5a17105e9441ecc13e37f0aaa4be2d3e4d763be773d4007f41e8fa4c1b746c9c

Observation 36b41260-1266-4bca-99de-4a7b79dfdc52 · outbound

This paper cites Moseley,Atlas of the World’s Languages in Danger.

Training-Free Voice Conversion with Factorized Optimal Transport Moseley,Atlas of the World’s Languages in Danger

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:46:58.311979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.250585Z digest=sha256:850ef98c6c3e6a3ecb2f8132cd9e1c2114373b382cfe436a982143a58bbb3d45

Pith citing papers

Observation 040479bc-4fb1-427b-bb1f-7fef9fdae6e1 · inbound

Training-Free Voice Conversion with Factorized Optimal Transport cites this paper.

Training-Free Voice Conversion with Factorized Optimal Transport Training-Free Voice Conversion with Factorized Optimal Transport

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:46:58.300939Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T04:46:58.165183Z digest=sha256:f2500195ecbfd97c32a7dd1ff1c5b4192add9d747a88551e76fa5884fc36f7b9