Pith. sign in

Paper Citation Record · LEDGER

Zero-shot Voice Conversion with Diffusion Transformers

As of 10 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 34 inbound Pith citation observations for arXiv:2411.09943.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.09943 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 34 of 34 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:58:57.726231Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T08:29:41.883949Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation e4e2bbb6-2e69-4f3a-8511-830be7742da8 · inbound

Kimi-Audio Technical Report cites this paper.

Kimi-Audio Technical Report Zero-shot Voice Conversion with Diffusion Transformers

Reference 46

Resolution
verified exact
arxiv_id, observed 2026-05-11T19:21:27.252453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-11T19:21:26.933349Z digest=sha256:765d3ff310f27af08a2f263ddb600aca22f739d5f5f379954e3a4f218f8ff249

Observation 0b9c0f93-9ca5-4759-9b6e-400d1b5d827d · inbound

EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion cites this paper.

EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T14:58:57.726231Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:58:57.726231Z digest=sha256:6068beb801089a364c17f6667e6c56cfbf2a814663e6226a026fde58ae3c50cb

Observation d4926be1-b5a6-4c69-9098-2f81b5e375b4 · inbound

IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech cites this paper.

IndexTTS2: A Breakthrough in Emotionally Expressive and Duration-Controlled Auto-Regressive Zero-Shot Text-to-Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-06T23:21:56.314992Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T23:21:56.314992Z digest=sha256:2ee9e067144b77d463a45fd0173ecedc4e09c185d68ec5622af07a55e8e4ddab

Observation 9f9872d1-24ce-41d6-8e04-0265e059ac42 · inbound

De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks cites this paper.

De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks Zero-shot Voice Conversion with Diffusion Transformers

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-06T20:29:21.074104Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T20:29:21.074104Z digest=sha256:2b397331f869a86907f0cf4797a316239a4c6afb5b9785ca1fb1b9beb363eaf8

Observation 6fcb33aa-9703-4594-a6f5-c5c54c7c1d50 · inbound

The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents cites this paper.

The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents Zero-shot Voice Conversion with Diffusion Transformers

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-06T17:45:45.111222Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:45:45.111222Z digest=sha256:5067c4b55f8fa04f3920567940bca45f42bff29dfda7e461176907995f556193

Observation fb5a0ba8-41c0-4959-a554-c5bd674b3d0c · inbound

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods cites this paper.

SpeechFake: A Large-Scale Multilingual Speech Deepfake Dataset Incorporating Cutting-Edge Generation Methods Zero-shot Voice Conversion with Diffusion Transformers

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T12:49:19.452862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T12:49:19.452862Z digest=sha256:d4a1c71a4d28f9ed273ca0accad4951191d21a042e7311ce7445314162a2ef46

Observation b728c6e2-5022-4993-a1fa-bde3f00b2e67 · inbound

REF-VC: Robust, Expressive and Fast Zero-Shot Voice Conversion with Diffusion Transformers cites this paper.

REF-VC: Robust, Expressive and Fast Zero-Shot Voice Conversion with Diffusion Transformers Zero-shot Voice Conversion with Diffusion Transformers

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T23:41:51.071923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T23:41:51.071923Z digest=sha256:5d0a806656a079e65a9303f7d4100f4b59d8868ec3b5f764a22549d4e3f466c7

Observation ac23beba-29e1-4989-9b7f-6b0eebea3e0b · inbound

Semantic-Aware Ship Detection with Vision-Language Integration cites this paper.

Semantic-Aware Ship Detection with Vision-Language Integration Zero-shot Voice Conversion with Diffusion Transformers

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T17:41:57.979302Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T17:41:57.979302Z digest=sha256:1522ba40b077dc31e5072fe0141c3604ce8ca266b54dbbe39fd444df2c4cd837

Observation 8c260918-7ea5-4d2c-8c26-e641f28fb380 · inbound

Entropy-based Coarse and Compressed Semantic Speech Representation Learning cites this paper.

Entropy-based Coarse and Compressed Semantic Speech Representation Learning Zero-shot Voice Conversion with Diffusion Transformers

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:05.343480Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:05.343480Z digest=sha256:056aef672a706713441da1fd594d124929e5141a40547fcd19e82d904cfc4273

Observation 0efb3374-c69e-4d13-9310-fc9df2c2a120 · inbound

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects cites this paper.

Intelligent Agents with Emotional Intelligence: Current Trends, Challenges, and Future Prospects Zero-shot Voice Conversion with Diffusion Transformers

Reference 95

Resolution
verified exact
arxiv_id, observed 2026-05-18T08:12:30.238340Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-18T08:11:31.181704Z digest=sha256:74512a7745f1859b7b9aa0acf36466bde07047678959c36793c6705d226d644c

Observation a6cdb647-77bb-46ec-b961-3b52975bdb6c · inbound

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis cites this paper.

QASA: Quality-Aware Semantic Augmentation for Robust Multimodal Sentiment Analysis Zero-shot Voice Conversion with Diffusion Transformers

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-03T11:19:33.944804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-03T11:19:33.944804Z digest=sha256:351dfd6b674403c99bf5ee29b1d93fb460b0b144d14af30d9dcdfc028b86b144

Observation c308a249-ce11-45e9-bd90-e015265ee38c · inbound

Universal Speech Content Factorization cites this paper.

Universal Speech Content Factorization Zero-shot Voice Conversion with Diffusion Transformers

Reference 44

Resolution
unresolved
no resolver link, observed 2026-07-15T12:21:48.698333Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-15T12:21:48.698333Z digest=sha256:77518512730fc13403081b4901d8ac81f3c34f2f79df056cf0945e625b4c88e0

Observation 7c8828e4-e183-4b8f-8777-4669fb22fe49 · inbound

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan cites this paper.

AT-ADD: All-Type Audio Deepfake Detection Challenge Evaluation Plan Zero-shot Voice Conversion with Diffusion Transformers

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:21:00.644713Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T17:40:10.657590Z digest=sha256:9dcad00bb60c747ddf6b08816821bd30c7ab93539303a3f8e4459896b2ea9e4b

Observation c463817e-4ee0-4514-900f-962ebbd3c4e3 · inbound

X-VC: Zero-shot Streaming Voice Conversion in Codec Space cites this paper.

X-VC: Zero-shot Streaming Voice Conversion in Codec Space Zero-shot Voice Conversion with Diffusion Transformers

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:20:29.712580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T14:20:01.885544Z digest=sha256:9f84d34493e515025340dfb200c1f580281204592cd6a275ecc080c72e6b0336

Observation 48646aca-6227-401e-acaa-d2dcad05656b · inbound

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction cites this paper.

From Seeing it to Experiencing it: Interactive Evaluation of Intersectional Voice Bias in Human-AI Speech Interaction Zero-shot Voice Conversion with Diffusion Transformers

Reference 18

Resolution
verified exact
arxiv_id, observed 2026-05-15T07:59:50.728859Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-15T07:59:22.416259Z digest=sha256:44d487cf69fe53b5facd59ed6c3fb857595f2413d42983b719deaffbce9311ab

Observation a5ef4e07-2da1-43d8-94f8-0df1aa567665 · inbound

How Far Are Video Models from True Multimodal Reasoning? cites this paper.

How Far Are Video Models from True Multimodal Reasoning? Zero-shot Voice Conversion with Diffusion Transformers

Reference 44

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T12:51:04.023382Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-10T02:44:52.920816Z digest=sha256:d4244ee0cec83a56b797930b487336ec92892a02e1ec1833b45fd15eb744bfec

Observation 2ff6df34-4f54-4a2a-b6f1-6ff99c44e7b2 · inbound

RTCFake: Speech Deepfake Detection in Real-Time Communication cites this paper.

RTCFake: Speech Deepfake Detection in Real-Time Communication Zero-shot Voice Conversion with Diffusion Transformers

Reference 16

Resolution
verified exact
arxiv_id, observed 2026-05-11T21:26:17.872375Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-05-08T05:29:45.895667Z digest=sha256:fbcaa42ca0847c573ad2301499814bf3ebcf65329e2510fd79146be9371f81fa

Observation 854b3826-4e6a-4e10-8493-9af9f355a1ac · inbound

Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling cites this paper.

Poly-SVC: Polyphony-Aware Singing Voice Conversion with Harmonic Modeling Zero-shot Voice Conversion with Diffusion Transformers

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-13T04:12:13.878150Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-05-13T04:09:48.737629Z digest=sha256:f363dbec372da4e47d850802d94225acead02bb6f937d7acba7e4335760248e1

Observation 0ea1d867-762f-42d1-b092-b9539be5185b · inbound

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing cites this paper.

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing Zero-shot Voice Conversion with Diffusion Transformers

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T01:06:23.865289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=arxiv_source observed=2026-06-28T12:54:20.815371Z digest=sha256:8b97de95efcb5e308c90f0015d5d69247b2c4323e2f24e0f2568d66ba90295c8

Observation dfd36285-41fe-4150-8876-fa50210360b2 · inbound

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data cites this paper.

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data Zero-shot Voice Conversion with Diffusion Transformers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-07-02T23:57:28.663888Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T17:39:36.966812Z digest=sha256:979c5539289bb298d26f8c373236cb2771ad472ab8aabdf720388915828b4cdf

Observation 4c810e7a-48ce-47fc-9a6f-d83682abe6fe · inbound

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data cites this paper.

From A to B to A: Palindromic Zero-Shot Voice Conversion with Non-Parallel Data Zero-shot Voice Conversion with Diffusion Transformers

Reference 21

Resolution
verified exact
arxiv_id, observed 2026-06-30T11:04:37.810024Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-30T10:57:44.041182Z digest=sha256:7068db16c56f689299aa66a374ba54f09aa6ed282ed29587e477e0d2ea3ec652

Observation 4ef1ae2b-b912-4f11-afe5-0354b55e4378 · inbound

MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion cites this paper.

MeanVC 2: Robust Low-Latency Streaming Zero-Shot Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-03T03:37:35.386457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T15:11:53.320505Z digest=sha256:11e13422b9eb4f93671d0a3a4c2606938e1e2f8b1b26fec5c6771881aac965f4

Observation c2d6f366-a499-4e1a-b1db-2a481ae472c7 · inbound

Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control cites this paper.

Vibrato Expression Control for Singing Voice Conversion with Improving Independent Control Zero-shot Voice Conversion with Diffusion Transformers

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T18:28:49.066595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-27T03:02:52.096117Z digest=sha256:63117651fd88ee160914bebc67e512fb783c61ff6a4ff8817b9ed324efbda944

Observation e19300d7-d149-422d-b35a-0fc20c528157 · inbound

Zero-VC: Zero-Lookahead Streaming Voice Conversion via Speaker Anonymization cites this paper.

Zero-VC: Zero-Lookahead Streaming Voice Conversion via Speaker Anonymization Zero-shot Voice Conversion with Diffusion Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T05:39:39.691633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T15:49:21.097255Z digest=sha256:8334ab5a43e1a0ff08e77e237c09bec1a47a00a6fdb83d35b87322d8fcabbe0f

Observation 99764790-3ccd-4382-9982-029011139817 · inbound

Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach cites this paper.

Speaker Identity in Non-Verbal Vocalizations: Conditional Distillation and Mixture of Experts Approach Zero-shot Voice Conversion with Diffusion Transformers

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-04T07:29:38.658890Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T13:27:42.302206Z digest=sha256:65187d6a002e1986d8d26b772251fc83562f5f6d4c5dd0502e2bd6031dc57520

Observation 53d830a9-414e-4786-b458-eedce1f71606 · inbound

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion cites this paper.

ProsoCodec: Prosody-Oriented Speech Codec for Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:19:44.602908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T11:49:04.308326Z digest=sha256:9a69b936803cc9f6b5f036e48d70279620d20d06637009d70826db32b6554a48

Observation bb720f14-37a2-42bd-b2fb-03139d65bf02 · inbound

AugCodec: A Low-Bitrate Disentangled Neural Speech Codec via Data Augmentation cites this paper.

AugCodec: A Low-Bitrate Disentangled Neural Speech Codec via Data Augmentation Zero-shot Voice Conversion with Diffusion Transformers

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-07-04T08:29:41.885749Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-06-26T11:37:39.960212Z digest=sha256:ac226dad298f41b29560a17caddad346262c460315e4e6e51590a25d626a830e

Observation 5a422359-a41d-4708-96e6-b045c0e9070f · inbound

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech cites this paper.

TRACE: Temporal Relationship-Aware Conversational Entrainment Detection in Dyadic Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 19

Resolution
unresolved
no resolver link, observed 2026-07-12T10:34:43.898772Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T10:34:43.898772Z digest=sha256:7c65c31e4fa0c00d8bbb459d4e884180c9fe0c696adf5692ac83cc392eaa8985

Observation bed75580-beb7-4272-b054-acaa075688fb · inbound

Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis cites this paper.

Enhancing Flow Matching with A Unified Guidance Framework for Efficient and Robust Speech Synthesis Zero-shot Voice Conversion with Diffusion Transformers

Reference 20

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T06:56:44.305504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-02T06:36:42.174254Z digest=sha256:9c9e32eb3d20dd619f90fa59e3c308743312bcacc774146b0b2820f434b021e2

Observation 93f6fa2d-b120-4162-88d0-753d170d0aab · inbound

Beyond Words: Towards Effective Modeling of Non-Verbal Vocalizations in ASR cites this paper.

Beyond Words: Towards Effective Modeling of Non-Verbal Vocalizations in ASR Zero-shot Voice Conversion with Diffusion Transformers

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-07-03T00:37:29.598161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-07-03T00:29:39.365107Z digest=sha256:bd62d7f925e1d0a85e265bd2981b39f5b35d0311e5326db44552bc93b25f31e8

Observation 298657dc-1185-40b7-9a76-4246648b1dcc · inbound

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech cites this paper.

GRAFT: Grafted Reference Audio for Fine-grained Pronunciation in Zero-shot Text-to-Speech Zero-shot Voice Conversion with Diffusion Transformers

Reference 32

Resolution
unresolved
no resolver link, observed 2026-07-12T08:16:37.711326Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T08:16:37.711326Z digest=sha256:fd0e5fc2d4bab2933892fd3b671b6f08f588ec4e78b48e7c9b8e9864fdf7b711

Observation 5f34df43-1869-47d7-b6a6-8d82c20fac19 · inbound

VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion cites this paper.

VoxENES 2026: Benchmarking Generalization of Speech Spoofing Detectors Against LLM-Era TTS and Voice Conversion Zero-shot Voice Conversion with Diffusion Transformers

Reference 25

Resolution
unresolved
no resolver link, observed 2026-07-14T03:43:46.836705Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-14T03:43:46.836705Z digest=sha256:87fcaf526e88e65f21100cc1110ed793240602bd5b437ea9e6a009014eb587b3

Observation d2287fb6-a75e-473e-9877-3627d630a0ef · inbound

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors cites this paper.

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors Zero-shot Voice Conversion with Diffusion Transformers

Reference 53

Resolution
unresolved
no resolver link, observed 2026-08-01T22:37:43.616198Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:37:43.616198Z digest=sha256:ab4dcc69b3ef11fa22bdfca7ff9aabf5bd74eb875223f9a374f17222d3066ced

Observation c19194d9-67a2-41f6-87b9-cd07aaaeab6b · inbound

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors cites this paper.

A Geometry-Limited Identification Floor and Its Consequences for Voice-Clone Attribution in Professional Voice Actors Zero-shot Voice Conversion with Diffusion Transformers

Reference 54

Resolution
unresolved
no resolver link, observed 2026-08-04T01:43:19.292962Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-04T01:43:19.292962Z digest=sha256:2d74786706d99acc68748ee1adb7d0e4f4ffa7d4b738ac304f3cb01227bb85b6