Pith. sign in

Paper Citation Record · LEDGER

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan

As of 22 August 2026, this Paper Citation Record lists 11 of 11 outbound references and 1 inbound Pith citation observation for arXiv:2505.09382.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.09382 v2

Coverage vector

measured 11 of 11 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:36:50.884359Z

measured 12 of 12 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-05T17:41:58.200870Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-05T17:42:00.136933Z

Reference resolution

11 of 11 outbound references displayed

  • verified exact1
  • verified fuzzy3
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 2670dec6-8cd7-4176-8fc8-1b9926d0078a · outbound

This paper cites VoxCeleb2: Deep Speaker Recognition.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan VoxCeleb2: Deep Speaker Recognition

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.843933Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.843933Z digest=sha256:d80c9044d58a9ee14157008bedf95805593919cb3071ed47bfab23f59c4f68a0

Observation 50730cd8-b205-4f97-80db-ea4d7c495fd8 · outbound

This paper cites ECAPA-TDNN: emphasized channel attention, propagation and aggregation in TDNN based speaker verification.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan ECAPA-TDNN: emphasized channel attention, propagation and aggregation in TDNN based speaker verification

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.847540Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.847540Z digest=sha256:bafd1f56560b9f087c00d9ba26074844938336459652c49db934b3342fcb274a

Observation 52486e4e-599c-4bda-9ccf-7351b4c4770f · outbound

This paper cites Prompttts: Controllable text-to-speech with text descriptions.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Prompttts: Controllable text-to-speech with text descriptions

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.851595Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.851595Z digest=sha256:771c1a3db087f6424fb870278abbacbd85f0b862f1bbfb5092713e7634423677

Observation 13abaa3f-fe0f-41f2-a3e8-05fba77e68a4 · outbound

This paper cites Naturalspeech 3: Zero-shot speech synthesis with factorized codec and diffusion models.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Naturalspeech 3: Zero-shot speech synthesis with factorized codec and diffusion models

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:36:51.088290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T21:36:50.856024Z digest=sha256:9cff6fc0fbfd28995bfaedbc84df915f26eb83220e00dc09e868520873acf125

Observation c0ee070a-3fb7-401f-80e4-bd782e16883f · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.859855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.859855Z digest=sha256:eed458670c05b4aedcd45aaa520ca585334913bd43524cbd3f3f9a74a61a2412

Observation 4501fd81-4ea7-42c8-8e5b-5750a72d9b17 · outbound

This paper cites Libri-light: A benchmark for asr with limited or no supervision.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Libri-light: A benchmark for asr with limited or no supervision

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:36:51.076140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T21:36:50.864115Z digest=sha256:3d682add8b40ddd1b0d19fb69e758bf0951ff04c3e26381e3771cedbde8fddc2

Observation 9faba2e7-2b25-4f2e-9e47-98ba4114f157 · outbound

This paper cites VoxCeleb: a large-scale speaker identification dataset.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan VoxCeleb: a large-scale speaker identification dataset

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.868048Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.868048Z digest=sha256:c894974c4fa4806a40a3429940817cbf6479be22053ef36bd009bed065a90527

Observation 3d4519cf-3dc2-4ef6-8975-bb57705ed6fd · outbound

This paper cites Voice attribute editing with text prompt.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Voice attribute editing with text prompt

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-15T21:36:51.063925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T21:36:50.872180Z digest=sha256:d22b00572054105fd82cf0ed63f05edcace4ad5957632a2a92eddf0a592ef0e9

Observation 9b97c1a5-2c6d-451c-a44f-7b39339ed603 · outbound

This paper cites Voice Attribute Editing with Text Prompt.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Voice Attribute Editing with Text Prompt

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-08-15T21:36:50.947363Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=arxiv_source observed=2026-08-15T21:36:50.876261Z digest=sha256:4122890807debe9c34eb5186588e3223e9888f8fbc6a870100945938c20adc96

Observation 644ccab7-7e45-4e12-aeef-e0e9337b2fac · outbound

This paper cites Unispeaker: A Unified Approach for Multimodality-driven Speaker Generation.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Unispeaker: A Unified Approach for Multimodality-driven Speaker Generation

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.880260Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.880260Z digest=sha256:7c89706f872aef9daf8c9547541d9c67bbee990614adbda3e26e70f09149160a

Observation 45dc1ae9-5531-4cea-8961-e8c894449367 · outbound

This paper cites Explainable Attribute-Based Speaker Verification.

The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan Explainable Attribute-Based Speaker Verification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T21:36:50.884359Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-15T21:36:50.884359Z digest=sha256:b4ae05004a84ef4bc32ba88c75896e9a7c3bbb0a82290a187e3240ebd88c525b

Pith citing papers

Observation 31548769-3034-4922-bb13-b2bd43b02219 · inbound

Semantic-Aware Ship Detection with Vision-Language Integration cites this paper.

Semantic-Aware Ship Detection with Vision-Language Integration The Voice Timbre Attribute Detection 2025 Challenge Evaluation Plan

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T17:42:00.252042Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.

source=pdf_text observed=2026-08-05T17:41:58.200870Z digest=sha256:1ad95991ab7c4016acbfed6c123f11bac840e97df6a1f03a1a3863cbfeca924f