Pith. sign in

Paper Citation Record · LEDGER

3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

As of 8 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 7 inbound Pith citation observations for arXiv:2306.15354.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2306.15354 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 7 of 7 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 7 of 7 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:43:04.029122Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T09:29:44.199281Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 5ad194ca-b72e-4f4d-9468-e213687d7f26 · inbound

The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition cites this paper.

The Multimodal Information Based Speech Processing (MISP) 2025 Challenge: Audio-Visual Diarization and Recognition 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T15:43:04.029122Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:43:04.029122Z digest=sha256:ca89272310b890d09b103be16f3c2c4ab87cf94b070d894a464a179f396580c2

Observation c2722b4e-f906-42c7-a0ce-6e054e013cac · inbound

Multi-Channel Sequence-to-Sequence Neural Diarization: Experimental Results for The MISP 2025 Challenge cites this paper.

Multi-Channel Sequence-to-Sequence Neural Diarization: Experimental Results for The MISP 2025 Challenge 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T15:04:50.995609Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:04:50.995609Z digest=sha256:451bd99db8bb3fed1abc96c13f2bfd4112189da1943d7dcc22b60e885741f5e5

Observation 05674538-aa95-487b-adc8-d195f47d14ff · inbound

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin cites this paper.

VoxAging: Continuously Tracking Speaker Aging with a Large-Scale Longitudinal Dataset in English and Mandarin 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-07T13:32:32.813536Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:32:32.813536Z digest=sha256:8b0346090a890e37acb2d329e77b690fac20c070875cbc836c70e22eb8c51307

Observation 60f6bdce-a975-4e39-94d4-1a574ca3b838 · inbound

Audio2Tool: Speak, Call, Act -- A Dataset for Benchmarking Speech Tool Use cites this paper.

Audio2Tool: Speak, Call, Act -- A Dataset for Benchmarking Speech Tool Use 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-05-10T07:47:13.060741Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-10T07:32:06.306966Z digest=sha256:97894e76e7acaa8499f2b2f1e19edf6c31488d4427db868bc6a61a834b4a7362

Observation 37562708-62ef-40ff-a7c4-569a13b094c1 · inbound

SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription cites this paper.

SoulX-Transcriber: A Robust End-to-End Framework for Multi-Speaker Speech Transcription 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 9

Resolution
verified exact
arxiv_id, observed 2026-07-02T01:06:24.520029Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-28T12:37:52.271987Z digest=sha256:89c32659710aceed84a8faa3a42fa0c39195c85cae9908ad0bfba818c5fe2e06

Observation a2b513cf-b347-4a4f-8746-82d1435e9877 · inbound

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification cites this paper.

Kiwano: A Cutting-Edge Open-Source Toolkit for Speaker Verification 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 32

Resolution
verified exact
arxiv_id, observed 2026-07-04T09:29:44.201695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T09:52:58.661879Z digest=sha256:5481519b7c370a0ca2a5c1f336604369681d8c7d68702392a23a50c163320533

Observation 505248d2-3a49-4223-b368-65e25321864d · inbound

StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis cites this paper.

StellarTTS: Sparse Temporal Embedding for Low-Latency and Robust Speech Synthesis 3D-Speaker: A Large-Scale Multi-Device, Multi-Distance, and Multi-Dialect Corpus for Speech Representation Disentanglement

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T11:33:14.026639Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T11:33:14.026639Z digest=sha256:76d82c22e0315c4192536b6b8d0c639baadd21b630fc10ffa3ac8dd7d2399c16