Pith. sign in

Paper Citation Record · LEDGER

Rhythm Features for Speaker Identification

As of 8 August 2026, this Paper Citation Record lists 32 of 32 outbound references and 2 inbound Pith citation observations for arXiv:2506.06834.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.06834 v1

Coverage vector

measured 32 of 32 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:07.435997Z

measured 34 of 34 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T05:54:04.008475Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T05:54:08.361418Z

Reference resolution

32 of 32 outbound references displayed

  • verified exact5
  • verified fuzzy20
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c47e5e7b-a891-4f4e-80c3-872e70f9e8fe · outbound

This paper cites an unresolved cited work.

Rhythm Features for Speaker Identification Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T05:54:13.271527Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:03.435043Z digest=sha256:c46f99cdf890ce093cf64d6e927f2e55e2f5643d44bcc68f934fdef8c0e8f996

Observation a3f4dd10-0d06-474a-bb20-03f7b3ec3f38 · outbound

This paper cites For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech.

Rhythm Features for Speaker Identification For example, [11] demonstrated that human listeners were able to identify familiar speakers based on only a sinusoidal encoding of the prosodic informa- tion in speech

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:13.054283Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:03.623277Z digest=sha256:79a42a75e4dbb37abaff637526d527ea05031b6f51bd42183ae681831b258630

Observation aa410efc-82a1-4641-9fff-3b368902c93f · outbound

This paper cites Rhythm Features for Speaker Identification.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:696d7e89757db6f7847781ec65aef8c764f1cd7288d616e91d81eb04849c74dc

Observation c6d38783-3ffe-4835-ba9f-9eb975717d90 · outbound

This paper cites Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets.

Rhythm Features for Speaker Identification Datasets We evaluate our approach on a closed-set SI task using two pop- ular speech datasets

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.691984Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.217453Z digest=sha256:45074f7532d4fc285e9f3ded4fc6b6e4d9748d5299c9a18463bfa3b19986b467

Observation c25e25d8-041c-4101-8ecc-61c15eef1b71 · outbound

This paper cites Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER).

Rhythm Features for Speaker Identification Additionally, it is better suited to our closed-set identi- fication task than other commonly used measures, such as equal error rate (EER)

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.366008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.600304Z digest=sha256:30d73ee86242babcb8dc03a959ade1ac7da305d16a12896353cb3c0b091a3775

Observation 406f3c22-f2a8-4119-bd03-8df349f8a56a · outbound

This paper cites Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity.

Rhythm Features for Speaker Identification Discussion Our results are consistent with prior works [2, 3] that have suggested rhythm features do convey useful information about a speaker’s identity

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.175705Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.771338Z digest=sha256:9860b5fa721412a54ae3517289cfe17f8ab1c5d93b0d3c6c2d1bc321f66fa11c

Observation 53377cd6-1458-4229-96e6-f00d78470ba2 · outbound

This paper cites Durations of context- dependent phonemes: A new feature in speaker verification,.

Rhythm Features for Speaker Identification Durations of context- dependent phonemes: A new feature in speaker verification,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.666295Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.885515Z digest=sha256:c48c1484f3ff895262f6e3983e4103dcdc2aa9620f7c02f10c088bc1527416ad

Observation 028b0074-465d-45ce-b6bd-e94900cf3c93 · outbound

This paper cites Prosodic pa- rameter for speaker identification,.

Rhythm Features for Speaker Identification Prosodic pa- rameter for speaker identification,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.441207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.937445Z digest=sha256:34db6eca89042405272f4f3d139cf1bb1befea67036dfa28c3777bdfb1d939b9

Observation 40c9c841-afce-47e3-88d1-a9483ff02afd · outbound

This paper cites Speaker recognition by machines and humans: A tutorial review,.

Rhythm Features for Speaker Identification Speaker recognition by machines and humans: A tutorial review,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.971180Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.981905Z digest=sha256:f06de3f0b7d57ccd8e1519b015c6ab603ec545ce71fae1fb386868fae9811cfc

Observation ab08149a-b143-4b5d-a9dc-5e658b131205 · outbound

This paper cites We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing.

Rhythm Features for Speaker Identification We used the pre-existing identifica- tion split, which uses roughly95%of the utterances for training and the remaining5%for testing

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.514529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.381391Z digest=sha256:c63031d4366b246f292bde1f2cf35169d8daa2246b4a4ca86b18bb70ccfff0a2

Observation e6651ab3-e85a-430b-beea-90175a08c77b · outbound

This paper cites Is phoneme length and phoneme energy useful in automatic speaker recognition?.

Rhythm Features for Speaker Identification Is phoneme length and phoneme energy useful in automatic speaker recognition?

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.788813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.128715Z digest=sha256:4fa765bee504acec6986029158a4c6c8b364e53c38cdb3775eb55cdd0bc62a52

Observation 79f82602-70e2-48e3-8f0e-de41d3af3c31 · outbound

This paper cites Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody.

Rhythm Features for Speaker Identification Other works, however, such as [13] have found that content can have a substantial influence on certain aspects of the speaker’s prosody

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:12.873829Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:03.842347Z digest=sha256:1d8e4126983e9c92623352cdf4d0f62e07f1eaea5c95f9ddb52084cc66809363

Observation a3fef487-3d24-482d-bc30-51e7fc883489 · outbound

This paper cites Intrinsic phone durations are speaker-specific,.

Rhythm Features for Speaker Identification Intrinsic phone durations are speaker-specific,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.491929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.294335Z digest=sha256:bbada6fb8ddf75c39cb8d31182b35fed5745746e2e3f2f86e396d40eda4d77bb

Observation 878ef969-b1df-4b60-8b17-4f754decae99 · outbound

This paper cites What determines duration-based rhythm measures: text or speaker?.

Rhythm Features for Speaker Identification What determines duration-based rhythm measures: text or speaker?

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:11.212587Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.428798Z digest=sha256:659e0ed994d51e2adcfef21440f1f5e74a0b540da43a04fca150e215f934338d

Observation 3d853395-be5f-418c-93af-4b809e211051 · outbound

This paper cites Extraction and representation of prosodic features for language and speaker recognition,.

Rhythm Features for Speaker Identification Extraction and representation of prosodic features for language and speaker recognition,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.964542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.739504Z digest=sha256:6b55e7fb1bbdb771e04f58d2ba98a7d022c0897041bc1a0988f71be3acf4328f

Observation ff9843ba-ad83-49a2-92a3-5a0cc28e648a · outbound

This paper cites Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization.

Rhythm Features for Speaker Identification Analysis of Speech Temporal Dynamics in the Context of Speaker Verification and Voice Anonymization

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.210687Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.801906Z digest=sha256:a85b71fead05bfe9ed0b976298faa44ed13452623a30abf38526d61a51f31183

Observation 95e1751b-1236-4cd7-9a5b-858ef3a85bc0 · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

Rhythm Features for Speaker Identification Robust Speech Recognition via Large-Scale Weak Supervision

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.944151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.944151Z digest=sha256:9fa4d93b4500fde54a67bef0d4b3a7de833576abc62be93a3988eb03512c209a

Observation 8cd7a73f-c587-42bf-b65f-ba31151625be · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Rhythm Features for Speaker Identification Lib- rispeech: an asr corpus based on public domain audio books,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:10.169887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.047270Z digest=sha256:621363f3da98510dd494ef0a77ef42ffb660096fcd7c6a793f9486c057caa680

Observation 2cc2a9a4-e143-4516-ac91-7356b31f31f9 · outbound

This paper cites V oxceleb: a large- scale speaker identification dataset,.

Rhythm Features for Speaker Identification V oxceleb: a large- scale speaker identification dataset,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.869062Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.126513Z digest=sha256:aedec8fe036f8615605971cfa063cf8e9e5c36b80ecfee3528e02c43e12cafe0

Observation 22708e30-6233-4650-af6a-ecf4cd757192 · outbound

This paper cites On the importance of pure prosody in the perception of speaker identity,.

Rhythm Features for Speaker Identification On the importance of pure prosody in the perception of speaker identity,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.615160Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.230664Z digest=sha256:1998a42fef5063342f3d5d223eb034723247ab1df7a216a4ac4996a919b48a52

Observation 1fc3e0fc-4607-4b20-82b0-67155870bfbb · outbound

This paper cites Speaker idiosyncratic rhythmic features in the speech signal,.

Rhythm Features for Speaker Identification Speaker idiosyncratic rhythmic features in the speech signal,

Reference 21

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.585336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.362057Z digest=sha256:ca71b309e455ff144649c35b65631364bc07b56ae39c53a7c294a66461a33f43

Observation a67c0dcb-b08a-4288-aac9-bb840b46e271 · outbound

This paper cites How stable are acoustic metrics of contrastive speech rhythm?.

Rhythm Features for Speaker Identification How stable are acoustic metrics of contrastive speech rhythm?

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.406348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.451841Z digest=sha256:32c385e585e38beb0239cd3b8ce6967dcbba4d7262a1d28c1c3caa8adf0d1971

Observation 701d603e-08e5-4652-96c2-2827259db652 · outbound

This paper cites Speech rate normalization used to improve speaker verification,.

Rhythm Features for Speaker Identification Speech rate normalization used to improve speaker verification,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:09.247335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.559650Z digest=sha256:6180651df0403c72c0f51849116f209f8d7dd20aa7da3c4739a74cf5b0428846

Observation a31148a8-0d21-46c6-8e8a-c7ef37354a60 · outbound

This paper cites Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis.

Rhythm Features for Speaker Identification Phoneme duration modeling using speech rhythm-based speaker embeddings for multi-speaker speech synthesis

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.995644Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:06.735474Z digest=sha256:6858e471b8b1071b84e7abfe6a2d374e505c0507c3ad4cfcf043e3b713f755f7

Observation e3af9745-c40e-4454-8921-92433606a7d6 · outbound

This paper cites Whisperx: Time-accurate speech transcription of long-form audio,.

Rhythm Features for Speaker Identification Whisperx: Time-accurate speech transcription of long-form audio,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:06.859629Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:06.859629Z digest=sha256:a33b96f7cda32f70b7a7e59d30d8873354c134a35d686a3a56e892ff470fd7e7

Observation a4b76104-1cff-4184-9d31-d430f5668a06 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Rhythm Features for Speaker Identification Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.029788Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.029788Z digest=sha256:7e036d661c81e02f77b4b80c0259e72c74ee5e9b25da1c81f898981446b3f8d4

Observation 8fa06a12-12ce-44d2-a4d7-055ac1174e82 · outbound

This paper cites SpeechBrain: A General-Purpose Speech Toolkit.

Rhythm Features for Speaker Identification SpeechBrain: A General-Purpose Speech Toolkit

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.103872Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.103872Z digest=sha256:75ab0049fdae275c1745986ed57c456b9f1d264e067baddd77145317398c91af

Observation 8d74c528-c8ac-4ed6-af2f-cc29a99eabe4 · outbound

This paper cites Metrics for Multi-Class Classification: an Overview.

Rhythm Features for Speaker Identification Metrics for Multi-Class Classification: an Overview

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T05:54:07.182515Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T05:54:07.182515Z digest=sha256:b9bf708f4e8d3dba32537d0875507a4e84c9c1ec924b9a94f19ef36029752cee

Observation 421f1a32-f7e0-47a4-9dc3-69e2765aab3e · outbound

This paper cites Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature.

Rhythm Features for Speaker Identification Improving Automatic Emotion Recognition from speech using Rhythm and Temporal feature

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:08.006889Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:07.275427Z digest=sha256:9f8f6d49b46202e0c510119d4ba888df64fef26a4a56e9fad30397a8f23592ff

Observation d64348cd-c582-481b-8c7d-222d99e010e1 · outbound

This paper cites Probing the in- formation encoded in x-vectors,.

Rhythm Features for Speaker Identification Probing the in- formation encoded in x-vectors,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T05:54:08.741765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:07.324574Z digest=sha256:6f1e2cd645fa27cfa2bd3721aa758140f42277bb77d5c8509799b5398f708c5c

Observation 8ba2e869-2962-4fd8-b72e-65056bfdd04d · outbound

This paper cites An empirical analysis of information encoded in disentangled neural speaker representations.

Rhythm Features for Speaker Identification An empirical analysis of information encoded in disentangled neural speaker representations

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-07T05:54:07.850121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:07.435997Z digest=sha256:4d4a1592705621264846a32418274106d7c508a9707e2ea423475525937307bd

Observation 2e035620-68bf-4465-95bf-3bfac055ce88 · outbound

This paper cites Available: https://doi.org/10.1515/lp-2013-0012.

Rhythm Features for Speaker Identification Available: https://doi.org/10.1515/lp-2013-0012

Reference 2013

Resolution
verified exact
doi, observed 2026-08-07T05:54:07.708316Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:05.581631Z digest=sha256:06fed42f83349a84b7f3d6d1928aca7bd829b54c2502e6c87985216ac98ae136

Pith citing papers

Observation aa410efc-82a1-4641-9fff-3b368902c93f · inbound

Rhythm Features for Speaker Identification cites this paper.

Rhythm Features for Speaker Identification Rhythm Features for Speaker Identification

Reference 3

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T05:54:08.482430Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T05:54:04.008475Z digest=sha256:696d7e89757db6f7847781ec65aef8c764f1cd7288d616e91d81eb04849c74dc

Observation 0e985fe9-4c6d-4ea1-a3fc-8ccaa30980b0 · inbound

Multimodal Speaker Verification as a Threat to Speaker Anonymization cites this paper.

Multimodal Speaker Verification as a Threat to Speaker Anonymization Rhythm Features for Speaker Identification

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-01T12:12:42.171456Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:12:42.171456Z digest=sha256:e53581c8215ed137c10130a2234840a06be806f81ab0eeae9ce3a2a0dd31cab5