Pith. sign in

Paper Citation Record · LEDGER

Self-Supervised Speech Representations are More Phonetic than Semantic

As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 6 inbound Pith citation observations for arXiv:2406.08619.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2406.08619 v1

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 6 of 6 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 6 of 6 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T19:32:52.550002Z

measured 1 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7fa9716a-4233-4141-9a93-650421c15ae1 · inbound

Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding cites this paper.

Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language Understanding Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:10:04.115891Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:10:04.115891Z digest=sha256:512dfc97708b0247b2069d9f61dc4c5cca323a8d6fb2ecf769240e1c5169ea04

Observation 683d8685-7551-4494-b098-9d356f18ddc0 · inbound

Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis cites this paper.

Continuous Autoregressive Modeling with Stochastic Monotonic Alignment for Speech Synthesis Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-09T16:43:08.683147Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T16:43:08.683147Z digest=sha256:b7284840e63733a0f336264cabfed8a45c03fd6a44c2d46ac869eaf3bf2919e7

Observation 1bf89544-a4f6-4900-a853-94601fe19a48 · inbound

From Words to Waves: Analyzing Concept Formation in Speech and Text-Based Foundation Models cites this paper.

From Words to Waves: Analyzing Concept Formation in Speech and Text-Based Foundation Models Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 45

Resolution
unresolved
no resolver link, observed 2026-08-07T11:55:23.641312Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:55:23.641312Z digest=sha256:cb9231c4c73feadd0d90a6376029534987b201f4458cddce9075a3877a7ad88e

Observation a99d50d4-4dd6-4f89-a014-36f0c61630a0 · inbound

Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching cites this paper.

Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T19:32:52.550002Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T19:32:52.550002Z digest=sha256:329a77c85022d7516c68ef214b8bf1f672eed24b7942d4198c3ce11c75217c0c

Observation c9ff6f51-2056-4e15-a5a1-b30733c5682a · inbound

Entropy-based Coarse and Compressed Semantic Speech Representation Learning cites this paper.

Entropy-based Coarse and Compressed Semantic Speech Representation Learning Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T13:36:04.348855Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T13:36:04.348855Z digest=sha256:e8eb7ed1c537be01a104d96970c1ad42091050455cb61aeb5a4210878ee070c2

Observation 61d8b787-28d6-47a4-93f2-44f99ca872bd · inbound

A framework for analyzing concept representations in neural models cites this paper.

A framework for analyzing concept representations in neural models Self-Supervised Speech Representations are More Phonetic than Semantic

Reference 188

Resolution
verified exact
arxiv_id, observed 2026-05-09T22:18:59.210149Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=arxiv_source observed=2026-05-09T14:49:22.776209Z digest=sha256:6cefc1845e32c0f07dbdc17aba8694ef31029f7d145ba7b243dd0a9b6782e47c