Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 30 inbound Pith citation observations for arXiv:2106.07447.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-10T20:27:22.494754Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
25
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation dc48ff5f-8d49-4f63-8941-bf4181efd1bc · inbound
WhiSPA: Semantically and Psychologically Aligned Whisper with Self-Supervised Contrastive and Student-Teacher Learning HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3efc365d-704a-4714-8d27-54e70b726133 · inbound
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c86a60e8-64fa-429d-b203-491345f93955 · inbound
Multimodal Zero-Shot Framework for Deepfake Hate Speech Detection in Low-Resource Languages HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b907b8c8-5dd0-4009-a434-2028a84c0a45 · inbound
Benchmarking Foundation Speech and Language Models for Alzheimer's Disease and Related Dementia Detection from Spontaneous Speech HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 3460
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2757860d-2e1a-4049-800c-f969eeb485a8 · inbound
Zero-Shot Cognitive Impairment Detection from Speech Using AudioLLM HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 762248e0-df10-4ea3-ac8f-a499ded69049 · inbound
XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca1dcac3-5898-4a56-8ae9-a51f4ae5a008 · inbound
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2f1a916e-e772-4f3d-8282-a493595b354f · inbound
Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff90399d-1651-42e3-97fb-5e6f8bcc7037 · inbound
Self-supervised learning of speech representations with Dutch archival data HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e26faa88-32ee-45c7-8e30-07a804b5bb7d · inbound
Leveraging Context for Multimodal Fallacy Classification in Political Debates HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 986ec86d-5bf2-489e-a7fd-de4c8516c67d · inbound
Step-Audio 2 Technical Report HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b7d3a70e-6d86-4d32-b58b-f91ce0aac5bf · inbound
A Concept-based approach to Voice Disorder Detection HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d20f6562-f71c-4c16-8d5f-bdf6236aaf4f · inbound
Seeing is Believing: Emotion-Aware Audio-Visual Language Modeling for Expressive Speech Generation HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9954816-dbc8-43e3-9707-98f3ef9225ec · inbound
Entropy-based Coarse and Compressed Semantic Speech Representation Learning HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a0f2b216-8a76-4676-9c08-b3ca6222aa23 · inbound
Where Do Backdoors Live? A Component-Level Analysis of Backdoor Propagation in Speech Language Models HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 14a60961-3961-41d0-a5f4-b6cea4c81c02 · inbound
A Two-Stage Dual-Modality Model for Facial Emotional Expression Recognition HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation baacf565-a834-462f-bf1b-b9e2c8474f3a · inbound
findsylls: A Language-Agnostic Toolkit for Syllable-Level Speech Tokenization and Embedding HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7728530-8560-4a6b-9fcb-4bc70517722d · inbound
Neural networks for Text-to-Speech evaluation HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation b1c58a5a-15e7-47fa-849e-a00281d74a84 · inbound
Meow-Omni 1: A Multimodal Large Language Model for Feline Ethology HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 590cc051-77a2-40c2-9753-b2bdf1d71bc4 · inbound
Fine-tuning language encoding models on slow fMRI improves prediction for fast ECoG HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 27588fde-9bd5-4255-a614-0f6e73963047 · inbound
F3-Tokenizer: Taming Audio Autoencoder Latents for Understanding and Generation HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation fcef0e10-8117-47e2-959c-c7bf9d3907e5 · inbound
Pretrained self-supervised speech models can recognize unseen consonants HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 46238d6e-7817-465c-bc14-d6371a8c32b9 · inbound
Light-weight Pronunciation Assessment via Discrete Speech Token Surprisal HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4cba52db-200a-4e08-9c0a-e31d1b391a8c · inbound
Interleaved Speech Language Models Latently Work In Text HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 12d732ad-24e4-49be-9020-97b89026ede6 · inbound
End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 05e6240c-cfc1-42e9-b23f-5ac9f84ef1b8 · inbound
Syntactic Belief Update as the Driver of Garden Path Processing Difficulty HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 288
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 28de5576-2bb0-49ca-9446-7276fb253d47 · inbound
Speaker-Aware Temporal Aggregation Strategies on Segment Representations for Depression Detection in Dyadic Interaction: A Benchmark Study HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e838a15-a1a9-403d-9d33-4fc1b0085c1f · inbound
Layer-wise Cross-Lingual Depression Detection from Speech: Analysis with Contrastive Alignment HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3df35c21-9c57-4186-97af-02c2123f6fe0 · inbound
Video2Reaction: Mapping Video to Audience Reaction Distribution in the Wild HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 0c6ac7ef-ce87-48c6-acb2-7e2bfde25cf9 · inbound
DONDO: Open w2v-BERT Speech-Recognition Base Models for African Languages HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.