Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:53:51.803073Z
Paper Citation Record · LEDGER
As of 7 August 2026, this Paper Citation Record lists 28 of 28 outbound references and 1 inbound Pith citation observation for arXiv:2506.20361.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:53:51.803073Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-06T22:53:48.450553Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-06T22:53:52.505268Z
28 of 28 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation cae30958-580b-4a12-898d-bb35d334a24c · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 24f8d8b4-f788-404d-b227-2091412a59d3 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Unresolved cited work
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation c0134223-de03-4e42-b025-6ce6f5285cbb · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Dataset We analyzed the phonetic decodability window on two datasets separately
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 61bea0c4-4e63-4d3f-b9ac-cf73c541fca9 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7eae74aa-131c-4dc9-8a0f-c363e80322ef · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models We found that A V-HuBERT’s encoding of speech temporal dynam- ics is dominated by its audio input
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 29e31935-0a1a-4314-8027-bc8c1e01f636 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation dc7c5188-171b-4f40-886a-d53c5020da7c · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models We would like to thank Biao Zeng from University of South Wales and Hao Tang, Sharon Goldwater from ILCC, Univer- sity of Edinburgh for useful discussion
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 64abd8f1-82eb-4fbe-abb1-9639bad530a5 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Using artificial neural networks to ask ‘why’ questions of minds and brains,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b48a8210-51ec-4cc5-9f5f-21253ce644b2 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Parallel hierarchical encoding of linguistic representations in the human auditory cortex and recur- rent automatic speech recognition systems,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 8ec8c1a4-95b3-4bb8-a244-04ffd366196b · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Hearing lips and seeing voices,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8cc38be9-dce2-4cfe-aa45-8ac763842dd8 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Multisensory integration: current issues from the perspective of the single neuron,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 0059e289-68da-490d-b5e5-6050a2e55bba · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Multimodal deep learning
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation b6c4587a-83bb-48dc-a28a-7e80954ebd17 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models On the role of noise in audiovisual integration: Evidence from artificial neural networks that exhibit the McGurk effect,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation fd50aea4-84af-4bc9-8022-d12eb1547d04 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28fdc799-899e-4acf-b4c6-6a96d44263e6 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models The natural statistics of audiovisual speech,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 71710634-2011-40e9-a40e-26b6ceaf68bf · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Bimodal speech: early suppressive visual effects in human auditory cor- tex,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 7fd2bde6-cde9-4a31-bc63-52dfa385c0da · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Visual speech speeds up the neural processing of auditory speech,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation d470de32-c3e2-4f2a-9d10-d283a1e640c9 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Asynchronicity between visual and auditory information in audiovisual speech: Evidence from four types of consonant- words/b/,/t/,/k/and/g,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e307496b-1ff3-4266-ad1e-6a2c3f165244 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models A predictive learning model can simulate temporal dynamics and context effects found in neural representations of continuous speech
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 38ebd5c8-3751-4a3f-8bd7-7411414dde71 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Neural dynamics of phoneme sequences reveal position-invariant code for content and order,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 123861fc-64e1-40ad-9645-3e6dc382826e · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d15b4bb8-4347-45bc-9b1a-8b7614b2d6f7 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Deep residual learning for image recognition,
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22583bb2-f733-4bba-b601-d0c22e9abf22 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Dynamic encoding of acoustic features in neural responses to continuous speech,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation e5c540e0-211b-4e72-91e0-98f4ac694089 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Representation Learning with Contrastive Predictive Coding
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4d1bdad-7f83-49d2-89a1-2787d06aa1af · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Montreal forced aligner: Trainable text-speech align- ment using Kaldi
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 1fb8ac9e-8336-43ee-b842-f7e5a04fb035 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models Prosodylab-aligner: A tool for forced alignment of laboratory speech,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation 216414d2-5e51-46c8-94e4-9dcd976e5bcb · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models An audio- visual corpus for speech perception and automatic speech recog- nition,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.
Observation a70d74ae-21d5-4051-92ac-96d592074ef8 · outbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models LRS3-TED: a large-scale dataset for visual speech recognition
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29e31935-0a1a-4314-8027-bc8c1e01f636 · inbound
The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models The role of audio-visual integration in the time course of phonetic encoding in self-supervised speech models
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.