Pith. sign in

Paper Citation Record · LEDGER

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2505.24059.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.24059 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:48.043254Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:39:41.644202Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T12:39:49.151368Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact3
  • verified fuzzy23
  • unresolved17
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a03eef48-8f97-425e-931d-ecfed85ff136 · outbound

This paper cites In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5].

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition In speech production, individ- ual variability in both articulation and its consequent acoustics is robust [4, 5]

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.910034Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:41.558755Z digest=sha256:4418acec01968abb49aba046c8352ddb95d41ba1cc664eae47c167cce9b35fe6

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · outbound

This paper cites Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:6f0c126825bd51c082852a79d57a8fb68548d4bdb135ebe94eedd3461f5df2c1

Observation 19660cb0-be24-429e-ab9b-effb73b85fca · outbound

This paper cites Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Dataset We use a new single speaker rtMRI corpus with simultaneously recorded audio and video

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.432731Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:41.990420Z digest=sha256:b7ffa032a6f7940b87e535bc375063776dedd2558a3649708c7b8d5097fdc429

Observation 303e5974-c565-4d3d-aa47-6264830f8626 · outbound

This paper cites Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Phoneme error rate (PER) The overall PER results on the held-out test set are presented in Table 1

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.067956Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.226483Z digest=sha256:ef0dbc87d60e70e25c5ea79c4c1004a0812ec235b584e3ab1b941646cacd0416

Observation 8bdbaa1d-acef-4eb8-a05f-8f629f3b9610 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.692447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.386227Z digest=sha256:fb64abb0b69385d15667eecf3a102653f7f8bc236ca43928635b42c7f07ea1c6

Observation 4f9464f1-031f-4df4-aa14-49e5721d453c · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:55.539004Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.468817Z digest=sha256:cbb84f50da6dac7d2437e9ae63e1ac4ce7bf1f3f63e322add9fdffc09cb81689

Observation b55612bd-f444-4527-91c0-4121046af81d · outbound

This paper cites The architecture of speech production and the role of the phoneme in speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The architecture of speech production and the role of the phoneme in speech processing,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.456393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:43.937557Z digest=sha256:0791c996036db3ee4e0bdadf9f53fcc3057b526f5b9c9e47e3b1f383d31fec08

Observation 02634c6e-5d51-4923-8e6b-70a56bd6e2f3 · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 8

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.199624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:44.133282Z digest=sha256:633e476b34b63095a894f2af42a08b0d9c8bc17a7e6f53c7d1bb368fbcf11cc7

Observation ad08ce4c-bade-4fbb-9f2a-e86bc7862945 · outbound

This paper cites Sensorimotor integration in speech processing: computational basis and neural organization,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Sensorimotor integration in speech processing: computational basis and neural organization,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.388476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.627165Z digest=sha256:75b745a67a20cfda4fe2c952f53ccc28e5174216d0536f6346f125c119397dc5

Observation cefcdd35-e84e-47e7-83a3-67e13141c8a9 · outbound

This paper cites Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Perception drives production across sensory modalities: A network for sensorimotor integration of visual speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:55.193380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.893417Z digest=sha256:5d26f5bf265e5f16e300430b14f5f6e8ea295c559c0fba3eb78c13b4a0cf95ce

Observation ba572b53-d737-4711-b1dc-fc85b808dfda · outbound

This paper cites The role of temporal modulation in sensorimotor interaction,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The role of temporal modulation in sensorimotor interaction,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.996114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:43.106505Z digest=sha256:c8297dcc930ff993e05cf8c19e09bdd2fb69af38f13c584fbb274d4b7d3a1eea

Observation 85e5706b-555a-4ae0-bd48-ce09cd3fa621 · outbound

This paper cites Variability of articulator positions and formants across nine english vowels,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Variability of articulator positions and formants across nine english vowels,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:54.823327Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:43.308671Z digest=sha256:f770e59e34e703169ce30094c6ae59308b6b1f4665d63adef93badd789037f8f

Observation 28164159-cfd1-489b-9d53-4391d0aed9db · outbound

This paper cites an unresolved cited work.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Unresolved cited work

Reference 13

Resolution
unresolved
raw_fallback, observed 2026-08-07T12:39:54.596321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:43.525990Z digest=sha256:133019dd1ec9283769881b015dbe292bd6e5c476ccf6e7b3d6dccbbe25494cd4

Observation 3583cc3e-bcdd-47fe-a774-a655e5c0234e · outbound

This paper cites Articulatory phonology: An overview,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory phonology: An overview,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:43.714834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:43.714834Z digest=sha256:99c30430c614c74e460260b00b59e06f9877b81e428e83665f84e7cbd0686ab6

Observation 1db0d81e-9b23-4164-93c8-43726580f67d · outbound

This paper cites Multimodal representations for syn- chronized speech and real-time mri video processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Multimodal representations for syn- chronized speech and real-time mri video processing,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.064691Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:45.396716Z digest=sha256:a6476bb9da3703642172a72d318749f3433759cdf2bb4629dfaf4928af58045f

Observation 2e3d5334-098b-416f-8c84-7c5c79b3c398 · outbound

This paper cites Towards speech classification from acous- tic and vocal tract data in real-time mri,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards speech classification from acous- tic and vocal tract data in real-time mri,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.867621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:45.545477Z digest=sha256:f562fde3f83022196529d6a61273c666004a0aefb5400417e8eae8fff69775c0

Observation e7a1966a-fbb3-4274-88fc-a3c1116e54fb · outbound

This paper cites Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:44.345646Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:44.345646Z digest=sha256:b2d11a155bfbff6b1a29bd491a94bcb51fa9e97e4bca2e75516f13fec0e43463

Observation f2a4bde9-cf0a-4d2d-8385-60c6cee01e8c · outbound

This paper cites Robust Audiovisual Speech Recognition Models with Mixture-of-Experts.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust Audiovisual Speech Recognition Models with Mixture-of-Experts

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.950065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:44.552105Z digest=sha256:6bc496e65d2e0bc4b3929f54c16e50e62d4c8f93ce7293cc76a4d900bc007161

Observation 002a3ab1-5feb-492e-a620-76fbaab3bf76 · outbound

This paper cites Speech production real-time mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Speech production real-time mri at 0.55 t,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.935947Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:44.741094Z digest=sha256:1d8a30a28559351c76afaea2c9a1a591a896d1a6ca8631293bccfa7032c1842a

Observation 84caf0c9-50a7-4904-89ba-54b5dbe728e1 · outbound

This paper cites A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A multispeaker dataset of raw and reconstructed speech production real-time mri video and 3d volumetric images,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.599113Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:44.921151Z digest=sha256:a8b702dcd884b40e59242e76ce5a2280db67e0824ccae860e67c25ffae4f9f76

Observation 943031f2-54a5-4935-baf7-120ab28b8497 · outbound

This paper cites The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The output of the Conformer is decoded by a single LSTM layer, with a final linear layer used for prediction

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.668434Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:41.818246Z digest=sha256:ce24b51acc1662ac5ef7965ef2e277560a92a8cc50a7325b8b33b0f4d15b5d56

Observation 53f86a46-9a4c-4b2a-b685-81efa2cfa5ac · outbound

This paper cites Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards Automatic Speech Identification from Vocal Tract Shape Dynamics in Real-time MRI

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:45.026461Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:45.026461Z digest=sha256:d18cb9a2e79e6f9733bf50327d462fdd5d0346da2026ec4bb13857f931562605

Observation 8f0f584e-2244-427e-894a-e5164b58361b · outbound

This paper cites Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Cnn-based phoneme classifier from vocal tract mri learns embed- ding consistent with articulatory topology

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:53.277765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:45.220360Z digest=sha256:74e96cbaac5461ffc5c07f42c527ad876ecfccd1b4f0d795976fa572fecd78e4

Observation c7435ec9-7a19-48f8-8a38-cca29cc75173 · outbound

This paper cites Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Eval- uation of a novel 8-channel rx coil for speech production mri at 0.55 t,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.456383Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:46.700802Z digest=sha256:c6b99d2d7290b5f1bed03045639317f7624695cfb3dd4efc385768c3a27fa880

Observation 3d723bf9-a38a-4dce-b296-52a9c168af75 · outbound

This paper cites Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Direct articulatory observa- tion reveals phoneme recognition performance characteristics of a self-supervised speech model,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.651440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:45.684601Z digest=sha256:f735dcb1286e71e3cd15ac02c7acb2aba248f81a7ee636eb3a721d9a6f7331b0

Observation 4a4267c1-21c3-48f6-8d48-5aa5897810fa · outbound

This paper cites Usc-timit: A database of multimodal speech production data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Usc-timit: A database of multimodal speech production data,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.392791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:45.838539Z digest=sha256:e883c703fbda877870d3d6ef4edf35f8a98b95f86c9a349c781f641e748d0f27

Observation 4a5597dc-eb9f-4ab4-835b-3445f4b1a323 · outbound

This paper cites Database of volumetric and real-time vocal tract mri for speech science.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Database of volumetric and real-time vocal tract mri for speech science

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:52.163254Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:46.054670Z digest=sha256:39a511617ea6922796c3e2b6d3822d60c0a2d6ea1ca05333e3c47c75745c5f53

Observation 93f597cf-99dd-42af-a453-18d0d01513ef · outbound

This paper cites Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Characterization of inter-speaker articulatory variability: A two- level multi-speaker modelling approach based on mri data,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.879504Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:46.187469Z digest=sha256:a24c6e29d8aacb5a32736b03cf74935f694f44d636a8ab6b3e314f6a80dc03bf

Observation 61506fae-6150-4f64-9c98-12aa2c30bd3e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.387249Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.387249Z digest=sha256:503086a14d277fc8ec8fef5cd540bca9edd782df5250cd99db74fcdd05231c8b

Observation 871471fd-e05a-40c1-99e4-0a104d85efe2 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.480237Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.480237Z digest=sha256:04ac32a75aa2bc2942db7d5af5573c3774ebd0602826ae546c1b488f0ee1e526

Observation 6f715e8c-6636-4581-b5c9-a3e2626e3e08 · outbound

This paper cites Simple and Effective Zero-shot Cross-lingual Phoneme Recognition.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Simple and Effective Zero-shot Cross-lingual Phoneme Recognition

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:46.592923Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:46.592923Z digest=sha256:a76c0f5775507a2bb400014e1bf0dc326945c6f8c6d19141f4de23ebc198b7bb

Observation e9410a79-ce06-4001-8534-f1eba9e98040 · outbound

This paper cites Robust and Efficient Medical Imaging with Self-Supervision.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Robust and Efficient Medical Imaging with Self-Supervision

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.585835Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.585835Z digest=sha256:42ca1a342ce23a0429607dd7b2b39c5567b545833cfdda98bd109814c6331ff7

Observation 445e246d-f9f9-4650-8dd7-e3230c14340d · outbound

This paper cites State-of-the-art speech production mri protocol for new 0.55 tesla scanners,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition State-of-the-art speech production mri protocol for new 0.55 tesla scanners,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:50.243216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:46.766700Z digest=sha256:86cf02a82f655ab1eaea5aa08b8fd102338d489ad925701e507590a22a92b5b8

Observation 93c1d063-cc4a-4c4c-88a2-c60ffcfbd0fb · outbound

This paper cites Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Announcing the electro- magnetic articulography (day 1) subset of the mngu0 articulatory corpus,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.890156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:46.910754Z digest=sha256:001085715fdbd0e39212505e1f3c1c20007e3d6a07de5304e41e29441e908b27

Observation 75c526be-bcfc-442f-bfee-473c2513d850 · outbound

This paper cites Real Time Speech Enhancement in the Waveform Domain.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Real Time Speech Enhancement in the Waveform Domain

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.014834Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.014834Z digest=sha256:d47d826877c19050011fcde9e7200f6da21cce7cca70a483f93122c54752dde3

Observation 1808e0f1-00ca-4fce-bb7b-e83a42107629 · outbound

This paper cites Wavlm: Large-scale self- supervised pre-training for full stack speech processing,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Wavlm: Large-scale self- supervised pre-training for full stack speech processing,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.089969Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.089969Z digest=sha256:6d9ce72f7079d631a79af8f93b54544b3d925efd375fdda7b9c2c1fd3fdf9944

Observation 951f1270-1d31-4711-90cc-71a59945f422 · outbound

This paper cites Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Evidence of Vocal Tract Articulation in Self-Supervised Learning of Speech

Reference 37

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:48.383512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:47.199046Z digest=sha256:4cdb25d5b1f710b8cee771eef2ac6e6776845a61115b877f603de1fc70fd2496

Observation 4d4b0c77-17ab-4de1-82bb-8b7beb2f3203 · outbound

This paper cites An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition An Image is Worth 16x16 Words: Transformers for Image Recognition at Scale

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.361459Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.361459Z digest=sha256:2d8ef767aefa7643b8c200e178035fcf90f663e6ce39f96edbe42530c3ddd326

Observation f2bc14b2-647d-4a89-8c33-86c79e04dbc2 · outbound

This paper cites A Simple Framework for Contrastive Learning of Visual Representations.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition A Simple Framework for Contrastive Learning of Visual Representations

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.436332Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.436332Z digest=sha256:175df87b151564ba28dff21865cf48bd9a8fa1a9097dbd4877fb689dec73497a

Observation 8a6cf77d-0e33-4af0-8982-9e0ea2a20e6d · outbound

This paper cites Visualizing data using t-sne.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Visualizing data using t-sne

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:47.724939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:47.724939Z digest=sha256:9ecd9e1d6ea44a4af231c86186add521524769814e0347a73d33f9cd7a7aa807

Observation c9039109-fc06-40f7-9e0d-fc49b769315d · outbound

This paper cites Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Toward articulatory- acoustic models for liquid approximants based on mri and epg data. part ii. the rhotics,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:49.529579Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:47.898973Z digest=sha256:47b0144464c72994ef495044176408ce5ed637145ca49607f1e75b1a1abe78db

Observation 1251a66c-05b6-4b09-ad8a-62368c93a71e · outbound

This paper cites Articulatory characterization of english liquid- final rimes,.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Articulatory characterization of english liquid- final rimes,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-08-07T12:39:48.043254Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:39:48.043254Z digest=sha256:c887c3d1380140a33ea3f385a93bdd1499f3314fa23d279ced1be83acd25f49c

Observation c6434ec6-3f83-45d8-b41f-4669a6b8313b · outbound

This paper cites The number of Conformer layers was set to 3.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition The number of Conformer layers was set to 3

Reference 256

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T12:39:56.210483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:42.111215Z digest=sha256:443d151052c4f233688cd9049957ca2249ead06083bca2cdfb94961ea52fa9db

Pith citing papers

Observation 4b43e443-4b31-4fb0-aec0-873a25c92896 · inbound

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition cites this paper.

Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition Towards disentangling the contributions of articulation and acoustics in multimodal phoneme recognition

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T12:39:49.286208Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T12:39:41.644202Z digest=sha256:6f0c126825bd51c082852a79d57a8fb68548d4bdb135ebe94eedd3461f5df2c1