Pith. sign in

Paper Citation Record · LEDGER

SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 5 inbound Pith citation observations for arXiv:1910.11559.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1910.11559 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 5 of 5 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 5 of 5 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T12:52:39.184008Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-17T22:45:24.330275Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation f26246e6-ec86-4efd-a0b7-f86657cc9403 · inbound

Spoken question answering for visual queries cites this paper.

Spoken question answering for visual queries SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-07T12:52:39.184008Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:52:39.184008Z digest=sha256:51c93d7af81ac8246a0ac64081e0692afeeab60ca0cabb060bde00e921013e46

Observation fd5d21fa-3fc5-49d1-85db-b13ec0454e93 · inbound

Reasoning-Based Approach with Chain-of-Thought for Alzheimer's Detection Using Speech and Large Language Models cites this paper.

Reasoning-Based Approach with Chain-of-Thought for Alzheimer's Detection Using Speech and Large Language Models SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

Reference 44

Resolution
unresolved
no resolver link, observed 2026-08-07T11:40:12.061556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:40:12.061556Z digest=sha256:f28e7bfb219bc3dc895ce5e446a69af846db03116f67cc4a0ed0a42b551d0d4d

Observation f8f00b46-0b68-4800-8e03-04f7680dc0a4 · inbound

DeepEmoNet: Building Machine Learning Models for Automatic Emotion Recognition in Human Speeches cites this paper.

DeepEmoNet: Building Machine Learning Models for Automatic Emotion Recognition in Human Speeches SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T18:31:36.143610Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T18:31:36.143610Z digest=sha256:6f68f58899279c97624d24775546f88e8446c2aabfd22040ee76b235d9cc6cd6

Observation d97dea7c-61bd-4a1b-bfe3-56310cf66e5d · inbound

Amplifying Emotional Signals: Data-Efficient Deep Learning for Robust Speech Emotion Recognition cites this paper.

Amplifying Emotional Signals: Data-Efficient Deep Learning for Robust Speech Emotion Recognition SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T15:51:22.740660Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T15:51:22.740660Z digest=sha256:ad0546b15730e43b8e5c0eee12bcaf54254c2bc424f7e9491df9311c361f3e5e

Observation d534fa49-fb0a-4f40-95e5-2624614331c9 · inbound

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering cites this paper.

End-to-end Contrastive Language-Speech Pretraining Model For Long-form Spoken Question Answering SpeechBERT: An Audio-and-text Jointly Learned Language Model for End-to-end Spoken Question Answering

Reference 8

Resolution
verified exact
arxiv_id, observed 2026-05-17T22:45:24.333218Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=arxiv_source observed=2026-05-17T22:44:49.949759Z digest=sha256:c1013dd82dfb2b5f2fee40832220c82a9cdd6776bf99bed70cc96b8e068b8158