Pith. sign in

Paper Citation Record · LEDGER

SUPERB: Speech processing Universal PERformance Benchmark

As of 12 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2105.01051.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2105.01051 v4

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00

measured 40 of 40 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-11T21:30:10.997692Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-07T14:53:55.761527Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 7f08a44d-91ff-4e09-8a10-aea7e0730d3b · inbound

CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing cites this paper.

CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing SUPERB: Speech processing Universal PERformance Benchmark

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T21:30:10.997692Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T21:30:10.997692Z digest=sha256:f53d3a68a719db1bf1c1359a88fe650367458b3adc5efb33d3e2db98f4dd89fd

Observation aa365929-7583-4da4-924a-e36fdb2f6d70 · inbound

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis cites this paper.

Speech-Forensics: Towards Comprehensive Synthetic Speech Dataset Establishment and Analysis SUPERB: Speech processing Universal PERformance Benchmark

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-11T17:24:22.432903Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-11T17:24:22.432903Z digest=sha256:dec4d124cd85be85fd9f609c2b8832550f14319d460b26412f0d2aedce2726af

Observation 482a3266-c427-437a-8a90-2b17a9a8bf9d · inbound

Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection cites this paper.

Noise-Agnostic Multitask Whisper Training for Reducing False Alarm Errors in Call-for-Help Detection SUPERB: Speech processing Universal PERformance Benchmark

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T18:06:18.040748Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T18:06:18.040748Z digest=sha256:5456ea04449517a971f98e2f30b8bc06a14a414d9c190d7a114d413affb89e35

Observation 0d6ff449-43b6-4c1b-8bc7-29e7b25f4afd · inbound

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey cites this paper.

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey SUPERB: Speech processing Universal PERformance Benchmark

Reference 153

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:19.889484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:19.889484Z digest=sha256:e29e182dfdf16d170be83444dd657ee5291979711a928488669326e192fe1458

Observation 36a1a052-2da7-43d5-a94e-3fa582a15f61 · inbound

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation cites this paper.

When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-09T19:16:37.865512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-09T19:16:37.865512Z digest=sha256:94738382825d4bc55ff39c41e9476d4cecd7de7331524a630800fbe0466b50ad

Observation 784e7f01-6faf-43e2-a3e3-16adc5fa465a · inbound

Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection cites this paper.

Comprehensive Layer-wise Analysis of SSL Models for Audio Deepfake Detection SUPERB: Speech processing Universal PERformance Benchmark

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-09T04:35:12.644162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-09T04:35:12.644162Z digest=sha256:844abca4abd2162f8dde10be2fd3cd6331e9d746fa40cafa0334cabf0b933f24

Observation 91d0c985-1dcc-433f-9059-aaa8909fc1e0 · inbound

Large Language Models based ASR Error Correction for Child Conversations cites this paper.

Large Language Models based ASR Error Correction for Child Conversations SUPERB: Speech processing Universal PERformance Benchmark

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T15:07:32.131218Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:07:32.131218Z digest=sha256:138e40cabea9a3cd120727183e5b2df4a49b5565ec2a875c47f4f73b2e2d659e

Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · inbound

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation cites this paper.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.829556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.829556Z digest=sha256:e7a8c61d8b7d014a3db4eed0bfdf8cf78822449a234591762c61437eb17b47c3

Observation 23de2174-1071-4659-8a82-a16a72715415 · inbound

StressTest: Can YOUR Speech LM Handle the Stress? cites this paper.

StressTest: Can YOUR Speech LM Handle the Stress? SUPERB: Speech processing Universal PERformance Benchmark

Reference 6

Resolution
metadata mismatch
arxiv_id, observed 2026-05-19T13:02:18.245043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-19T13:00:23.002962Z digest=sha256:f49413ad29f31bd4d48aa378e1d7d76f6cd61c5bbaf3e886777a98debf643a41

Observation b7fa7fea-a8de-4bd5-b118-5815cf44f159 · inbound

Continual Speech Learning with Fused Speech Features cites this paper.

Continual Speech Learning with Fused Speech Features SUPERB: Speech processing Universal PERformance Benchmark

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:46:59.069395Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:46:59.069395Z digest=sha256:4c35c24ea1e8e4af4c78b461f10dd09e82e30100182b14d8a7e5e2ddb61a125d

Observation d8de2a0e-6168-4d88-ab94-a6a625aaacae · inbound

Joint ASR and Speaker Role Tagging with Serialized Output Training cites this paper.

Joint ASR and Speaker Role Tagging with Serialized Output Training SUPERB: Speech processing Universal PERformance Benchmark

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T04:32:30.437423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:32:30.437423Z digest=sha256:e3e841324f14c633b44d9b208c3c909c87aafc35ec9dc02e72cdd19fff9fe8b7

Observation f4a97ade-ec12-49e6-9a14-2af222fe9663 · inbound

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases cites this paper.

FixTalk: Taming Identity Leakage for High-Quality Talking Head Generation in Extreme Cases SUPERB: Speech processing Universal PERformance Benchmark

Reference 84

Resolution
unresolved
no resolver link, observed 2026-08-06T20:58:31.169287Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T20:58:31.169287Z digest=sha256:6fc10543e143704407a4d751b3af9a5cafc3c534fd01a473e080436a24399ac1

Observation 724aaa79-c9b0-4550-bf4b-3bd285639038 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation SUPERB: Speech processing Universal PERformance Benchmark

Reference 48

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:59.287586Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:59.287586Z digest=sha256:d2510ee33ce1ed9d3497a1d3854842981e7cfb120fda18c699c2d3f9cc6d0b50

Observation 07b7f3c5-28e0-411e-80db-09449c0dffa8 · inbound

Towards Robust Speech Recognition for Jamaican Patois Music Transcription cites this paper.

Towards Robust Speech Recognition for Jamaican Patois Music Transcription SUPERB: Speech processing Universal PERformance Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-06T17:26:27.022505Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:26:27.022505Z digest=sha256:899636766a5eca2d7c460b0863cf93ab4e21e15d59429858d2aaed27e234c1e6

Observation 4660fb3d-fdf7-40e6-89ae-fe74d8854c43 · inbound

Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space cites this paper.

Multi-Modal Motion Retrieval by Learning a Fine-Grained Joint Embedding Space SUPERB: Speech processing Universal PERformance Benchmark

Reference 42

Resolution
unresolved
no resolver link, observed 2026-08-06T11:07:20.126818Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T11:07:20.126818Z digest=sha256:b7ba2ec7f5f772bf280f3ca238f22ac80db56e2d73952168751c2a5197edc87d

Observation d35525ca-2a30-43d0-9329-ae571f573311 · inbound

Representing Speech Through Autoregressive Prediction of Cochlear Tokens cites this paper.

Representing Speech Through Autoregressive Prediction of Cochlear Tokens SUPERB: Speech processing Universal PERformance Benchmark

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T19:54:58.546041Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:54:58.546041Z digest=sha256:e99badc16ae6d4f7c43ea854380ec81c69be75fb10bfdcbdbebadbb978fdb5f7

Observation b381d03e-da15-48d3-bd19-0adf626e619d · inbound

Deformation Driven Suction Cups: A Mechanics-Based Approach to Wearable Electronics cites this paper.

Deformation Driven Suction Cups: A Mechanics-Based Approach to Wearable Electronics SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T19:46:48.209823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:46:48.209823Z digest=sha256:64e8a16368c6c0f40cf8f76f880e4c6fe9e049826e9d6709d6da37b3b73d2d34

Observation 925214f4-052a-4fc9-9a96-b3d59c39ffe5 · inbound

AVEX: What Matters for Animal Vocalization Encoding cites this paper.

AVEX: What Matters for Animal Vocalization Encoding SUPERB: Speech processing Universal PERformance Benchmark

Reference 27

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T22:16:51.709697Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T22:15:48.885339Z digest=sha256:74d71783e1d04837fec52bfc91fd28b28517faca031542dc0ae026df33c0053b

Observation 5f4738f6-a9f9-4743-b03f-76a552e53c59 · inbound

Multiple-Noise-Resilient Nonadiabatic Geometric Quantum Control of Solid-State Spins in Diamond cites this paper.

Multiple-Noise-Resilient Nonadiabatic Geometric Quantum Control of Solid-State Spins in Diamond SUPERB: Speech processing Universal PERformance Benchmark

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T19:39:22.296512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T19:39:22.296512Z digest=sha256:f72a4993fa52f6af1ac578ee1a04b8e271a5120cf9c6aa6406ff1346e1dffa1e

Observation 5d9a2068-4691-4b0e-a7a7-c09ff5e2cc0c · inbound

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation cites this paper.

AudioCodecBench: A Comprehensive Benchmark for Audio Codec Evaluation SUPERB: Speech processing Universal PERformance Benchmark

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T11:41:02.892997Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T11:41:02.892997Z digest=sha256:fa2147837ecd8dc369d84b5f0744e85097a17cb61c1eebe79c4463695ee89d5e

Observation a697c031-e64b-4621-b4ed-b4c604146834 · inbound

Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition cites this paper.

Joint Learning using Mixture-of-Expert-Based Representation for Speech Enhancement and Robust Emotion Recognition SUPERB: Speech processing Universal PERformance Benchmark

Reference 50

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T18:11:42.900862Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-18T18:07:34.965356Z digest=sha256:e007f5e406c520218151a7fa138437ed2afd17eec8b7bfd2bf9161314cc7f285

Observation f984d490-1f1f-4575-8b54-3c91fb536ed3 · inbound

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining cites this paper.

ParaSpeechCLAP: A Dual-Encoder Speech-Text Model for Rich Stylistic Language-Audio Pretraining SUPERB: Speech processing Universal PERformance Benchmark

Reference 26

Resolution
unresolved
no resolver link, observed 2026-07-13T16:09:50.207908Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-13T16:09:50.207908Z digest=sha256:5ecbe2ee62899bf8c2d1bd1db19329c722182ab81685d0f7c1b3672d278dee6e

Observation 9102b529-8e2c-43ce-90e9-6f0d6c407315 · inbound

ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals cites this paper.

ULTRAS -- Unified Learning of Transformer Representations for Audio and Speech Signals SUPERB: Speech processing Universal PERformance Benchmark

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:25:54.281562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-10T18:30:30.431777Z digest=sha256:6561a3c293a13c9237faf7de952f768b2d5b88cb275de1a11bbd5cef9657a38d

Observation 1d174128-92b1-4b5d-90be-656fa86a94bc · inbound

Alethia: A Foundational Encoder for Voice Deepfakes cites this paper.

Alethia: A Foundational Encoder for Voice Deepfakes SUPERB: Speech processing Universal PERformance Benchmark

Reference 45

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T15:41:33.505238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-09T19:27:59.124425Z digest=sha256:7c6cb8758d11c9156bd3544051852d56d39d64102c44d789eeaee6c5c996e760

Observation 2e8d27b2-891d-4f3f-bca1-e5eff96c860f · inbound

Multi-layer attentive probing improves transfer of audio representations for bioacoustics cites this paper.

Multi-layer attentive probing improves transfer of audio representations for bioacoustics SUPERB: Speech processing Universal PERformance Benchmark

Reference 9

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T06:31:28.421913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-12T04:11:17.994180Z digest=sha256:8ae487216ed4bb4485a8c7ace180139ea460888dcaa3a3e3f1d30858b194500c

Observation 4bcd3ac3-e213-4e36-8687-991e32b4b297 · inbound

AudioMosaic: Contrastive Masked Audio Representation Learning cites this paper.

AudioMosaic: Contrastive Masked Audio Representation Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-15T01:53:28.854680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-05-15T01:52:01.164694Z digest=sha256:3cad5bec1f964628ad5213d0e92369cffa2e6b1536ff0c829ee382f5f1b2ef8a

Observation 458ee8da-b048-4de5-9728-65bcad93a99e · inbound

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages cites this paper.

Multilingual and Multimodal LLMs in the Wild: Building for Low-Resource Languages SUPERB: Speech processing Universal PERformance Benchmark

Reference 228

Resolution
verified exact
arxiv_id, observed 2026-05-20T14:38:21.758423Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=arxiv_source observed=2026-05-20T14:33:36.100966Z digest=sha256:a40f9459723ab2b03812c7660dcea617511d8a5ee566c745a14d1247a164cb74

Observation 26975453-eb90-4d85-87ab-71f476c7aee6 · inbound

A Unified and Reproducible Experimentation Framework for Speech Understanding cites this paper.

A Unified and Reproducible Experimentation Framework for Speech Understanding SUPERB: Speech processing Universal PERformance Benchmark

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:16:12.193979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:3b354d37379d86378808d5b88ba52001e6834b402edf739db9c1f4de17b338cb

Observation 82c14792-db0e-40de-9be4-3c637d99cb18 · inbound

MOSS-Audio Technical Report cites this paper.

MOSS-Audio Technical Report SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T00:56:25.070239Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-28T13:05:29.813707Z digest=sha256:d0d91153742224df2b57f66b341fd795432c0d2066d1e9b1f148aa56eaec2c61

Observation 3342e72b-077b-4df4-ae0a-95680c353790 · inbound

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails cites this paper.

SEAM: Shortcut-Aware Real-Time Detection of Scripted vs. Spontaneous Speech for Interview Guardrails SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-02T19:47:19.562187Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-27T21:19:56.932689Z digest=sha256:f230c8d44a5204a7f3b337d9edb3fe90f6c05239f61e8a18dbf785b56faf816a

Observation 2906d0af-dab0-419d-b7fb-6f3e014dc63a · inbound

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning cites this paper.

S-JEPA : Soft Clustering Anchors for Self-Supervised Speech Representation Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:19:24.216627Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T19:46:47.653439Z digest=sha256:0be70032e27422b4b9c51507b0c5774b73a58e0878271005c53a9927d51ec9dc

Observation fdcd89da-18a7-46e6-8099-6ac99c19f2e2 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T06:39:37.328046Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T14:19:59.573391Z digest=sha256:6dc3d8e5975ccae9ef306f9ce8852ad2de7c34388e33c7a00e1a790025f482ae

Observation c6760762-9e33-4456-8c11-2dda8ff85c86 · inbound

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack cites this paper.

Impact Analysis of Speech Representation Learning Models for Acoustic Side-Channel Attack SUPERB: Speech processing Universal PERformance Benchmark

Reference 21

Resolution
metadata mismatch
arxiv_id, observed 2026-06-29T19:33:54.408623Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-29T04:44:52.537405Z digest=sha256:34eccf98144a0c34a300b3950e95297f44b0d46d6b760b17115b39bf0e334795

Observation e04976fb-eddd-48d7-817e-3474b7469bb9 · inbound

MSU-Bench: Towards Speaker-Centric Understanding in Conversational Multi-Speaker Scenarios cites this paper.

MSU-Bench: Towards Speaker-Centric Understanding in Conversational Multi-Speaker Scenarios SUPERB: Speech processing Universal PERformance Benchmark

Reference 18

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T11:49:50.798715Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T07:36:34.307652Z digest=sha256:bd611799e27c7342e0c7043725f20fd0ff9da6ba68c04c6435736723151c4cd5

Observation cd021a1a-d2c2-4ef3-a5e2-6faea7688acd · inbound

End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users cites this paper.

End-to-End Voice Intent Recognition for Spontaneous Human-Drone Interaction with Naive Users SUPERB: Speech processing Universal PERformance Benchmark

Reference 31

Resolution
metadata mismatch
arxiv_id, observed 2026-07-04T07:29:38.212507Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-26T13:30:12.101045Z digest=sha256:21cc8c5ca4b21e85bdb8bf3b476dbb38619782a153f9bc36e8b9c280b749e4d3

Observation 15289bd6-f251-4997-93be-5a9d1e8d1798 · inbound

SIGMA: Saliency-Guided Sparse Mask Attacks for Speech Emotion Recognition cites this paper.

SIGMA: Saliency-Guided Sparse Mask Attacks for Speech Emotion Recognition SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
metadata mismatch
arxiv_id, observed 2026-06-30T16:24:57.677906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-06-30T04:44:59.279562Z digest=sha256:16f615f086591fd45849aa6866121db78d273edd307cffc4b38bab9f4bdb3225

Observation f5bf67c4-e13c-4e07-ba87-265be4b59ea9 · inbound

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning cites this paper.

From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning SUPERB: Speech processing Universal PERformance Benchmark

Reference 111

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T05:56:39.882231Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-02T05:52:55.818877Z digest=sha256:4f6645d134898bd40813a6cbb0a2f4fd309f742f628c873f8133c09de63c98ed

Observation 7963645b-9cd7-4068-b5d1-df5bcd427f8c · inbound

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models cites this paper.

SPEARBench: A Benchmark for Naturalness Evaluation in Streaming Speech-to-Speech Language Models SUPERB: Speech processing Universal PERformance Benchmark

Reference 41

Resolution
metadata mismatch
local_arxiv, observed 2026-07-07T14:53:55.763035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.

source=pdf_text observed=2026-07-07T14:53:07.512543Z digest=sha256:9ea2fb425ed8a302ad7f82ffd6d510191c0ad46a80af030be05893a1655d8a2d

Observation b307710e-4c92-4ad4-8562-05cadfad0ae0 · inbound

Multi-Phonation Graph Learning with Self-Supervised Speech Embeddings for ALS Detection and Progression Prediction cites this paper.

Multi-Phonation Graph Learning with Self-Supervised Speech Embeddings for ALS Detection and Progression Prediction SUPERB: Speech processing Universal PERformance Benchmark

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-01T02:57:44.725464Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:57:44.725464Z digest=sha256:2c4b4684c0b423c9cd1c8f838669274001567a425de8602b9e66e9d6e3ba3ed9

Observation 00876789-57bd-4e7b-8726-22dd09ef36fa · inbound

Speaker Verification Under Real Classroom Conditions for English Speech cites this paper.

Speaker Verification Under Real Classroom Conditions for English Speech SUPERB: Speech processing Universal PERformance Benchmark

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-05T15:52:34.565904Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T15:52:34.565904Z digest=sha256:ca9873ef46558dc04fb4b6d53afdcd40fb29de0a16c76a471b5ed3bb013a37ac