Pith. sign in

Paper Citation Record · LEDGER

A Unified and Reproducible Experimentation Framework for Speech Understanding

As of 8 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2605.30899.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.30899 v1

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:20:16.428207Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-28T21:20:16.428207Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-01T20:16:12.170480Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact14
  • verified fuzzy0
  • unresolved23
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch5

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a08b0b3f-88ad-4c12-b587-f550f99826ed · outbound

This paper cites A Unified and Reproducible Experimentation Framework for Speech Understanding.

A Unified and Reproducible Experimentation Framework for Speech Understanding A Unified and Reproducible Experimentation Framework for Speech Understanding

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T20:16:12.171901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:211f7703f18fd871e557f5d056ab43479ce9cd6af6d1a4822b82678dfc9ac5e1

Observation c5ac17ea-1c7a-4c73-8da7-278055a21a5c · outbound

This paper cites paper + code.

A Unified and Reproducible Experimentation Framework for Speech Understanding paper + code

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:332b9b3c84d1e131496c9ca4bd301bce3fb2cabc828380044b459a6c7ad7525d

Observation 611098ca-aa15-45f7-baa9-c4075538a39b · outbound

This paper cites paper + code.

A Unified and Reproducible Experimentation Framework for Speech Understanding paper + code

Reference 3

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:aee37f4611042868c16858d54d317043a1f7ddee94367c6d3b9751a0f80f2857

Observation 7ca8c767-824d-4dfa-9268-2b79bd94ea7d · outbound

This paper cites an unresolved cited work.

A Unified and Reproducible Experimentation Framework for Speech Understanding Unresolved cited work

Reference 4

Resolution
malformed identifier
arxiv_id, observed 2026-07-01T20:16:12.178153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:2ed6311f7c104dd33c7a580880e20899cc7ef93a69e5d4a0cd9a4e39cb556a21

Observation ee1d9c1a-8bdc-4c60-83e3-9b3540d6affb · outbound

This paper cites an unresolved cited work.

A Unified and Reproducible Experimentation Framework for Speech Understanding Unresolved cited work

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:4674e7e534ee6b8d7b13f5399d9f3f881745f935886f60280751239f38abb310

Observation cad2034d-c349-4852-9d23-293119a8e38b · outbound

This paper cites paper + code.

A Unified and Reproducible Experimentation Framework for Speech Understanding paper + code

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:09b165c9a8f0abb0ee8e115670fb12574b550e58b8f411a349be1ade97e39bf1

Observation f0ce2598-9a3d-4844-8d93-f8ecb5b7b572 · outbound

This paper cites pa- per + code.

A Unified and Reproducible Experimentation Framework for Speech Understanding pa- per + code

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:a64c52a8a5d34aa6d44cd56ab2378e12fb6ece802134374d4a8f4263ceac2bf5

Observation 1b46f2b4-f9bf-42b8-b641-69056181a1fc · outbound

This paper cites On the landscape of spoken language models: A comprehensive survey,.

A Unified and Reproducible Experimentation Framework for Speech Understanding On the landscape of spoken language models: A comprehensive survey,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:b8543c7d6a5ebe891a7f4a71823eda61b2871e1d6da146125ca0c8541a96d2a8

Observation b6ef50bc-739b-4261-913d-0f7de7a8ab23 · outbound

This paper cites On The Landscape of Spoken Language Models: A Comprehensive Survey.

A Unified and Reproducible Experimentation Framework for Speech Understanding On The Landscape of Spoken Language Models: A Comprehensive Survey

Reference 9

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T20:16:12.174887Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:95410f0c29c0ddcf30f99eb4712bc3dc44b0004ea7ad6da535d069dfe089c4a1

Observation c9fedc37-37d6-4cd0-9544-9e7b67f2dd62 · outbound

This paper cites A survey on speech large language models for understanding.

A Unified and Reproducible Experimentation Framework for Speech Understanding A survey on speech large language models for understanding

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-06-28T21:22:38.016455Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:d87ac6ed33dbed6a255a1542aaa735a5a50d432ee779adf086a5bcfc1c07282a

Observation fcbbb6b4-3be4-4ff5-bf22-cfe277e9552e · outbound

This paper cites Measuring the accuracy of automatic speech recognition solutions,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Measuring the accuracy of automatic speech recognition solutions,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:2de53659cbfc539d0dd1194bf6d8173136753676e0b80049510a4cecc0923cd3

Observation 34581df1-2447-40aa-8f3c-310a40420344 · outbound

This paper cites Le Page, R.

A Unified and Reproducible Experimentation Framework for Speech Understanding Le Page, R

Reference 12

Resolution
metadata mismatch
doi, observed 2026-06-28T21:22:38.018502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:9ecb81264195ace3cf8c5cf3d98f4a1e91f047811e5abfc0829cddaec769fc2c

Observation 028abfee-17f1-4f7a-9327-34837f28b584 · outbound

This paper cites Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual Speech Recognition Evaluation.

A Unified and Reproducible Experimentation Framework for Speech Understanding Open ASR Leaderboard: Towards Reproducible and Transparent Multilingual Speech Recognition Evaluation

Reference 13

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.197182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:6b61cdba8101444dfd78a925f04e57cf74c1da1431acf469981deb7756d012c7

Observation fd395b8f-7161-46be-a5c2-834210b4977c · outbound

This paper cites Methodologies for the evaluation of speaker diariza- tion and automatic speech recognition in the presence of overlap- ping speech,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Methodologies for the evaluation of speaker diariza- tion and automatic speech recognition in the presence of overlap- ping speech,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:de3d4e13ac3c10196fe43a3fc6adecd593f96123a4d5c15d5c61af02d7744d53

Observation 4a290e82-d760-457e-bb0a-9a35c539a185 · outbound

This paper cites On the Evaluation of Speech Foundation Models for Spoken Language Understanding.

A Unified and Reproducible Experimentation Framework for Speech Understanding On the Evaluation of Speech Foundation Models for Spoken Language Understanding

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.214339Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:049f5fc85aa9a692dc170b718b030afc76cd054607f96ca5576bed94f3594eed

Observation 1f03175d-1081-4afc-9402-c9ad969f983d · outbound

This paper cites Speechr: A benchmark for speech reasoning in large audio-language models,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Speechr: A benchmark for speech reasoning in large audio-language models,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:af672a080c346d00da3bd76d7419df6d96d7e9c9a0dda3bc89ab35e1a94f76e0

Observation 94866213-157c-48b7-ae26-64f2e45be034 · outbound

This paper cites SpeechR: A Benchmark for Speech Reasoning in Large Audio-Language Models.

A Unified and Reproducible Experimentation Framework for Speech Understanding SpeechR: A Benchmark for Speech Reasoning in Large Audio-Language Models

Reference 17

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:16:12.181582Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:aaced335275a6c0046641fccee13d733c1be4db32910aa1e68f6a2d2530b0781

Observation 5cd47468-86f6-4c71-95c6-d379a91c8e6b · outbound

This paper cites SUPERB: Speech Processing Universal PERformance Benchmark,.

A Unified and Reproducible Experimentation Framework for Speech Understanding SUPERB: Speech Processing Universal PERformance Benchmark,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:5098e8d8127152d6007a08450e3b1123de048d979d2928a7e6c0ad53b3fb7ff6

Observation 64b9a5d4-a7a5-43a7-a8e7-24c7dd9b23bd · outbound

This paper cites Dynamic-SUPERB: Towards a Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Dynamic-SUPERB: Towards a Dynamic, Collaborative, and Comprehensive Instruction-Tuning Benchmark for Speech,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:94c603ee9a17e255e5690e919c19c9531d5cde0e23d44c61b0ade652c599c6a7

Observation dcbab749-a30d-4172-953e-aeee35d999ad · outbound

This paper cites AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension,.

A Unified and Reproducible Experimentation Framework for Speech Understanding AIR-Bench: Benchmarking Large Audio-Language Models via Generative Comprehension,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:4848e807f7eaf247e5dd1af500fe7f509c5c9829c1bbca9084cdc0203519f7e1

Observation 39901a07-00e3-4ca5-9b6d-7dbd6eee20d5 · outbound

This paper cites AudioBench: A Universal Benchmark for Audio Large Language Models,.

A Unified and Reproducible Experimentation Framework for Speech Understanding AudioBench: A Universal Benchmark for Audio Large Language Models,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:c46d8f4e57a7900a3c6169feb481d197930459302a8fd223aaa55df29df20eef

Observation 26975453-eb90-4d85-87ab-71f476c7aee6 · outbound

This paper cites SUPERB: Speech processing Universal PERformance Benchmark.

A Unified and Reproducible Experimentation Framework for Speech Understanding SUPERB: Speech processing Universal PERformance Benchmark

Reference 22

Resolution
metadata mismatch
arxiv_id, observed 2026-07-01T20:16:12.193979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:34306b5f4b30f75392bf6336bd0d1a2d13acd54c357a3778f56f35e0b878a3e8

Observation 3482c4a0-d002-48cb-9eff-b7b4591b78bd · outbound

This paper cites MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark.

A Unified and Reproducible Experimentation Framework for Speech Understanding MMAU: A Massive Multi-Task Audio Understanding and Reasoning Benchmark

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:16:12.207674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:5833fff52c1566dc1ffb4d35a3a997210798c0a929e72dff178f05e70c04d356

Observation 69d7a8f7-47e5-4d33-b6c2-5d1e157bf6b8 · outbound

This paper cites MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix.

A Unified and Reproducible Experimentation Framework for Speech Understanding MMAR: A Challenging Benchmark for Deep Reasoning in Speech, Audio, Music, and Their Mix

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.221545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:e78d11a73da0ffbd2cb361a9bb72042bddde771834dd6a06ed17e25d27753d34

Observation 7d969661-0ce0-4a91-b55e-2dd43ed37aee · outbound

This paper cites MeetEval: A Toolkit for Computation of Word Error Rates for Meeting Transcription Systems.

A Unified and Reproducible Experimentation Framework for Speech Understanding MeetEval: A Toolkit for Computation of Word Error Rates for Meeting Transcription Systems

Reference 25

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.224798Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:25a795c52a586808f33249ea46a2f608db4f444abfd04e5ddc638d7b45e2ff82

Observation 296ec7f2-ae2d-4e9f-b3f5-8716a85cd31e · outbound

This paper cites A call for clarity in reporting BLEU scores,.

A Unified and Reproducible Experimentation Framework for Speech Understanding A call for clarity in reporting BLEU scores,

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:daeb9496288a58bc088f0caaa99adbb72732deea6d6a6cd64222739a64f50cd6

Observation afe61cb0-722c-4aac-b777-01fe72d86fb1 · outbound

This paper cites VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.

A Unified and Reproducible Experimentation Framework for Speech Understanding VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.228406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:c32a63c9ccd13ca7894324f39dc0d5341711b1ce30ca376cc496b4ca16bb689e

Observation d3548ec0-aed3-4855-8934-4fd9c98a649c · outbound

This paper cites AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition.

A Unified and Reproducible Experimentation Framework for Speech Understanding AISHELL-5: The First Open-Source In-Car Multi-Channel Multi-Speaker Speech Dataset for Automatic Speech Diarization and Recognition

Reference 28

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.217425Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:e70f0e4719f7516fe63ed6e1185e92bd696fa912625f411617956f18868a453f

Observation 9a936dab-ecb5-4359-aa0d-0dccd86b425e · outbound

This paper cites The AMI meeting corpus,.

A Unified and Reproducible Experimentation Framework for Speech Understanding The AMI meeting corpus,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:ef24c4c1abfe24f1a582c7e5977c5985419f9a7ddeefd6a434b80fb00bcedd8f

Observation 5d7c5471-3f32-4c18-b587-d5149b49e0d0 · outbound

This paper cites M2MeT: The ICASSP 2022 multi-channel multi-party meeting transcription challenge,.

A Unified and Reproducible Experimentation Framework for Speech Understanding M2MeT: The ICASSP 2022 multi-channel multi-party meeting transcription challenge,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:deefe1e6b3db2dc5f04c24b9742a4c62442e58ec88116561243282913fe26a09

Observation c2a86ac6-f407-4206-9009-bfd1395d35d4 · outbound

This paper cites CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition.

A Unified and Reproducible Experimentation Framework for Speech Understanding CS-Dialogue: A 104-Hour Dataset of Spontaneous Mandarin-English Code-Switching Dialogues for Speech Recognition

Reference 31

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.204210Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:5f056be5e59da3a4c5870edbc9da71d03d15b95ddbc5c8aebd795e7cc6bfa1e3

Observation fb65c358-ef66-49dd-a175-f9c4b0ccb4b3 · outbound

This paper cites Ke- speech: An open source speech dataset of mandarin and its eight subdialects,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Ke- speech: An open source speech dataset of mandarin and its eight subdialects,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:d7bf8fd8db19f88018cd11bc1caf48aeda5e26fc4558842d46c639bcaa45190f

Observation 3274ed34-a22f-45c1-a99d-ea714381e61a · outbound

This paper cites ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark.

A Unified and Reproducible Experimentation Framework for Speech Understanding ContextASR-Bench: A Massive Contextual Speech Recognition Benchmark

Reference 33

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.204617Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:541f178d57a22a6b43fd1e9ec83ccc51f4e007d3428ec6d377a992bf6dd382c3

Observation 5618bc1a-9b9d-4e38-b275-6a2830f16ad7 · outbound

This paper cites VIBEVOICE-ASR technical report.

A Unified and Reproducible Experimentation Framework for Speech Understanding VIBEVOICE-ASR technical report

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.211050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:b48ab64025fff1ec6491184ad650ec9dc9d9838537b50583b72775f1bb05e3d8

Observation e8deb034-0f0e-4e2d-bff3-6c1b411910f6 · outbound

This paper cites Covost 2 and massively multilingual speech translation.

A Unified and Reproducible Experimentation Framework for Speech Understanding Covost 2 and massively multilingual speech translation

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:2a6aefcc8369ccb97f724feb51a31e810f237924b21ea1e0c681b6d5156883eb

Observation 9612d6d9-0555-4344-8feb-fdd26c7c38bc · outbound

This paper cites Swift: a scalable lightweight infras- tructure for fine-tuning,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Swift: a scalable lightweight infras- tructure for fine-tuning,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:77e3488cd49d87aeba45e5cd1d961ab28f6e92e9b1daa3647419aed3bc022f1a

Observation 6909e9fc-10d9-4ccc-a492-32d8a4700f93 · outbound

This paper cites Iemocap: Interactive emotional dyadic motion capture database,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Iemocap: Interactive emotional dyadic motion capture database,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:205c064f79c7e6cc28c7818585cbff1abf958b8ac83164af79bb148c90d51fcd

Observation 65b865b1-96b6-4c76-af54-b2f3f79292eb · outbound

This paper cites Meld: A multimodal multi-party dataset for emo- tion recognition in conversations,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Meld: A multimodal multi-party dataset for emo- tion recognition in conversations,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:a500459818f07c5d97b3e4c3f7d781c9ea5db14dca00b5150df8481025c57d7d

Observation b0c5d0c6-97f4-42db-9485-0aa650900d87 · outbound

This paper cites Slurp: A spoken language understanding resource package,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Slurp: A spoken language understanding resource package,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:98afe0ed2b1efd5b49cfde41e699718add6f69ea523897721b64379ed270e55e

Observation 962b0ebd-c620-4413-995a-90165ce58c4e · outbound

This paper cites MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark.

A Unified and Reproducible Experimentation Framework for Speech Understanding MMSU: A Massive Multi-task Spoken Language Understanding and Reasoning Benchmark

Reference 40

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:16:12.184460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:d09827c7f15783b4641d4736144def8257ebd90d2100782c6bfea9bb680c5eb9

Observation d25732a5-3a4d-42a7-98a9-d0d5d9ac5554 · outbound

This paper cites Qwen2-Audio Technical Report.

A Unified and Reproducible Experimentation Framework for Speech Understanding Qwen2-Audio Technical Report

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-07-01T20:16:12.190584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:8533b0b3dacd1cf32930b14355bd1fc427e0484860ee362dc326a447fdc8a30b

Observation 1ab4ecf5-88b8-478b-94ba-d2cf3e7171d9 · outbound

This paper cites Tasu: Text-only alignment for speech understanding,.

A Unified and Reproducible Experimentation Framework for Speech Understanding Tasu: Text-only alignment for speech understanding,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-28T21:20:16.428207Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:f6d0d68a0bd13107dd60ee10df5ea51ac71bea35b80c3cc6a77dcddad92fbcf2

Observation e49f5dfa-6b1f-4d10-aa27-266b3af7f066 · outbound

This paper cites Available: https://arxiv.org/abs/2511.03310.

A Unified and Reproducible Experimentation Framework for Speech Understanding Available: https://arxiv.org/abs/2511.03310

Reference 43

Resolution
verified exact
arxiv_id, observed 2026-07-01T20:16:12.187566Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:61462956887455b4a7e4a2b34fee1f2a0d07604b3ab1bc33b4670cf470cf9349

Pith citing papers

Observation a08b0b3f-88ad-4c12-b587-f550f99826ed · inbound

A Unified and Reproducible Experimentation Framework for Speech Understanding cites this paper.

A Unified and Reproducible Experimentation Framework for Speech Understanding A Unified and Reproducible Experimentation Framework for Speech Understanding

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-01T20:16:12.171901Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-06-28T21:20:16.428207Z digest=sha256:211f7703f18fd871e557f5d056ab43479ce9cd6af6d1a4822b82678dfc9ac5e1