Pith. sign in

Paper Citation Record · LEDGER

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

As of 18 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 2 inbound Pith citation observations for arXiv:2505.15061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15061 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:28:03.066429Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:27:58.440889Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:59:56.411210Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · outbound

This paper cites SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:da787e28429b5b6b5887b043838ff1ec2aec88de515964506830c6aa7d381b72

Observation 17b03289-f42a-4ff1-9120-16048fd7ddba · outbound

This paper cites Speech Quality Estimation: Models and Trends,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech Quality Estimation: Models and Trends,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.115842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.485160Z digest=sha256:a18ae8cd0f83471ffa215512c7964bba6e55686bfda97722e4656606928a56af

Observation 6cb42b73-4b7c-4e0d-b8d1-ca4a168bef27 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:08.125815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:58.531292Z digest=sha256:b1a18f4eb647da6bdb8ec89d79d3b122bdde87630f1002e626999df890fa5c49

Observation acdeb8db-a514-4603-be5a-a3073b640b40 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.650385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.054106Z digest=sha256:839c3c74ece9748c6347d9d28c8151e20fb36054c32344ed226e3ed1a2f99617

Observation 02d4cf45-0602-4757-924f-479bd72ce4c7 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.455618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.224815Z digest=sha256:a148c3fde4d553c12c3e4cfcd7f0164d3ea43a07a5bdf852ceb98ecf81022941

Observation 6343ec3b-5f95-4d1f-8882-45224c0b29f6 · outbound

This paper cites Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.671159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.021869Z digest=sha256:fb3b2c0a3ab6e76cf99dd551c1daa1a34149aff09bec4b1eeac209422a1d8f4f

Observation 6dc804ba-5b49-4974-bdae-b4568eb53419 · outbound

This paper cites Speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech quality assessment,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.310950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.348457Z digest=sha256:e9753a3071168ae4f6f555167fd8492f38bf286108c7bb837823f34866bd347c

Observation 81616ce7-c9a7-4040-8ebb-725042a8ae56 · outbound

This paper cites MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.572907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.299044Z digest=sha256:11e699ca261ca7c48c33fc3db08d34e71a468be1decf4bdd61bc0d847e2beb9d

Observation 68135304-236a-4e45-b0f4-f58a04201fbe · outbound

This paper cites A review on subjective and objective evaluation of synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A review on subjective and objective evaluation of synthetic speech,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.931041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.598535Z digest=sha256:47ef7670adf23d00503d09362b3bd0ccf2c9b745bf4b9455fd8accc37640220f

Observation 687b92b2-576d-49de-8a40-44c29a674ede · outbound

This paper cites An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.792909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:59.715325Z digest=sha256:aafad27b62059d17a911920b5cf94cf8a8bd2e90cc725168ad043cb1ad603f0f

Observation de63a51d-0cb1-4d18-93a7-9c8d608a27fc · outbound

This paper cites Mel-cepstral distance measure for objective speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Mel-cepstral distance measure for objective speech quality assessment,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:59.834694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:59.834694Z digest=sha256:290ce0ab208412d29875cdfc4b9b99847606adcd8538a0a02dcc5691081eacf8

Observation fdf9fadf-3cfe-4cf3-aabe-06025d193970 · outbound

This paper cites The Voicemos Challenge 2024: Beyond Speech Quality Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2024: Beyond Speech Quality Prediction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.136858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.865900Z digest=sha256:3399f777cb17258516c5b18f9010df86f67fc423a0f0011572744bfb6bad1b1c

Observation 3d832d82-6e22-446a-a8ad-6fc86894596d · outbound

This paper cites AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:00.137013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:00.137013Z digest=sha256:1e4b1e4dddda971cf21e0a6cd12dc4bf81799e8b606f261c0ee895f08ff9324d

Observation 112ea42e-d165-4702-93a4-b64fc554ff7d · outbound

This paper cites Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.985689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.998429Z digest=sha256:9a1a2bfbc60d3d02704a692b62915736c11d9a734330aab053b9256caa56861c

Observation 8e877be0-1bf8-40f4-93e3-24bc5b71b792 · outbound

This paper cites DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.428593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.477433Z digest=sha256:29ce4b96e102867fcbe3a46d130acd92f4be7df13aa7d9cd3f914bd40be7b165

Observation 2d9b6c6b-c28b-4b88-aa88-7e6ee7546c46 · outbound

This paper cites The VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The VoiceMOS Challenge 2022,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.306170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.619961Z digest=sha256:78271db468a68a71b08d05a2f96e624de7848bedbb18b4320fe7a66234d0b1c9

Observation 2482ec1a-57d7-4a76-97db-7b8ccef0660a · outbound

This paper cites The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.208074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.754218Z digest=sha256:d912002b4b8341a8c0402a220fddd2e6feaeb62b17d6af94d81b2e3fff52753b

Observation 6e572b51-e5f8-42c1-a471-a64f910c57d0 · outbound

This paper cites Generaliza- tion ability of MOS prediction networks,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Generaliza- tion ability of MOS prediction networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.792093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.379773Z digest=sha256:57913212d751917ac10fb63661a78fb4584daf252384f99b9af461c408e8f056

Observation 5c1ce8ff-68a5-48db-97cb-bbf5d3200416 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.056204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:00.932454Z digest=sha256:5fa615ffd9838d03d7ee2eb738047ebcf89c39eb3f528e163ad9a1878a98bc6d

Observation 0e4c87d8-7628-4902-9837-75848e9a1bba · outbound

This paper cites How do voices from past speech synthesis challenges compare today?.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit How do voices from past speech synthesis challenges compare today?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.602221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.602221Z digest=sha256:9420a2ed59b36f978a7fe981d8bb38d3e084f59e07fc76102ac6e6972d536df3

Observation d2e3d3b2-f745-46f9-929f-7e5ec3b75e7b · outbound

This paper cites SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.916549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.102877Z digest=sha256:a2e871403864b621bc5ac3830ad53a626a6da0f6af77c55cd9b403b66bbaf468

Observation 2ca48c2f-7018-4db2-8bf8-a2f6be8817d2 · outbound

This paper cites VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.173993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.173993Z digest=sha256:7119efb83987d3c9ef71be37f2d6fdd3c66f280ff96c9e46d4c2025a33bf904c

Observation 09396629-52f3-4690-b895-9a39664efe2e · outbound

This paper cites NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.862548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.260758Z digest=sha256:3ca142ce81ef725feade3e34e13ac212e78fc2e6241286f693b4945b626e9dff

Observation 6453a74d-c0b7-4f19-b1d8-8f0fac00084d · outbound

This paper cites Self-supervised speech representation learning: A review,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Self-supervised speech representation learning: A review,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:02.135446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:02.135446Z digest=sha256:bdc1080b4b471cc287ede3754dfe1dca1c002a292c71eb5a05a67e4cfd8bcc98

Observation c2b7b22f-4ad8-4c73-91b0-4c51f9f0dc13 · outbound

This paper cites RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.714739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.445768Z digest=sha256:95588f9066ec15ee596b1f88b8cccb476363d88824765d3efcc90c4e7f3aaf56

Observation 936f9241-75e6-43bb-8913-339eb3930deb · outbound

This paper cites The Singing Voice Conversion Challenge 2023,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Singing Voice Conversion Challenge 2023,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.059446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.269699Z digest=sha256:1188203148be9a5a487a439947fb5f477c3d1e4d58f1b57ab121cde56c417c1e

Observation 6294b4b8-8c7e-4460-af43-589f6412482f · outbound

This paper cites The Kaldi Speech Recognition Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Kaldi Speech Recognition Toolkit,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.645976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.743471Z digest=sha256:6099ecc3da895ce5578a1df3dfea5fa51c5f1b6479569cb9638f560fb19a919d

Observation 86aaaa1a-b7c7-4d64-82a6-0341515bb59b · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit ESPnet: End-to-End Speech Processing Toolkit,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.535217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:01.880783Z digest=sha256:2db42afec2c17d5b0368707a5f42258e961dad5374b474b5647e604bc5041db4

Observation 07f8f61e-1b87-42df-9ecf-7ae14f9ea84e · outbound

This paper cites An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.422256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.016260Z digest=sha256:64f01bbbe3503413416a119c8a5bd579978e530a01f023bdfda14de2400f01e7

Observation 0d1a0ec6-69e4-4859-98c9-6c1d8c39dd40 · outbound

This paper cites Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.228324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.452285Z digest=sha256:0ac4fa5dc64dda622d652e0ac074967e7d50714102a96dd802fedca1453fa1ff

Observation fd44c759-bdfc-4bd8-b71d-c3ecc88d89b8 · outbound

This paper cites Utilizing Self-Supervised Representations for MOS Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Utilizing Self-Supervised Representations for MOS Prediction,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.259311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.239871Z digest=sha256:2337521e02c582a8dea22f915f06cc19205929cf224c01907415a9e7167f1c3b

Observation 23d6e7d9-910d-4aeb-bd38-8534d5b49be5 · outbound

This paper cites On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.871031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:58.911703Z digest=sha256:9d5774c7a8500c97af496d8db528c514f12361b0bf3d087802633086dec1f367

Observation d9a2b204-f6f0-48eb-92f0-453313c43342 · outbound

This paper cites MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.888066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.282769Z digest=sha256:dfae991818d01af73fc200f7143ee288e447284877769e0b8b10b58af598d39e

Observation ccae44fd-d877-47cd-ac64-edd7305536fd · outbound

This paper cites LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.675473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.320774Z digest=sha256:b4e3eb4f1bd397c139123e048ba64436e77354f5983c15206ca44cff301bbee5

Observation 41da7b5e-66d8-4ff5-b61e-2f57760076c8 · outbound

This paper cites Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.443873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.393441Z digest=sha256:5b7c17d6f8f9149e78fc5f7e5725bfc86aceb13ea58ef0af0ad58164d8218ff8

Observation b1532b15-934d-4afe-9c62-277a48251fe8 · outbound

This paper cites VoxSim: A perceptual voice similarity dataset,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VoxSim: A perceptual voice similarity dataset,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.451526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.779806Z digest=sha256:d3b4d670f73d7fb19fea2bd08a0d9443c5c91b67cc1b94b287360e49ff7c1eee

Observation 87b739b1-812e-4f75-9d76-43b334dc497b · outbound

This paper cites A Large- Scale Evaluation of Speech Foundation Models,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A Large- Scale Evaluation of Speech Foundation Models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.985643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.508013Z digest=sha256:49b4743580b7892e4f0550f4de83d0071f6af35cafbea4a06798f3d4341cbc97

Observation fc11a089-7fa0-45a6-b6b9-c2ddf825a917 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.805292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.558645Z digest=sha256:d666dcb46baa56a5521fd4c726dad41e31034037f4b6159597c452bf82bc931e

Observation e82ef2e9-c752-404e-98af-c00c7a450908 · outbound

This paper cites WavLM: Large-scale self-supervised pre-training for full stack speech processing,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit WavLM: Large-scale self-supervised pre-training for full stack speech processing,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.662765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.605498Z digest=sha256:3ef01030ad919903db5e7bed0dcb618a0a3a0053ca15a5193974d502f3eeeb97

Observation 6f803eb1-b3c2-4419-912d-140668090ed4 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.596764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.647836Z digest=sha256:404c99014c5cda595aef40b1e734bd876a71faf31a0293caaaebde0ec3471a9f

Observation bbc1f6f4-160d-4abb-816d-fa9e27ff09f0 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.514832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.696851Z digest=sha256:68d285b2b125cf3bacb890f582e3c37b22f700737f2815aa0b0a20f86a0aa107

Observation e3f0447e-2af3-4381-a9c1-b093fdf438f9 · outbound

This paper cites PAM: Prompting Audio- Language Models for Audio Quality Assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit PAM: Prompting Audio- Language Models for Audio Quality Assessment,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.392977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:02.943151Z digest=sha256:56e8b26ef90df52b76fdb0881ade692dbd907f440a8b8451b624ebad232b9e9a

Observation 72a379e7-d0ab-465d-ad96-9a05c8242d2d · outbound

This paper cites Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.305134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:28:03.066429Z digest=sha256:2ca0a29ff794773000ce6e47714e9fc9c67372b3d16def5a65b529ad640eaaeb

Observation 964e5b53-eb6d-44da-b0a7-50f33ac1377f · outbound

This paper cites Each sample was rated by 8 distinct listeners.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Each sample was rated by 8 distinct listeners

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.963154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T15:27:58.708724Z digest=sha256:cda2c5c5d036a9e985218722d80ed34159a07d4d3e068874f8f8d673bbf0d16c

Pith citing papers

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · inbound

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit cites this paper.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:da787e28429b5b6b5887b043838ff1ec2aec88de515964506830c6aa7d381b72

Observation 7b40b310-2311-414f-955d-0e1c5e80c18d · inbound

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment cites this paper.

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:59:56.456405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T00:59:55.643986Z digest=sha256:1efec016ed0f22f2994abcabefada743eafcfbe0a6cc0e0a7b37d5cf1fec6237