Pith. sign in

Paper Citation Record · LEDGER

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

As of 8 August 2026, this Paper Citation Record lists 44 of 44 outbound references and 2 inbound Pith citation observations for arXiv:2505.15061.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.15061 v1

Coverage vector

measured 44 of 44 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:28:03.066429Z

measured 46 of 46 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:27:58.440889Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:59:56.411210Z

Reference resolution

44 of 44 outbound references displayed

  • verified exact0
  • verified fuzzy35
  • unresolved9
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · outbound

This paper cites SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:f4933aad85940086c30c69d3430f3b12d9d44bc2f914586f62051899b9ff6a67

Observation 17b03289-f42a-4ff1-9120-16048fd7ddba · outbound

This paper cites Speech Quality Estimation: Models and Trends,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech Quality Estimation: Models and Trends,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.115842Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.485160Z digest=sha256:987de05ac6fd66a7c8ac37d5055b56b7108dc8d9f219f90a526c3d472db64c17

Observation 6cb42b73-4b7c-4e0d-b8d1-ca4a168bef27 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:08.125815Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:58.531292Z digest=sha256:d4d18b69daaed33c49baca898136df700f4c6a7a7e1caf2c5deea8114c7b8a86

Observation acdeb8db-a514-4603-be5a-a3073b640b40 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.650385Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.054106Z digest=sha256:d30082c2c762ccc27e04d2119689fd01ee05325e37115cb35083b37d24b45cec

Observation 02d4cf45-0602-4757-924f-479bd72ce4c7 · outbound

This paper cites an unresolved cited work.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Unresolved cited work

Reference 5

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:28:07.455618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.224815Z digest=sha256:1b74eb24701ba404cebaf23c669848c9fee1c00d63eb7ab54160952ebf17e526

Observation 6343ec3b-5f95-4d1f-8882-45224c0b29f6 · outbound

This paper cites Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Perceptual evaluation of speech quality (PESQ)-a new method for speech quality assessment of telephone networks and codecs,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.671159Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.021869Z digest=sha256:3aa4c01c11c96a79737c861afed506f74c0c88bd93d2341ecfd32559679dd409

Observation 6dc804ba-5b49-4974-bdae-b4568eb53419 · outbound

This paper cites Speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Speech quality assessment,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.310950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.348457Z digest=sha256:c9e8dd8427c705492a7dc3d9b57885f08a03c1cb3c62d3b3f293385d830feb4a

Observation 81616ce7-c9a7-4040-8ebb-725042a8ae56 · outbound

This paper cites MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MOSNet: Deep Learning-Based Objective Assessment for Voice Conversion,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.572907Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.299044Z digest=sha256:f6a2de35f5bbf5966f5bcba82afde2a6a73688ff61d60aef79dd1cb4943af28f

Observation 68135304-236a-4e45-b0f4-f58a04201fbe · outbound

This paper cites A review on subjective and objective evaluation of synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A review on subjective and objective evaluation of synthetic speech,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.931041Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.598535Z digest=sha256:9bbda6698524e7af4a5af6092f939c5d304c895c2c3622fb46bb41284636fc36

Observation 687b92b2-576d-49de-8a40-44c29a674ede · outbound

This paper cites An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An Al- gorithm for Intelligibility Prediction of TimeFrequency Weighted Noisy Speech,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.792909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:59.715325Z digest=sha256:7e54da5d718addf255570bf119c2e7131cd1ec83eeb8c7c189bae3f4e5823312

Observation de63a51d-0cb1-4d18-93a7-9c8d608a27fc · outbound

This paper cites Mel-cepstral distance measure for objective speech quality assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Mel-cepstral distance measure for objective speech quality assessment,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:59.834694Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:59.834694Z digest=sha256:d2da427011099728025be36300df0df34b38f1cbd09247ed6d77e83f685805e6

Observation fdf9fadf-3cfe-4cf3-aabe-06025d193970 · outbound

This paper cites The Voicemos Challenge 2024: Beyond Speech Quality Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2024: Beyond Speech Quality Prediction,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.136858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.865900Z digest=sha256:377ff46bad42824a428f1e2a4b900b0430f0eb0da966aa28ad96a71e57d0a2c5

Observation 3d832d82-6e22-446a-a8ad-6fc86894596d · outbound

This paper cites AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit AutoMOS: Learning a non-intrusive assessor of naturalness-of-speech

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:00.137013Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:00.137013Z digest=sha256:4f6c3e405a450bc4b312f1d4f47f38ffcc637433983a56c1a121587a20bf18d3

Observation 112ea42e-d165-4702-93a4-b64fc554ff7d · outbound

This paper cites Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Torchaudio-squim: Reference-less speech quality and intelligibility measures in torchaudio,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.985689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.998429Z digest=sha256:80d2969e377c3ff7012e0ea23e78801e2e265028e8627d536f9ccb9cfa835744

Observation 8e877be0-1bf8-40f4-93e3-24bc5b71b792 · outbound

This paper cites DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit DNSMOS: A Non- Intrusive Perceptual Objective Speech Quality Metric to Evaluate Noise Suppressors,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.428593Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.477433Z digest=sha256:4af33dc2724de90132be48a342649fef4f22cab6cacdf9a1afdcf5f3f2cc6da8

Observation 2d9b6c6b-c28b-4b88-aa88-7e6ee7546c46 · outbound

This paper cites The VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The VoiceMOS Challenge 2022,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.306170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.619961Z digest=sha256:7b0e27b44f251df9b4020c217e1f344aa8e86011a99da47d4a11fc23a24e2101

Observation 2482ec1a-57d7-4a76-97db-7b8ccef0660a · outbound

This paper cites The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Voicemos Challenge 2023: Zero-Shot Subjective Speech Quality Prediction for Multiple Domains,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.208074Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.754218Z digest=sha256:4a2488593430595030d11deff47a084b8e676e2413cab0c679cdc47c7e92fac5

Observation 6e572b51-e5f8-42c1-a471-a64f910c57d0 · outbound

This paper cites Generaliza- tion ability of MOS prediction networks,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Generaliza- tion ability of MOS prediction networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.792093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.379773Z digest=sha256:b70d1c2a8a85c66088e014369dd947aaf443fe54e732d92e62ff4dd4f762dc0f

Observation 5c1ce8ff-68a5-48db-97cb-bbf5d3200416 · outbound

This paper cites UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit UTMOS: UTokyo-SaruLab System for VoiceMOS Challenge 2022,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:06.056204Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:00.932454Z digest=sha256:82f7ae701e81376defc2da2d73602393fe76cf7fdb8d4b2761b9e3d713de6693

Observation 0e4c87d8-7628-4902-9837-75848e9a1bba · outbound

This paper cites How do voices from past speech synthesis challenges compare today?.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit How do voices from past speech synthesis challenges compare today?

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.602221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.602221Z digest=sha256:d6b49ae51a2ce3316e74fae7d79ba9e55c080e0ec7e7ec4b79507db14487a2f7

Observation d2e3d3b2-f745-46f9-929f-7e5ec3b75e7b · outbound

This paper cites SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.916549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.102877Z digest=sha256:e44d0e892cc6e21fb626e24dee6c9cf2c865e417e3002e25d6c48a1c9a196436

Observation 2ca48c2f-7018-4db2-8bf8-a2f6be8817d2 · outbound

This paper cites VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VERSA: A Versatile Evaluation Toolkit for Speech, Audio, and Music

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:01.173993Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:01.173993Z digest=sha256:07bf2b410657c8f16159d5bc4cd4b23d7cd54e30ac13c09c3355f5fe350d52a1

Observation 09396629-52f3-4690-b895-9a39664efe2e · outbound

This paper cites NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.862548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.260758Z digest=sha256:5b7c18deac7ee74165db6545cda6fff11c9ca31cbf48257ea8ac3998b6b64e43

Observation 6453a74d-c0b7-4f19-b1d8-8f0fac00084d · outbound

This paper cites Self-supervised speech representation learning: A review,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Self-supervised speech representation learning: A review,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-07T15:28:02.135446Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:28:02.135446Z digest=sha256:73dff91fe9ccfbb801281c82b1b48feb0560fed10d28c9976e8ebc35061300d4

Observation c2b7b22f-4ad8-4c73-91b0-4c51f9f0dc13 · outbound

This paper cites RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit RAMP: Retrieval- Augmented MOS Prediction via Confidence-based Dynamic Weighting,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.714739Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.445768Z digest=sha256:e018f9022900f80c1e84a102bebd000bb66f5f75e8caaced32aceb9fa4a63fea

Observation 936f9241-75e6-43bb-8913-339eb3930deb · outbound

This paper cites The Singing Voice Conversion Challenge 2023,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Singing Voice Conversion Challenge 2023,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.059446Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.269699Z digest=sha256:1a312330dd7bebd1724ea87adb702667a165f8d3259f65bac4a3212d5b469fc7

Observation 6294b4b8-8c7e-4460-af43-589f6412482f · outbound

This paper cites The Kaldi Speech Recognition Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit The Kaldi Speech Recognition Toolkit,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.645976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.743471Z digest=sha256:dbde898c57be23a76fd79ef69b6fad459bb0a7ee772a54e928c31820ea846ac0

Observation 86aaaa1a-b7c7-4d64-82a6-0341515bb59b · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit ESPnet: End-to-End Speech Processing Toolkit,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.535217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:01.880783Z digest=sha256:7dcdc05900485905b831a8ea5690b9609f46976b168b6db30af89bd51f6aaa2c

Observation 07f8f61e-1b87-42df-9ecf-7ae14f9ea84e · outbound

This paper cites An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit An End- To-End Non-Intrusive Model for Subjective and Objective Real- World Speech Assessment Using a Multi-Task Framework,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.422256Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.016260Z digest=sha256:d6f8ec982379090b657778f340979ea744a8943d62aed9ea5dadbd6e83ce3897

Observation 0d1a0ec6-69e4-4859-98c9-6c1d8c39dd40 · outbound

This paper cites Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Espnet-TTS: Unified, Reproducible, and Integratable Open Source End-to-End Text-to- Speech Toolkit,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.228324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.452285Z digest=sha256:2f55f93d8deb2c9ab76574e732eacf8b2c3e544701ab8f2e2723f7344cbb1614

Observation fd44c759-bdfc-4bd8-b71d-c3ecc88d89b8 · outbound

This paper cites Utilizing Self-Supervised Representations for MOS Prediction,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Utilizing Self-Supervised Representations for MOS Prediction,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:05.259311Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.239871Z digest=sha256:2cd265affab1bb7621f66ad75a8443351556c04bd07671fc7198a10618cdf450

Observation 23d6e7d9-910d-4aeb-bd38-8534d5b49be5 · outbound

This paper cites On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit On the NISQA dataset, on average, the WavLM large model [33] and the XLS-R 1b model [34] achieved the best and second best scores on the Sys MSE and Sys SRCC metrics, respectively

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.871031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:58.911703Z digest=sha256:48d25761b6714abc2b8455a5930616fa396ab03d4a782af938f66b1bc0e02f93

Observation d9a2b204-f6f0-48eb-92f0-453313c43342 · outbound

This paper cites MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit MB- NET: MOS Prediction for Synthesized Speech with Mean-Bias Network,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.888066Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.282769Z digest=sha256:83e3cdb00c6cf177d88d14e3901181f2186e1aafd9398fe20a2eb1368be1a5b5

Observation ccae44fd-d877-47cd-ac64-edd7305536fd · outbound

This paper cites LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit LDNet: uni- fied listener dependent modeling in MOS prediction for synthetic speech,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.675473Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.320774Z digest=sha256:1a0aaafbb991e4159eb32ced1b16b75821423319ff62d5ac288ee76573284459

Observation 41da7b5e-66d8-4ff5-b61e-2f57760076c8 · outbound

This paper cites Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Alignnet: Learning dataset score align- ment functions to enable better training of speech quality estima- tors,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:04.443873Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.393441Z digest=sha256:fdc0ab8cef2a1c63c0df9096303449be7e184305f87decbd0aea6b0f25c11ecd

Observation b1532b15-934d-4afe-9c62-277a48251fe8 · outbound

This paper cites VoxSim: A perceptual voice similarity dataset,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit VoxSim: A perceptual voice similarity dataset,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.451526Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.779806Z digest=sha256:a72012b8592b501c6e7b81f1626c22e99f10c7167a9558a082a2662106a99e9e

Observation 87b739b1-812e-4f75-9d76-43b334dc497b · outbound

This paper cites A Large- Scale Evaluation of Speech Foundation Models,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit A Large- Scale Evaluation of Speech Foundation Models,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.985643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.508013Z digest=sha256:317a94136ab9d7c53f6109ee71ad64a83531b0d8ff4214cd336512de3718cd7b

Observation fc11a089-7fa0-45a6-b6b9-c2ddf825a917 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.805292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.558645Z digest=sha256:ea9e0c41e460f7b09ed66cacc37860ff643da72304e0815ddb9b521e40e13b2a

Observation e82ef2e9-c752-404e-98af-c00c7a450908 · outbound

This paper cites WavLM: Large-scale self-supervised pre-training for full stack speech processing,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit WavLM: Large-scale self-supervised pre-training for full stack speech processing,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.662765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.605498Z digest=sha256:b5bbcc437f316bbb1c35148e74071f246cfed3849a73045d2a81fd62554b674d

Observation 6f803eb1-b3c2-4419-912d-140668090ed4 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit XLS-R: Self-supervised Cross-lingual Speech Rep- resentation Learning at Scale,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.596764Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.647836Z digest=sha256:a1d03632a4ff45f154c80ff10a8b356242f13d35e049975a6f990f762764edae

Observation bbc1f6f4-160d-4abb-816d-fa9e27ff09f0 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.514832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.696851Z digest=sha256:ac619f58b7a52cf535dc782699a3b1c917d2073852c3350def075e73598a9f20

Observation e3f0447e-2af3-4381-a9c1-b093fdf438f9 · outbound

This paper cites PAM: Prompting Audio- Language Models for Audio Quality Assessment,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit PAM: Prompting Audio- Language Models for Audio Quality Assessment,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.392977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:02.943151Z digest=sha256:c566bcc983de9f3097a38b634e68c4a07863d5281d37e0c8763de27eff1f368d

Observation 72a379e7-d0ab-465d-ad96-9a05c8242d2d · outbound

This paper cites Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Audio Large Language Models Can Be Descriptive Speech Quality Evaluators,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:03.305134Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:28:03.066429Z digest=sha256:c05d6636b4aaf71c9b73e5bc73e83a7e0787c88501c0d476b2136528e8ae9e28

Observation 964e5b53-eb6d-44da-b0a7-50f33ac1377f · outbound

This paper cites Each sample was rated by 8 distinct listeners.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit Each sample was rated by 8 distinct listeners

Reference 2022

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:28:07.963154Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T15:27:58.708724Z digest=sha256:7bbed85263889fd962c29531d6fdc94134df74be66e3b7dd6ec8c2496335e3d7

Pith citing papers

Observation c51a7411-f08c-4fc3-a1fc-95f2ca59c90d · inbound

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit cites this paper.

SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T15:27:58.440889Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:27:58.440889Z digest=sha256:f4933aad85940086c30c69d3430f3b12d9d44bc2f914586f62051899b9ff6a67

Observation 7b40b310-2311-414f-955d-0e1c5e80c18d · inbound

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment cites this paper.

Improving Speech Enhancement with Multi-Metric Supervision from Learned Quality Assessment SHEET: A Multi-purpose Open-source Speech Human Evaluation Estimation Toolkit

Reference 41

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:59:56.456405Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T00:59:55.643986Z digest=sha256:e7c8fbb3dedd4755bd704e98f5901fb4858d1e9bc891084b3d78411868b8cfb8