Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 24 inbound Pith citation observations for arXiv:2104.09494.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T06:05:11.436227Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T16:08:37.813294Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation be99caf8-6851-409d-8e79-0bcec4553505 · inbound
Towards Improved Objective Perceptual Audio Quality Assessment -- Part 1: A Novel Data-Driven Cognitive Model NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 772a9847-6895-4a12-a802-c361164763e7 · inbound
Overview of the Amphion Toolkit (v0.2) NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 78
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be93759f-f1ab-42ed-97ef-c47e7ef11280 · inbound
Audio Large Language Models Can Be Descriptive Speech Quality Evaluators NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c12aaf43-7de6-45ee-adfd-7cfa2152ef9d · inbound
Metis: A Foundation Speech Generation Model with Masked Generative Pre-training NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0d180f06-89bc-49a4-8a9c-4eeb242cd7b7 · inbound
Muyan-TTS: A Trainable Text-to-Speech Model Optimized for Podcast Scenarios with a $50K Budget NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1926ba2a-b443-45ad-9567-ba7ca98d6dc9 · inbound
Towards Flow-Matching-based TTS without Classifier-Free Guidance NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 01cb78cd-3329-4f01-bac8-a992308224c2 · inbound
RoVo: Robust Voice Protection Against Unauthorized Speech Synthesis with Embedding-Level Perturbations NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation acce009c-0393-49ea-a7a9-d118aa8da34c · inbound
Audio Jailbreak Attacks: Exposing Vulnerabilities in SpeechGPT in a White-Box Framework NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c9c1005d-6da5-4701-ad4b-a3d196cfb89f · inbound
A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 270
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 07980a83-9295-4dc1-a55c-f5146b5b06e7 · inbound
AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 66d5267e-3512-4829-a45e-6f977e03a8cb · inbound
Schr\"odinger Bridge Mamba for One-Step Speech Enhancement NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bc17981-6d85-46b7-9710-5c344c87022d · inbound
VABench: A Comprehensive Benchmark for Audio-Video Generation NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b5c520e3-968d-4498-a4f7-828dc82af443 · inbound
GenTSE: Enhancing Target Speaker Extraction via a Coarse-to-Fine Generative Language Model NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation accabfca-7a05-4f46-97bc-7fe84ad890bc · inbound
LLM-Guided Reinforcement Learning for Audio-Visual Speech Enhancement NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fba3f46-deb1-4120-b68d-1e010ce83738 · inbound
Discrete Token Modeling for Multi-Stem Music Source Separation with Language Models NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation 286239c9-47a8-4c7a-ae16-bbadbbfbeb63 · inbound
Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation d1606140-ce0a-4f09-a6dc-4b181af79233 · inbound
JASTIN: Aligning LLMs for Zero-Shot Audio and Speech Evaluation via Natural Language Instructions NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation acb0c766-be82-4ead-aa38-b0a6b470f1f4 · inbound
A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation fd924b34-dd36-43b5-8c1a-a366a3914541 · inbound
Beyond Waveform Robustness: Robust Feature-Vocoder Adversarial Attacks on Automatic Speech Recognition NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation b6454594-db1d-4567-bde7-bf1accec18b5 · inbound
Feature-Aligned Speech Watermarking for Robustness to Reconstruction Distortions NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation c8922d62-d0fb-4c6a-990d-4c59296eed29 · inbound
Emo-LiPO: Listwise Preference Optimization for Fine-Grained Emotion Intensity Control in LLM-based Text-to-Speech NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.
Observation a3de0a8f-1d1c-4804-bd6a-984d304ecd90 · inbound
CS-ETS: Chaos-Inspired Samba-Based EMG-To-Speech Synthesis with Nonlinear Chaotic Losses NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb0b5c6e-2eb4-4a5e-9dac-db672dcc5553 · inbound
CallScreenBench: Benchmarking On-Device Models as Phone Secretaries NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19601c4e-067d-4dc4-ac06-f51e5edc8609 · inbound
Beyond Naturalness: Probing Automated Text-To-Speech Evaluators on Linguistically Grounded Dimensions NISQA: A Deep CNN-Self-Attention Model for Multidimensional Speech Quality Prediction with Crowdsourced Datasets
Reference 2026
Source-reported events for the cited work
Unavailable: canonical work link unavailable.