Pith. sign in

Paper Citation Record · LEDGER

MOSNet: Deep Learning based Objective Assessment for Voice Conversion

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 14 inbound Pith citation observations for arXiv:1904.08352.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1904.08352 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 14 of 14 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 14 of 14 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-15T21:07:38.924508Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-19T20:27:53.680114Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation a14080d4-9d88-469f-b961-6561f243f3f9 · inbound

Scaling Transformers for Low-Bitrate High-Quality Speech Coding cites this paper.

Scaling Transformers for Low-Bitrate High-Quality Speech Coding MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-12T05:56:42.817221Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-12T05:56:42.817221Z digest=sha256:7ce74a06b21a8387855c78d8bfe0ebc006266c053c97de2459e0b2512ef596dc

Observation 4ee01c43-168f-4f9c-b1ca-a1a4bb7c612e · inbound

DiffAttack: Diffusion-based Timbre-reserved Adversarial Attack in Speaker Identification cites this paper.

DiffAttack: Diffusion-based Timbre-reserved Adversarial Attack in Speaker Identification MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-10T21:21:12.169911Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:21:12.169911Z digest=sha256:ced342e6518499c538dc7a4714d9e4c45ced0247c03bd5b311a10f52b45bd539

Observation e338897b-139b-4f6c-89cd-c03b2dc76461 · inbound

Audio Large Language Models Can Be Descriptive Speech Quality Evaluators cites this paper.

Audio Large Language Models Can Be Descriptive Speech Quality Evaluators MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T12:30:52.080500Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T12:30:52.080500Z digest=sha256:daf977c9f6c1c2aa7ebc7fa277fef206e8fc3918fe98e89437d47d16222ad0a7

Observation 35500a2e-48ca-4ba1-8b71-f261a0836319 · inbound

Visual-based spatial audio generation system for multi-speaker environments cites this paper.

Visual-based spatial audio generation system for multi-speaker environments MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-08T12:30:26.197293Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-08T12:30:26.197293Z digest=sha256:0540ef023bd8c2847a5cf505c4fde58c0e22f39b616c6dc39ed2f3a69f57c850

Observation f77055db-272c-4d31-98c1-a1fa46df0db5 · inbound

SongEval: A Benchmark Dataset for Song Aesthetics Evaluation cites this paper.

SongEval: A Benchmark Dataset for Song Aesthetics Evaluation MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-15T21:07:38.924508Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T21:07:38.924508Z digest=sha256:029761eb49de689b16c68e0346ba5ce6d668c44d54d84460a8a5a7509ace03ec

Observation ed080602-d705-4148-958f-1e36bc095878 · inbound

SALF-MOS: Speaker Agnostic Latent Features Downsampled for MOS Prediction cites this paper.

SALF-MOS: Speaker Agnostic Latent Features Downsampled for MOS Prediction MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T11:44:17.680497Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:44:17.680497Z digest=sha256:2d05793faa70fd8db0d76aa0cfd1abc7ab75006e0825fb187122575ade40c605

Observation 26cf5952-8d46-4441-89ac-e9ce9fafdf9e · inbound

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations cites this paper.

A Survey of Automatic Evaluation Methods on Text, Visual and Speech Generations MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 229

Resolution
unresolved
no resolver link, observed 2026-08-07T10:17:47.065365Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:17:47.065365Z digest=sha256:b530b46519293ba7951f3bc998ad9583e68d877be5c0a3dc328af3ed5c413518

Observation f4941141-288b-45c5-a6f4-b395f75250d2 · inbound

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation cites this paper.

AudioJudge: Understanding What Works in Large Audio Model Based Speech Evaluation MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-06T16:47:56.555197Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-06T16:47:56.555197Z digest=sha256:72cd9fdcf09ebec280b977adb3fab0e3bf8424c43c60a38d58f3a3490c0e96b3

Observation 10b34981-1018-4791-b2c0-c0e76fd61647 · inbound

Neural networks for Text-to-Speech evaluation cites this paper.

Neural networks for Text-to-Speech evaluation MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 4

Resolution
verified exact
arxiv_id, observed 2026-05-15T09:49:54.671517Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-15T09:46:26.884551Z digest=sha256:6d320f45d54c1a95215ac4fae3e255446ce7ed86373ea60d539fe1fc36236749

Observation fa24a8eb-4b44-4c14-9ecb-1da1c7e80c75 · inbound

Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment cites this paper.

Voice Mapping of Text-to-Speech Systems: A Metric-Based Approach for Voice Quality Assessment MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-10T01:10:09.302242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-10T01:07:50.243903Z digest=sha256:13071ad9cc961627baaac3345c618fbce55cf5709981315d66b181a4cd00aafb

Observation 366c75ff-6f51-404e-b485-900a1c3b501d · inbound

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models cites this paper.

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 41

Resolution
verified exact
arxiv_id, observed 2026-05-19T20:27:53.684089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-19T20:26:59.049472Z digest=sha256:866b5ddb6027bb80d344f3260352f9496860b24f4fc7c5b772ba4f492a7db0cd

Observation 6e732adf-c477-44e6-9729-136ae39cab44 · inbound

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers cites this paper.

Memory Efficient Audio Synthesis with Decoupled Temporal Depth Diffusion Transformers MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 2019

Resolution
unresolved
no resolver link, observed 2026-07-30T11:48:49.502742Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-30T11:48:49.502742Z digest=sha256:4273b553685837b4f20a9bd23db29167d9cf736230bef46aa6f38023ebff9519

Observation b7e114da-fed1-48fc-a2d2-c61d0c4e4166 · inbound

Do Audio Language Models Use Paralinguistic Evidence? Counterfactual Audits for Response Evaluation cites this paper.

Do Audio Language Models Use Paralinguistic Evidence? Counterfactual Audits for Response Evaluation MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-10T21:57:54.730415Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T21:57:54.730415Z digest=sha256:1d89b0537443df1fffc8a8183010c83f8d68249ea42987fec5b4372e5f7b4aa2

Observation 778db9d8-3f12-4828-bdcb-562d76064021 · inbound

Do Audio Language Models Use Paralinguistic Evidence? Counterfactual Audits for Response Evaluation cites this paper.

Do Audio Language Models Use Paralinguistic Evidence? Counterfactual Audits for Response Evaluation MOSNet: Deep Learning based Objective Assessment for Voice Conversion

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-14T04:43:45.022396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-14T04:43:45.022396Z digest=sha256:423092ba9fba027e3b58ad40b05897f541a3e148b9e86e1de93c214e9edb1744