Pith. sign in

Paper Citation Record · LEDGER

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

As of 15 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2505.20506.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20506 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:28.634322Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-15T06:32:42.880941+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:25.595644Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:00:28.778797Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d9f8115b-88fa-4f4b-9ba0-7b5d2b52230e · outbound

This paper cites an unresolved cited work.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:33.316583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.491796Z digest=sha256:ef2fb57a52c67387333f2ace6a3095a26769912670352f2e8fe8994e08586f8d

Observation b3b01c96-e5a6-431e-a440-48ac43f529a0 · outbound

This paper cites ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:00:28.939908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.595644Z digest=sha256:28033076bc6bc0d63e09cb4f91e934f720856f00c623de27fa8c3ae9086f52f5

Observation dfb1b0c9-46c1-48b4-8a93-5b50995bb039 · outbound

This paper cites In this section, we describe each part of ArV oice and provide justifica- tion for design decisions where applicable.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis In this section, we describe each part of ArV oice and provide justifica- tion for design decisions where applicable

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:33.048002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.719894Z digest=sha256:325a0f90d1ab12313fafbabd52890d897d94de034655b5bcc65be8f249f128ae

Observation 53e56dec-cefd-45d0-a40e-a8f05d00d8b3 · outbound

This paper cites an unresolved cited work.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:32.857585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.864240Z digest=sha256:3b4c5312f37c55398dc4c2c60a1787137ae8d2f63f2d476802301d08138a35e4

Observation 504024d5-b896-4714-bf34-0f8b1f7a0a16 · outbound

This paper cites The dataset consists of 11 voices in total, 7 of which are human voices, and 4 are syn- thetic with parallel text.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis The dataset consists of 11 voices in total, 7 of which are human voices, and 4 are syn- thetic with parallel text

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.522013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.057794Z digest=sha256:85af772ba8c15a70c354a3ec14d45a2192a978f8ac5b7136db50dbc701dcc1a2

Observation 027b4f30-a22f-4281-9271-061b715c7bcc · outbound

This paper cites This work was partially funded by a Google research award (11/2023).

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis This work was partially funded by a Google research award (11/2023)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.321259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.148273Z digest=sha256:6c76ab99ec3a4c5d334f8e444f21763ed31a90ed8c8da85218e67c09eb80a5d2

Observation ab7b15dd-9064-48a4-b593-a83a87fad8c9 · outbound

This paper cites Cmu wilderness multilingual speech dataset,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Cmu wilderness multilingual speech dataset,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.404796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.742493Z digest=sha256:26b9ce32afc1a7c1a3407dd4c5e2dc93c71da8c8dc59f02aa18011acc547a0dd

Observation 86a121d7-422a-40da-86a4-9c12d8eed7b6 · outbound

This paper cites Speech recognition challenge in the wild: Arabic mgb-3,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Speech recognition challenge in the wild: Arabic mgb-3,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:26.224154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:26.224154Z digest=sha256:bac4dd8e924a5ccbec54f7922e1ee2207a4b9881913776616edfe7596014929f

Observation 5ee82de3-d52f-495f-a59d-c9f82452c7e5 · outbound

This paper cites QASR: QCRI aljazeera speech resource a large scale annotated Arabic speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis QASR: QCRI aljazeera speech resource a large scale annotated Arabic speech corpus,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.148453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.295851Z digest=sha256:0c5626fbf36ad4d234bd125087571f9650354cae2693175af09a7443358e0bc6

Observation e69c566f-2255-4adc-82e5-6b2a1379ae4b · outbound

This paper cites Masc: Massive arabic speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Masc: Massive arabic speech corpus,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.972932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.415432Z digest=sha256:7725b685de65164c4ae94e845c6084f4151f474e3c6e9c7bae9ac9c19c08e0ca

Observation 053ef7d6-8f61-4a79-bc8f-df7c3f39f83b · outbound

This paper cites Diacritic recognition perfor- mance in arabic asr,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Diacritic recognition perfor- mance in arabic asr,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.748774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.475402Z digest=sha256:57cfb9a3758bd3df3cf6ca52e2a0185bae9104fd7699eec08d4ae2c23761f666

Observation 9112c2e5-c618-47fe-ad9a-eb3823aa28cf · outbound

This paper cites Automatic restora- tion of diacritics for speech data sets,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Automatic restora- tion of diacritics for speech data sets,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.573106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.595759Z digest=sha256:1346846ccd341f38fc629f3fa8cf5e4e453dcaac0d28da73f10bc78b3efc426c

Observation dd85b093-b194-469c-afc5-51f7b77b07cf · outbound

This paper cites Clartts: An open-source classical arabic text-to-speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Clartts: An open-source classical arabic text-to-speech corpus,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.492215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.674924Z digest=sha256:6fabd272d1ae736984721965562357130468a7beeb5f9458ddf1f7706c14d2e3

Observation 53a911d2-71d0-4978-a8aa-4c5e4be5be7b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Robust speech recognition via large-scale weak supervision,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.089431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.682750Z digest=sha256:2a450f32e7d0f40dd9387ed1202cf1bc86813259ff3f7aa3a35c858a7f8734e6

Observation 9ef20df5-6677-456b-8174-27a1f6c87197 · outbound

This paper cites ArzEn: A speech corpus for code-switched Egyptian Arabic-English,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArzEn: A speech corpus for code-switched Egyptian Arabic-English,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.239238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.831943Z digest=sha256:b736aaa8dac03f1c31ad980bf2d5cc223469e21c13110fd5e824f4bfe811a2f9

Observation 1e8c6ac5-a457-4e26-8358-c331085d93d5 · outbound

This paper cites V oxblink: A large scale speaker verification dataset on camera,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis V oxblink: A large scale speaker verification dataset on camera,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.089268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:26.916277Z digest=sha256:592164050bc53802fa1035235b98ac4466e693afb3d0ba02b8d9a96af8098937

Observation 65af7c26-2278-48fa-9ce9-047a8ad23143 · outbound

This paper cites Modern standard arabic phonetics for speech synthesis,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Modern standard arabic phonetics for speech synthesis,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.900960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.019303Z digest=sha256:589be84665baf891af5e216da9973960f606a431b7244e8b356bdd7aa5e3c265

Observation f620d47a-d187-42f4-93e0-6e13081bcf3c · outbound

This paper cites Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.199944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.199944Z digest=sha256:be59cb7b0aa2551308437a22e532532aedef1a7eadb560cf819c95dd8011c1dc

Observation ffc5d9f1-5ed1-4090-880c-fae2b0018c2f · outbound

This paper cites Tashkeela: Novel corpus of arabic vo- calized texts, data for auto-diacritization systems,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Tashkeela: Novel corpus of arabic vo- calized texts, data for auto-diacritization systems,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.602628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.241330Z digest=sha256:3b75c36b4cc6ebaa059341cef383ad9bc2338a13e2b8984988db43dc464c72af

Observation 4f860300-c013-4d5e-93cc-f81f59662c3d · outbound

This paper cites We also fine-tuned KNN-VC [21], a non-parallel VC model that converts source into target speech by replacing each frame (a) VITS (w.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis We also fine-tuned KNN-VC [21], a non-parallel VC model that converts source into target speech by replacing each frame (a) VITS (w

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.690016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.989216Z digest=sha256:8fd8d2aaecfc4b0bb847c4a0bf540e1ab3f1c01c6b6abe47fcbfd6db51cec55b

Observation ba7c7e28-106d-4fb4-a4b0-37b977092f62 · outbound

This paper cites Comparison of topic identification methods for arabic language,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Comparison of topic identification methods for arabic language,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.443055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.415410Z digest=sha256:66994a33a0a3edbea0024c352bc0695a27ad064ff0fca960d743751585090bf7

Observation 595761de-0765-4661-924a-299b8c43c8dc · outbound

This paper cites STTATTS: Unified speech-to-text and text-to-speech model,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis STTATTS: Unified speech-to-text and text-to-speech model,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.288429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.583427Z digest=sha256:6617615f2f1a806a25398aa66dafb9c29ebd3b2daac79a3008d18117f320b8b3

Observation 25212192-0b16-4a2a-86c8-024f93293fcf · outbound

This paper cites ArTST: Arabic text and speech transformer,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArTST: Arabic text and speech transformer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.737572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.836847Z digest=sha256:3363d29008553d2c3ec73147d3c454ea764694595a558a7d5c2d14236eed7493

Observation 15679c2d-fd08-41d1-87b3-e00a189b3c50 · outbound

This paper cites Fast conformer with linearly scalable attention for efficient speech recognition,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Fast conformer with linearly scalable attention for efficient speech recognition,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.484852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.957211Z digest=sha256:b12eba3f95e283a0b97e766df53cba19ffb0eda7cfc509f1b86173320d702e41

Observation 68242d48-3888-4ae0-a34e-f2e53dc222d1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.076071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.076071Z digest=sha256:bb3c5ec4dd8f1f129dbc14b23b7fa3b6327b6bf0706d33a818a42a3819c13b83

Observation ca834ef0-2608-411c-8690-a878ddb3dd7e · outbound

This paper cites AAS-VC: On the Generalization Ability of Automatic Alignment Search based Non-autoregressive Sequence-to-sequence Voice Conversion.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis AAS-VC: On the Generalization Ability of Automatic Alignment Search based Non-autoregressive Sequence-to-sequence Voice Conversion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.372172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.372172Z digest=sha256:ccaa2a3a63db91c44cc2ee34b54e9d0bec3a15d9a22f59b246faa48a0ab2e63c

Observation 6b959bc7-994d-4b22-8339-aeb77a222c66 · outbound

This paper cites Parallel wavegan: A fast waveform generation model based on generative adversarial net- works with multi-resolution spectrogram,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Parallel wavegan: A fast waveform generation model based on generative adversarial net- works with multi-resolution spectrogram,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.248170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:28.521972Z digest=sha256:26c1cef367700461556fd3a8ef2d1d0a41c6d827934c05e4ce4b9504dc3ca6bc

Observation cf97cd27-ac38-4b85-983d-b3806904a469 · outbound

This paper cites V oice conversion with just nearest neighbors,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis V oice conversion with just nearest neighbors,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.075673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:28.634322Z digest=sha256:5d6db185451d89630def2cd6ba990fbcdf0ee374fb0c5daad5722d57d42cd402

Observation a26d98ca-5205-48f0-99ea-61164e471c45 · outbound

This paper cites Available: https://eprints.soton.ac.uk/409695/.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Available: https://eprints.soton.ac.uk/409695/

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.781759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:27.112518Z digest=sha256:aad4a5696744f8d06f073440d840d6b0d45adf8294e275901cceb953ac1226df

Pith citing papers

Observation b3b01c96-e5a6-431e-a440-48ac43f529a0 · inbound

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis cites this paper.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:00:28.939908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-15T06:32:42.880941+00:00.

source=pdf_text observed=2026-08-07T14:00:25.595644Z digest=sha256:28033076bc6bc0d63e09cb4f91e934f720856f00c623de27fa8c3ae9086f52f5