Pith. sign in

Paper Citation Record · LEDGER

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

As of 19 August 2026, this Paper Citation Record lists 29 of 29 outbound references and 1 inbound Pith citation observation for arXiv:2505.20506.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.20506 v1

Coverage vector

measured 29 of 29 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:28.634322Z

measured 30 of 30 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:00:25.595644Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:00:28.778797Z

Reference resolution

29 of 29 outbound references displayed

  • verified exact0
  • verified fuzzy22
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d9f8115b-88fa-4f4b-9ba0-7b5d2b52230e · outbound

This paper cites an unresolved cited work.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Unresolved cited work

Reference 1

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:33.316583Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.491796Z digest=sha256:4fa886de7b0c5dc8b544409093ac971df43b7a8a586a30ad594c34b21315ab3e

Observation b3b01c96-e5a6-431e-a440-48ac43f529a0 · outbound

This paper cites ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:00:28.939908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.595644Z digest=sha256:41620081dc6e5f76d819025484dd46853f25584d70c26dc0bd5d23264ff4e524

Observation dfb1b0c9-46c1-48b4-8a93-5b50995bb039 · outbound

This paper cites In this section, we describe each part of ArV oice and provide justifica- tion for design decisions where applicable.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis In this section, we describe each part of ArV oice and provide justifica- tion for design decisions where applicable

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:33.048002Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.719894Z digest=sha256:ab23d2fcc686d1fc47a07819403f111955d6dc607973395721200062edbadc6d

Observation 53e56dec-cefd-45d0-a40e-a8f05d00d8b3 · outbound

This paper cites an unresolved cited work.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:00:32.857585Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.864240Z digest=sha256:8e9932a9b41364b5fab6e05b5dcbcf924978e313acc6d2c55035a2d3721008fe

Observation 504024d5-b896-4714-bf34-0f8b1f7a0a16 · outbound

This paper cites The dataset consists of 11 voices in total, 7 of which are human voices, and 4 are syn- thetic with parallel text.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis The dataset consists of 11 voices in total, 7 of which are human voices, and 4 are syn- thetic with parallel text

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.522013Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.057794Z digest=sha256:d3f98baadf7a7990d6f05d6766a9b54c2278656794aa44756ea6986fb7271cbe

Observation 027b4f30-a22f-4281-9271-061b715c7bcc · outbound

This paper cites This work was partially funded by a Google research award (11/2023).

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis This work was partially funded by a Google research award (11/2023)

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.321259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.148273Z digest=sha256:7e5ea426c2363f6f2c95fb92ff1e53b67d33d9c549773991a72ccbe0180de8e2

Observation ab7b15dd-9064-48a4-b593-a83a87fad8c9 · outbound

This paper cites Cmu wilderness multilingual speech dataset,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Cmu wilderness multilingual speech dataset,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.404796Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.742493Z digest=sha256:ed9fead8fc2b66e7d7a265fead9e86ac2f87dfd531034e3f733a39b3bbec004d

Observation 86a121d7-422a-40da-86a4-9c12d8eed7b6 · outbound

This paper cites Speech recognition challenge in the wild: Arabic mgb-3,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Speech recognition challenge in the wild: Arabic mgb-3,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:26.224154Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:26.224154Z digest=sha256:3ef5f470d3d70a623c041d656b6d537b6492da15f0813e8c79b8e66adf030d4f

Observation 5ee82de3-d52f-495f-a59d-c9f82452c7e5 · outbound

This paper cites QASR: QCRI aljazeera speech resource a large scale annotated Arabic speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis QASR: QCRI aljazeera speech resource a large scale annotated Arabic speech corpus,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.148453Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.295851Z digest=sha256:f28391ce7506cc5c0db87c0ae2c8a43c7696b6d1e00c68010d246d1303e0c826

Observation e69c566f-2255-4adc-82e5-6b2a1379ae4b · outbound

This paper cites Masc: Massive arabic speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Masc: Massive arabic speech corpus,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.972932Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.415432Z digest=sha256:c6b166396d708b0557aaac76c9113d6c4c58a876feee41618d0657120c39023c

Observation 053ef7d6-8f61-4a79-bc8f-df7c3f39f83b · outbound

This paper cites Diacritic recognition perfor- mance in arabic asr,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Diacritic recognition perfor- mance in arabic asr,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.748774Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.475402Z digest=sha256:0f1375a8b1c38b91fade84c67be26b670141d99c0120e18f6e49ba5019712ce9

Observation 9112c2e5-c618-47fe-ad9a-eb3823aa28cf · outbound

This paper cites Automatic restora- tion of diacritics for speech data sets,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Automatic restora- tion of diacritics for speech data sets,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.573106Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.595759Z digest=sha256:9f3b83205315c9b39527d26134f33040999978b2a8f46cfe0b18dad51f7391fe

Observation dd85b093-b194-469c-afc5-51f7b77b07cf · outbound

This paper cites Clartts: An open-source classical arabic text-to-speech corpus,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Clartts: An open-source classical arabic text-to-speech corpus,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.492215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.674924Z digest=sha256:557282649bba134ec2f49b3025042277f9ccc47f01b53bc699622bcd58216732

Observation 53a911d2-71d0-4978-a8aa-4c5e4be5be7b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Robust speech recognition via large-scale weak supervision,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.089431Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.682750Z digest=sha256:8fec5018f50ac52dd400b4186b391152c7c2bd2f57d0aa05a1ed78b07b306291

Observation 9ef20df5-6677-456b-8174-27a1f6c87197 · outbound

This paper cites ArzEn: A speech corpus for code-switched Egyptian Arabic-English,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArzEn: A speech corpus for code-switched Egyptian Arabic-English,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.239238Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.831943Z digest=sha256:dbec7cad2b13587100695567ac56edfbc07beeb81fe622268fb3465ffb942e87

Observation 1e8c6ac5-a457-4e26-8358-c331085d93d5 · outbound

This paper cites V oxblink: A large scale speaker verification dataset on camera,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis V oxblink: A large scale speaker verification dataset on camera,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:31.089268Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:26.916277Z digest=sha256:4f5923a89e85af7b86c7f8095cafcc9e5d069cdbd28b5426aa40e930552d0c12

Observation 65af7c26-2278-48fa-9ce9-047a8ad23143 · outbound

This paper cites Modern standard arabic phonetics for speech synthesis,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Modern standard arabic phonetics for speech synthesis,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.900960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.019303Z digest=sha256:6bcef804b3bc56c9c27de139dd068ada070b2d4ef19ea5c0b82c14d1337b9cb3

Observation f620d47a-d187-42f4-93e0-6e13081bcf3c · outbound

This paper cites Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Fish-Speech: Leveraging Large Language Models for Advanced Multilingual Text-to-Speech Synthesis

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.199944Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.199944Z digest=sha256:d051f89cab228d7db967180eb344b31f5e37919afdae8547a1212a53a3922a84

Observation ffc5d9f1-5ed1-4090-880c-fae2b0018c2f · outbound

This paper cites Tashkeela: Novel corpus of arabic vo- calized texts, data for auto-diacritization systems,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Tashkeela: Novel corpus of arabic vo- calized texts, data for auto-diacritization systems,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.602628Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.241330Z digest=sha256:ecac9f77af039b06456494711cd10275e34bd533c9b56245a87a63b7b3c970c8

Observation 4f860300-c013-4d5e-93cc-f81f59662c3d · outbound

This paper cites We also fine-tuned KNN-VC [21], a non-parallel VC model that converts source into target speech by replacing each frame (a) VITS (w.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis We also fine-tuned KNN-VC [21], a non-parallel VC model that converts source into target speech by replacing each frame (a) VITS (w

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:32.690016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.989216Z digest=sha256:03a25e4bdba2652b42313a5614cf2c5d15355f06f0dfb394ecc42efc922bbf02

Observation ba7c7e28-106d-4fb4-a4b0-37b977092f62 · outbound

This paper cites Comparison of topic identification methods for arabic language,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Comparison of topic identification methods for arabic language,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.443055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.415410Z digest=sha256:b903689481eb5b7106c7e29f3349c5e6f05f040419f9b23c90f4a611e421813d

Observation 595761de-0765-4661-924a-299b8c43c8dc · outbound

This paper cites STTATTS: Unified speech-to-text and text-to-speech model,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis STTATTS: Unified speech-to-text and text-to-speech model,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.288429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.583427Z digest=sha256:2488ee00e8e3071ff422076fcb04e193887dab03b565d3bbe4b19b1c168d07e2

Observation 25212192-0b16-4a2a-86c8-024f93293fcf · outbound

This paper cites ArTST: Arabic text and speech transformer,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArTST: Arabic text and speech transformer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.737572Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.836847Z digest=sha256:820e233361f4b3fd000f46096fbec6a95973093a5fa0e70bf58f6832997cc6fe

Observation 15679c2d-fd08-41d1-87b3-e00a189b3c50 · outbound

This paper cites Fast conformer with linearly scalable attention for efficient speech recognition,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Fast conformer with linearly scalable attention for efficient speech recognition,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.484852Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.957211Z digest=sha256:af920e50ecb475e33b856f11316fcc6a057df0e2dad8de209f326aa2dd27aa1c

Observation 68242d48-3888-4ae0-a34e-f2e53dc222d1 · outbound

This paper cites Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Conditional variational autoencoder with adversarial learning for end-to-end text-to-speech,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.076071Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.076071Z digest=sha256:cc391dc64979744e561eff8b6ce60ebdf46f0efe4a0bd10206d7669aca6a0780

Observation ca834ef0-2608-411c-8690-a878ddb3dd7e · outbound

This paper cites AAS-VC: On the Generalization Ability of Automatic Alignment Search based Non-autoregressive Sequence-to-sequence Voice Conversion.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis AAS-VC: On the Generalization Ability of Automatic Alignment Search based Non-autoregressive Sequence-to-sequence Voice Conversion

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T14:00:28.372172Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:00:28.372172Z digest=sha256:76fae130f028917acf3b217acc25c594ac033f1f4e77525e1d648c2e1c1681db

Observation 6b959bc7-994d-4b22-8339-aeb77a222c66 · outbound

This paper cites Parallel wavegan: A fast waveform generation model based on generative adversarial net- works with multi-resolution spectrogram,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Parallel wavegan: A fast waveform generation model based on generative adversarial net- works with multi-resolution spectrogram,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.248170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:28.521972Z digest=sha256:60c1ece43c26a94e9725a4f8befbd402f3aa75104c672fc9729a0146700e2928

Observation cf97cd27-ac38-4b85-983d-b3806904a469 · outbound

This paper cites V oice conversion with just nearest neighbors,.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis V oice conversion with just nearest neighbors,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:29.075673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:28.634322Z digest=sha256:7f8ba29fa1847a2ad503f66c2923c8785fc0cff29ebfd1cfe74fdf3efe834214

Observation a26d98ca-5205-48f0-99ea-61164e471c45 · outbound

This paper cites Available: https://eprints.soton.ac.uk/409695/.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis Available: https://eprints.soton.ac.uk/409695/

Reference 2016

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:00:30.781759Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:27.112518Z digest=sha256:03205912fa0f1cd6f79d3a91a5a8beea9fe835f0458bbb0408b387ab37d91927

Pith citing papers

Observation b3b01c96-e5a6-431e-a440-48ac43f529a0 · inbound

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis cites this paper.

ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:00:28.939908Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T14:00:25.595644Z digest=sha256:41620081dc6e5f76d819025484dd46853f25584d70c26dc0bd5d23264ff4e524