Pith. sign in

Paper Citation Record · LEDGER

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

As of 19 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2505.21578.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.21578 v1

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:21.783339Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T13:49:18.926273Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T13:49:22.026276Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy11
  • unresolved10
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 86228da0-13b9-456f-bba9-b3845192c19b · outbound

This paper cites open-source.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use open-source

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.394179Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:18.821110Z digest=sha256:88272e5c7ff17130fda0b70c2d5ee537a26a02fe4e007b9bb1c25106f4273c53

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · outbound

This paper cites Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:287e323bc2ce8ac64d1a72894dab814764551bb7ba4b4b16701f6227166524fe

Observation 3cb75ae8-7806-45ed-8591-b197b1d11987 · outbound

This paper cites dev-other.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use dev-other

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.889982Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.100288Z digest=sha256:8ca6c6fab431cf7bc905805d0742977326d864104fda567f74aa9d41004fa1f0

Observation 2f61fc9c-bdac-46c9-bef6-fb1aa54c9974 · outbound

This paper cites an unresolved cited work.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Unresolved cited work

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T13:49:23.743847Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.188453Z digest=sha256:b68af500353bb013418bbddc4f52984c4a8c0ea2ea8927b792b5eec660fc1226

Observation 6921c4be-1ce9-4e8e-8e98-955461dd57aa · outbound

This paper cites The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The dataset is easy to reproduce thanks to the SpeechBrain re- leased source code and can be loaded in a single line of code

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.596584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.273591Z digest=sha256:7d8905c451fedf9c397bf67b4dfa2d3cd8b618bfdf0ee8d4174e07d0e146159d

Observation 7bb05ae4-89c5-4403-a3fd-02391cb282d9 · outbound

This paper cites Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Mosel: 950,000 hours of speech data for open-source speech foundation model train- ing on eu languages,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.936707Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:20.155009Z digest=sha256:5c45693b65ddce205eb1c1350bd728c3c55911b286214f656c77f8ac3dc53d36

Observation 3c5899e0-7220-43a1-9edb-d87e6474afe9 · outbound

This paper cites The design for the wall street journal- based csr corpus,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The design for the wall street journal- based csr corpus,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.282469Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.415192Z digest=sha256:df0941f6c42d5fde3982b359c7a0f18e68689d95852a6419a158e94302fa9cb4

Observation 21e858af-ea56-477b-b161-b649dc65e523 · outbound

This paper cites Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Darpa timit acoustic-phonetic continous speech corpus cd-rom. nist speech disc 1-1.1,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:23.109696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.540643Z digest=sha256:d94d2636f56c91da87d0179bc5a91f711a2d60261f11ea4805e52013b4f1c5fb

Observation 2833979d-5341-4859-b931-91574bd846d9 · outbound

This paper cites Lib- rispeech: an asr corpus based on public domain audio books,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Lib- rispeech: an asr corpus based on public domain audio books,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.722000Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.722000Z digest=sha256:71eb54dbbf64f71c94c77f28f8b631cecf16769d44624d335f5aa125375785e8

Observation e3b28d05-9402-41ee-a114-f234b230419b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Robust speech recognition via large-scale weak supervision,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:19.868605Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:19.868605Z digest=sha256:3b20e070aa08db163e6748ec5cc79643d7308f8b0ef88136111cf92b9bcd2075

Observation 27e30a3f-6914-4aef-b1de-f9656e08d8ce · outbound

This paper cites Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Conformer-1: Robust ASR via Large-Scale Semisupervised Bootstrapping

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.016438Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.016438Z digest=sha256:e39e40c34d9a6a086728f0b6a9cd9c6b574b6f83b311023a04573919bdb1e59a

Observation 48f8d9f6-776d-4b89-85bd-820eb4c85620 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.093321Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.093321Z digest=sha256:29b7406d3dd8943a8c52117847fb52f08be400d2d4f55591d42706c8d079bb63

Observation 06f76e27-7e8b-4b08-bc88-fe1b898c0c92 · outbound

This paper cites Yodas: Youtube-oriented dataset for audio and speech,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Yodas: Youtube-oriented dataset for audio and speech,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.351770Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.351770Z digest=sha256:0ec6cb3ea212686aad80caa8d08b8aa61af3b30b2132c465ef0f979c5c5871ae

Observation f10eea87-3652-4171-bb8e-2ae55c4a4af5 · outbound

This paper cites Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Libriheavy: a 50,000 hours asr corpus with punc- tuation casing and context,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.780955Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:20.544196Z digest=sha256:b01847bc226b3d6b90bed24dcd7d2740cf9d00568ba404c179e5e61049db4a54

Observation 287cf320-8e23-46a3-8c7c-cad2c419ef2f · outbound

This paper cites People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use People’s Speech [9] and YODAS [7] are the most recent attempts at overcoming the read versus spontaneous speech issue at large scale

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:24.045529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:19.029598Z digest=sha256:bf97400cc1dabd7d093ce39adbee6cd1d2ace7326b0874513aab64c95bbf46f0

Observation d94107cd-9278-4f9c-abfd-72870ccb72d7 · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.652179Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.652179Z digest=sha256:f9455181e53ecd6062559412f750272ea3a50333ebf5356dbad8af58bec11cb3

Observation fc3ba030-0bf9-4472-86b6-df1c827b9927 · outbound

This paper cites SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use SpeechStew: Simply Mix All Available Speech Recognition Data to Train One Large Neural Network

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.806789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.806789Z digest=sha256:d529f54c7bbd49a795f98006b3fa8a9b8c25fc3fae12789a7cd825b36e6a9cab

Observation 4794e4e2-7e1b-4afe-96ca-bf96996520d8 · outbound

This paper cites Reproducing whisper-style training using an open-source toolkit and publicly available data,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Reproducing whisper-style training using an open-source toolkit and publicly available data,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:20.960789Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:20.960789Z digest=sha256:1f9320681c6bbf34fd9b365b21695da34e68ef0c4d139825077a92f50a7dfda6

Observation 9dc565a7-c1be-46a6-9342-50d71a9f83fe · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Common Voice: A Massively-Multilingual Speech Corpus

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.198546Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.198546Z digest=sha256:a726bbfab9cbb831cc34d5ef53fcf5cf4b20d6e49bb4ceb27514089ae99126f3

Observation 39e6ca45-5e08-4ef6-acb3-1f1219484f0a · outbound

This paper cites Open-source conversational ai with speechbrain 1.0,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Open-source conversational ai with speechbrain 1.0,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.623249Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:21.352107Z digest=sha256:8d0552ce70d1ff9d232e995e3e452c0313519a202b60838dd9f50c440eb26d0f

Observation 80bb602e-4d01-4b5d-82c3-8a56169d6738 · outbound

This paper cites MLS: A Large-Scale Multilingual Dataset for Speech Research.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use MLS: A Large-Scale Multilingual Dataset for Speech Research

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T13:49:21.471158Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T13:49:21.471158Z digest=sha256:e31122286c9b6d55a5ae1ad8f196ef57804177333c1db2df7ecd705934006b0a

Observation 44b33b1d-a72a-4c23-94b9-0f3acd44b9ea · outbound

This paper cites V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use V oxpopuli: A large-scale multilingual speech corpus for representation learn- ing, semi-supervised learning and interpretation,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.464684Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:21.645978Z digest=sha256:3a2bd564296306df969dcad254ca6dce026b25dbe16164552351aa4360cecdec

Observation a99041cf-c46f-4689-bd42-82de3fd80fbb · outbound

This paper cites Joint ctc-attention based end-to-end speech recognition using multi-task learning,.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Joint ctc-attention based end-to-end speech recognition using multi-task learning,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T13:49:22.319240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:21.783339Z digest=sha256:efb37c01cadf57e1f6e56718646d8a4b7f021412ef86a55598830b43752ba288

Pith citing papers

Observation 2232aa19-443b-418f-9e68-3cfda600b704 · inbound

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use cites this paper.

Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use Loquacious Set: 25,000 Hours of Transcribed and Diverse English Speech Recognition Data for Research and Commercial Use

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-08-07T13:49:22.160440Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.

source=pdf_text observed=2026-08-07T13:49:18.926273Z digest=sha256:287e323bc2ce8ac64d1a72894dab814764551bb7ba4b4b16701f6227166524fe