Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:33:54.762787Z
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2411.17100.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T12:33:54.762787Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:40:57.467662Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:40:58.287735Z
43 of 43 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d2b53f19-ec63-4db8-a1f7-28dca0f00e5d · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning wav2vec: Unsu- pervised pre-training for speech recognition,
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation fa3224e0-bee8-43ea-9bcc-cdd6c37bb8c7 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning wav2vec 2.0: A framework for self-supervised learning of speech representations,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6927b8a-9ca8-4414-9036-895fcd163390 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 5cf27144-77a7-4c3d-a378-f0e4034163de · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning WavLM: Large-scale self-supervised pre-training for full stack speech processing,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 342d76dc-29fd-4fa4-b31a-9563cfd23959 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning data2vec: A general framework for self-supervised learning in speech, vision and language,
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation bbc42e51-8097-4dea-ac66-afb6b425acb8 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Efficient self-supervised learning with contextualized target rep- resentations for vision, speech and language,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 934b9a84-9ce5-4b69-bf4d-6d0673e91aa4 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Self-supervised speech representation learning: A review,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation e16f9946-999a-48be-9277-a4c0d9b63e28 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Recent advances in end-to-end automatic speech recog- nition,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 42449c39-9fd3-47b8-b5ef-6e394a872dc5 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Attention is all you need,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f5c5aef2-d7a6-4c18-b15d-965a43962906 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning SUPERB: Speech processing universal performance benchmark,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6e57c97f-e4e0-4c3d-a05e-d12d98a7b220 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 200fc060-2c57-4bc7-90bb-4e4e8674e9e1 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning V oicebox: Text-guided multilingual universal speech generation at scale,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d3d2acce-03ea-413d-bbf5-9d535f30cb5a · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Scaling Speech Technology to 1,000+ Languages
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f4d3ca29-61b5-404d-8622-40764ef3ad5b · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Robust speech recognition via large-scale weak supervision,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a394ca53-1cbc-4082-b787-9cc1d567b4e8 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc4884a6-8239-4a36-9873-54a4efad921a · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 11c877bf-f11c-4e6b-b1a0-208b70127424 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8f7593d9-29cf-424d-ac50-f2c6ea86443d · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning LibriheavyMix: a 20,000-hour dataset for single-channel reverberant multi-talker speech separation, ASR and speaker diarization,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 05831b48-4977-4f7e-a54f-fdf954c257c9 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning fairseq: A fast, extensible toolkit for sequence modeling,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6ef83d78-220e-4244-9e37-d846ddb86a78 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning ESPnet: End-to-end speech processing toolkit,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 931d557b-3162-48e4-863a-9a4421f83536 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning The S3PRL toolkit: Self-supervised speech pre-training and representation learning,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 3361a864-7dbe-4546-a945-8162c46032d6 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Zipformer: A faster and better encoder for automatic speech recognition,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0bbc40bd-6f9f-4d78-a734-65ac912e91a4 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Librispeech: an ASR corpus based on public domain audio books,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ae3351bd-9d35-4871-b91b-19785096c599 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Libri-Light: A benchmark for ASR with limited or no supervision,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 066939f3-c4e7-4978-9097-c1d62569031b · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Reducing barriers to self-supervised learning: HuBERT pre-training with academic compute,
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation de871f11-888e-4b79-bb61-ce9fb1727915 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning MelHuBERT: A simplified hubert on mel spectrograms,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ac58aebd-d653-4297-bb09-a5d60b695483 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Fast-HuBERT: an efficient training framework for self-supervised speech representation learning,
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 778a5ea2-6a12-44ff-b266-be723a7fd002 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Speech pre-training with acoustic piece,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 788523a1-5fb4-4796-987c-2aed7caa300d · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Towards universal speech discrete tokens: A case study for ASR and TTS,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation f5e7ccc9-bf0b-40aa-8d1f-d7839d12de21 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Ctcbert: Advancing hidden-unit bert with ctc objectives,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 97b844af-8cd5-4220-98b9-010226ab61f5 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Supervision-guided codebooks for masked prediction in speech pre- training,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 03c16d2a-92f5-4d3f-bf21-2d0b33a4157f · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pushing the limits of unsupervised unit discovery for SSL speech representation,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4644650e-4918-4039-94ee-cad6ac975b35 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Towards end-to-end unsupervised speech recognition,
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a1a9998c-da2d-482f-8dcc-d20f44003046 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Connec- tionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d77fda06-291b-4798-b641-d74b37498b85 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pruned RNN-T for fast, memory-efficient ASR training,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation ed823f41-8986-4f45-baf4-4625ebaa1632 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Speech recognition with deep recurrent neural networks,
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2a3c8865-9f80-4558-8620-efdc63e03de7 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Lhotse: a speech data representation library for the modern deep learning ecosystem,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 13b462f5-4e08-4526-a83a-6d92aad1d630 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Neural machine translation of rare words with subword units,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation a31fc833-9d95-4a7a-a240-700a2f22533a · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Fast and parallel decoding for transducer,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 0e6a96dc-b2c0-47c3-9240-eaba7d53421a · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Wav2letter++: A fast open-source speech recognition system,
Reference 40
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 72ecdffa-ed46-47df-b80e-254a9e2e5206 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Transformer transducer: A streamable speech recognition model with transformer encoders and RNN-T loss,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 59003f19-c582-4cec-86fe-11ec091d15da · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Conformer: Convolution-augmented transformer for speech recognition,
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 84c75abf-cbd5-4a4b-accf-44af8ff8f274 · outbound
k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pushing the limits of semi- supervised learning for automatic speech recognition,
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2ff276ab-f86e-4d6c-8636-b72a67c6d58f · inbound
VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.