Pith. sign in

Paper Citation Record · LEDGER

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning

As of 17 August 2026, this Paper Citation Record lists 43 of 43 outbound references and 1 inbound Pith citation observation for arXiv:2411.17100.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.17100 v2

Coverage vector

measured 43 of 43 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T12:33:54.762787Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:40:57.467662Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:40:58.287735Z

Reference resolution

43 of 43 outbound references displayed

  • verified exact0
  • verified fuzzy38
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation d2b53f19-ec63-4db8-a1f7-28dca0f00e5d · outbound

This paper cites wav2vec: Unsu- pervised pre-training for speech recognition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning wav2vec: Unsu- pervised pre-training for speech recognition,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.142227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.651112Z digest=sha256:950b32a75a0b37d320c4cbfd66857350744b5790836252519e845e61094a3519

Observation fa3224e0-bee8-43ea-9bcc-cdd6c37bb8c7 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning wav2vec 2.0: A framework for self-supervised learning of speech representations,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T12:33:54.654886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:33:54.654886Z digest=sha256:1a2b8658b3819e055755b0136ed9f1d8d676c930eb5fae126bd4a41828297e5f

Observation d6927b8a-9ca8-4414-9036-895fcd163390 · outbound

This paper cites HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning HuBERT: Self-supervised speech representation learning by masked prediction of hidden units,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.129398Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.657731Z digest=sha256:4f1987b7ec0fab4fd73b36dc86932c75a34b61581b236079aede4512595f45fb

Observation 5cf27144-77a7-4c3d-a378-f0e4034163de · outbound

This paper cites WavLM: Large-scale self-supervised pre-training for full stack speech processing,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning WavLM: Large-scale self-supervised pre-training for full stack speech processing,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.121261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.660563Z digest=sha256:4e65eb7b6fc78cc89e4918d117bf3b1e903b91886be0bc5a7f3902df2d0abd8a

Observation 342d76dc-29fd-4fa4-b31a-9563cfd23959 · outbound

This paper cites data2vec: A general framework for self-supervised learning in speech, vision and language,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning data2vec: A general framework for self-supervised learning in speech, vision and language,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.113373Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.663414Z digest=sha256:4db45fea7f79aa4e7a81b3a23fc7c94608411eb1d56b8a82fed0c66b78c59db6

Observation bbc42e51-8097-4dea-ac66-afb6b425acb8 · outbound

This paper cites Efficient self-supervised learning with contextualized target rep- resentations for vision, speech and language,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Efficient self-supervised learning with contextualized target rep- resentations for vision, speech and language,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.104641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.666493Z digest=sha256:c91d9b459866f2d8d3a37ade7b7df675c74ae694c2768fe8df0916a78269d623

Observation 934b9a84-9ce5-4b69-bf4d-6d0673e91aa4 · outbound

This paper cites Self-supervised speech representation learning: A review,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Self-supervised speech representation learning: A review,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.095429Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.669461Z digest=sha256:15bd4e932393ac9ea37ae3dd53282569f427d5ff4a70b8a604f735de289a069e

Observation e16f9946-999a-48be-9277-a4c0d9b63e28 · outbound

This paper cites Recent advances in end-to-end automatic speech recog- nition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Recent advances in end-to-end automatic speech recog- nition,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.088394Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.672234Z digest=sha256:3741c98a74d26ec8bf01e197512bec5d37a0f51dab734e489ea9b323e1e880e9

Observation 42449c39-9fd3-47b8-b5ef-6e394a872dc5 · outbound

This paper cites Attention is all you need,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Attention is all you need,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.080286Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.674742Z digest=sha256:76b13c3da076d43986d8360482cf9cc96bcccf8774ae5e036b3a80e03c82b5be

Observation f5c5aef2-d7a6-4c18-b15d-965a43962906 · outbound

This paper cites SUPERB: Speech processing universal performance benchmark,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning SUPERB: Speech processing universal performance benchmark,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.071189Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.677322Z digest=sha256:51ec57cebc158b1284a152e618d3c8acde1f6e55ebdbee8cf51ad2d8da51304e

Observation 6e57c97f-e4e0-4c3d-a05e-d12d98a7b220 · outbound

This paper cites BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning BASE TTS: Lessons from building a billion-parameter Text-to-Speech model on 100K hours of data

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T12:33:54.679939Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:33:54.679939Z digest=sha256:1062a63abb91f1c77fd705d3f610fd865cf3b204679d0125c058b2b4aaf4dd58

Observation 200fc060-2c57-4bc7-90bb-4e4e8674e9e1 · outbound

This paper cites V oicebox: Text-guided multilingual universal speech generation at scale,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning V oicebox: Text-guided multilingual universal speech generation at scale,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.063525Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.683033Z digest=sha256:8d8a76e77705e7d64ff711b3a2f97e82aa191a466cd35e46663dc96c6ee16fad

Observation d3d2acce-03ea-413d-bbf5-9d535f30cb5a · outbound

This paper cites Scaling Speech Technology to 1,000+ Languages.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Scaling Speech Technology to 1,000+ Languages

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T12:33:54.685478Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:33:54.685478Z digest=sha256:e094f628a17a456f97a87542a6b2e3dec6c59c0481fae1ddfbf5698484c74f19

Observation f4d3ca29-61b5-404d-8622-40764ef3ad5b · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Robust speech recognition via large-scale weak supervision,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.055633Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.688434Z digest=sha256:dc6b0585762e336bcc5860e959f5f98f29bb0e56f79d9ab5a910a86f07feb141

Observation a394ca53-1cbc-4082-b787-9cc1d567b4e8 · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-12T12:33:54.691186Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:33:54.691186Z digest=sha256:1f87b6d2e23c78718c3597d9cdbf8f973553ced40d3b15852d5b339b4264d80f

Observation fc4884a6-8239-4a36-9873-54a4efad921a · outbound

This paper cites Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Libriheavy: a 50,000 hours ASR corpus with punctuation casing and context,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.046457Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.694125Z digest=sha256:ead0a22dd464cc3daaf7ec4be13ba1598faa04febafe527ab9b4ddf25dc0a392

Observation 11c877bf-f11c-4e6b-b1a0-208b70127424 · outbound

This paper cites GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning GigaSpeech 2: An Evolving, Large-Scale and Multi-domain ASR Corpus for Low-Resource Languages with Automated Crawling, Transcription and Refinement

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-12T12:33:54.696668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T12:33:54.696668Z digest=sha256:27942d3a5979558acb6c89a694cf2e3111107de6135a188d6aac9701b680669d

Observation 8f7593d9-29cf-424d-ac50-f2c6ea86443d · outbound

This paper cites LibriheavyMix: a 20,000-hour dataset for single-channel reverberant multi-talker speech separation, ASR and speaker diarization,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning LibriheavyMix: a 20,000-hour dataset for single-channel reverberant multi-talker speech separation, ASR and speaker diarization,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.038271Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.699600Z digest=sha256:aa52060436433fdfc5f60aa114444d865f9bc890983e0623e9d5d18045209228

Observation 05831b48-4977-4f7e-a54f-fdf954c257c9 · outbound

This paper cites fairseq: A fast, extensible toolkit for sequence modeling,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning fairseq: A fast, extensible toolkit for sequence modeling,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.030292Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.702237Z digest=sha256:5528ad8867aab2845c0d322da567f421e27eb0457d6dd7334123cadc5a1ffb3d

Observation 6ef83d78-220e-4244-9e37-d846ddb86a78 · outbound

This paper cites ESPnet: End-to-end speech processing toolkit,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning ESPnet: End-to-end speech processing toolkit,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.021899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.705107Z digest=sha256:f9e42b4cc95637cb505579f3ed26b08b43bf16a00dce79dfb722261b84c4faec

Observation 931d557b-3162-48e4-863a-9a4421f83536 · outbound

This paper cites The S3PRL toolkit: Self-supervised speech pre-training and representation learning,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning The S3PRL toolkit: Self-supervised speech pre-training and representation learning,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.012657Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.707852Z digest=sha256:146163fe94f8c74355422c4ef674c72c3999dabe0ea6ce51ff6da9bb29083b3c

Observation 3361a864-7dbe-4546-a945-8162c46032d6 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Zipformer: A faster and better encoder for automatic speech recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:55.003289Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.710599Z digest=sha256:fff9b0ae69468ed94dabac1f663e3f9a9be0f861c6a5301047d2dce914b096c0

Observation 0bbc40bd-6f9f-4d78-a734-65ac912e91a4 · outbound

This paper cites Librispeech: an ASR corpus based on public domain audio books,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Librispeech: an ASR corpus based on public domain audio books,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.993447Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.712871Z digest=sha256:4e85fc97456d80d5252d71112791d7c700302777df8f3a60d774ddb32671a44a

Observation ae3351bd-9d35-4871-b91b-19785096c599 · outbound

This paper cites Libri-Light: A benchmark for ASR with limited or no supervision,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Libri-Light: A benchmark for ASR with limited or no supervision,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.985413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.715003Z digest=sha256:2904f37f7d929a98bca4c59fd2ecf15d3f5635c0604afac0237215825d8c23a5

Observation 066939f3-c4e7-4978-9097-c1d62569031b · outbound

This paper cites Reducing barriers to self-supervised learning: HuBERT pre-training with academic compute,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Reducing barriers to self-supervised learning: HuBERT pre-training with academic compute,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.977110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.717196Z digest=sha256:62de26bdf17bf2e9dcf5bd3749b0dc02bfffecafb02134e7522c261de19e2d64

Observation de871f11-888e-4b79-bb61-ce9fb1727915 · outbound

This paper cites MelHuBERT: A simplified hubert on mel spectrograms,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning MelHuBERT: A simplified hubert on mel spectrograms,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.968529Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.719336Z digest=sha256:56f41dc69eca381179c331a44c64715453249ee1d2e38ea7d62fe4718d2a9907

Observation ac58aebd-d653-4297-bb09-a5d60b695483 · outbound

This paper cites Fast-HuBERT: an efficient training framework for self-supervised speech representation learning,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Fast-HuBERT: an efficient training framework for self-supervised speech representation learning,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.959364Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.721569Z digest=sha256:f0d87322391589f0659d44f1cc79fff2be8a8b77b2fb4608c96de2a9a4493248

Observation 778a5ea2-6a12-44ff-b266-be723a7fd002 · outbound

This paper cites Speech pre-training with acoustic piece,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Speech pre-training with acoustic piece,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.950358Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.723637Z digest=sha256:4a2bc749c740305a888af0b18ecadc4b2cd3a7db1f2778a57d8c101a167a81dc

Observation 788523a1-5fb4-4796-987c-2aed7caa300d · outbound

This paper cites Towards universal speech discrete tokens: A case study for ASR and TTS,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Towards universal speech discrete tokens: A case study for ASR and TTS,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.943061Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.725853Z digest=sha256:7b0116cda1721a1266cfe6ccb6a7190ef2eb0961a4c524b0ac81cce578311d42

Observation f5e7ccc9-bf0b-40aa-8d1f-d7839d12de21 · outbound

This paper cites Ctcbert: Advancing hidden-unit bert with ctc objectives,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Ctcbert: Advancing hidden-unit bert with ctc objectives,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.935114Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.728172Z digest=sha256:0e9d78d0c8eeb8eec76d7b5f27729b9fed4dc247d56c552675bd9732668a9f2e

Observation 97b844af-8cd5-4220-98b9-010226ab61f5 · outbound

This paper cites Supervision-guided codebooks for masked prediction in speech pre- training,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Supervision-guided codebooks for masked prediction in speech pre- training,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.926575Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.730965Z digest=sha256:e84f4a11b3cfb4a3c7bfff4c952ae8805eed1a1165e4177c3c23c9cd3dc2e792

Observation 03c16d2a-92f5-4d3f-bf21-2d0b33a4157f · outbound

This paper cites Pushing the limits of unsupervised unit discovery for SSL speech representation,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pushing the limits of unsupervised unit discovery for SSL speech representation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.917603Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.733613Z digest=sha256:249df3edfab6cc58b08316c1f7431d337a0c8b0b2aaf5a325ef641802663c4a5

Observation 4644650e-4918-4039-94ee-cad6ac975b35 · outbound

This paper cites Towards end-to-end unsupervised speech recognition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Towards end-to-end unsupervised speech recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.908185Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.736219Z digest=sha256:71976fccc09d239149f4a978edff1c14a20a104a1d6dc29131095d1e6b6969b1

Observation a1a9998c-da2d-482f-8dcc-d20f44003046 · outbound

This paper cites Connec- tionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Connec- tionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.899630Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.738757Z digest=sha256:31e6408566c2b3e6be79429d0930cc4f3f347c021102086074a678174652a7b2

Observation d77fda06-291b-4798-b641-d74b37498b85 · outbound

This paper cites Pruned RNN-T for fast, memory-efficient ASR training,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pruned RNN-T for fast, memory-efficient ASR training,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.892110Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.741363Z digest=sha256:20e96b6ad2788ff55fbb00d8ecd02c559627517473fb3d22c10a55949f4203d8

Observation ed823f41-8986-4f45-baf4-4625ebaa1632 · outbound

This paper cites Speech recognition with deep recurrent neural networks,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Speech recognition with deep recurrent neural networks,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.882442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.744050Z digest=sha256:94d48caf13591b154a40d8103fb22030738d3b468986450001cca832393ea7a0

Observation 2a3c8865-9f80-4558-8620-efdc63e03de7 · outbound

This paper cites Lhotse: a speech data representation library for the modern deep learning ecosystem,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Lhotse: a speech data representation library for the modern deep learning ecosystem,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.873325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.746733Z digest=sha256:e03cbc62850f7086e0aba721615ac0c9326339baf8a2257fa49f1ac576e06364

Observation 13b462f5-4e08-4526-a83a-6d92aad1d630 · outbound

This paper cites Neural machine translation of rare words with subword units,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Neural machine translation of rare words with subword units,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.862680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.749446Z digest=sha256:9376cc3116caa97be8854a9d06e3b276df24b6642a0f459a11446e99698f198d

Observation a31fc833-9d95-4a7a-a240-700a2f22533a · outbound

This paper cites Fast and parallel decoding for transducer,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Fast and parallel decoding for transducer,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.851992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.752100Z digest=sha256:7fe2ba1e2911e3fc94d6eebb088a6223ba1336d5f9f70bd03dd919a9d1ed6d3c

Observation 0e6a96dc-b2c0-47c3-9240-eaba7d53421a · outbound

This paper cites Wav2letter++: A fast open-source speech recognition system,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Wav2letter++: A fast open-source speech recognition system,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.843290Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.754822Z digest=sha256:e5aca7abcd29c30df8cfdafea17f5274262ba62a95c876f7b87f3599dfdc669b

Observation 72ecdffa-ed46-47df-b80e-254a9e2e5206 · outbound

This paper cites Transformer transducer: A streamable speech recognition model with transformer encoders and RNN-T loss,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Transformer transducer: A streamable speech recognition model with transformer encoders and RNN-T loss,

Reference 41

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.834581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.757441Z digest=sha256:13b430d18d06259e9912fe0cd2cc3c227a413180bdf1800953138302c953893d

Observation 59003f19-c582-4cec-86fe-11ec091d15da · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Conformer: Convolution-augmented transformer for speech recognition,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.826678Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.760183Z digest=sha256:3165a8cfbbd8425c1095fad3701e0d00c2424225cc3f0c1bfaa75841edbfd2d8

Observation 84c75abf-cbd5-4a4b-accf-44af8ff8f274 · outbound

This paper cites Pushing the limits of semi- supervised learning for automatic speech recognition,.

k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning Pushing the limits of semi- supervised learning for automatic speech recognition,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T12:33:54.818178Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-12T12:33:54.762787Z digest=sha256:20ada2f873d8062d3cf04e7a95492558a80ec68c3d3792c8aa0a967b97e7ee24

Pith citing papers

Observation 2ff276ab-f86e-4d6c-8636-b72a67c6d58f · inbound

VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining cites this paper.

VietASR: Achieving Industry-level Vietnamese ASR with 50-hour labeled data and Large-Scale Speech Pretraining k2SSL: A Faster and Better Framework for Self-Supervised Speech Representation Learning

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:40:58.375212Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T14:40:57.467662Z digest=sha256:14a01b33973488e0de51cb2e233ce814ea60e57c02337b28c8044961d70f4ff1