Pith. sign in

Paper Citation Record · LEDGER

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

As of 19 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2505.19774.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19774 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:49.357205Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:46.355892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:28:39.508010Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved21
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f25726a-38b9-46a3-aa95-c855e7e20be6 · outbound

This paper cites Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:53.104993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.192370Z digest=sha256:2084250a640e4cb9d9c0152ba3ded853145fda90a3867bfa0c3de3b535c92637

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · outbound

This paper cites DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:e8b44313237fbb5fba4580de7a0bcb1185b06cf9d36021d3f6e9f95c3c1f603a

Observation a6ad7272-62b5-4be4-9a9a-02d1a43c21e6 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.685483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.467017Z digest=sha256:dc65be786ee823aa34a788970b8230540e6117d76fb4505f4f43984be2a8f9ac

Observation fa9dcbb8-7061-43e1-b826-bdaec77172bb · outbound

This paper cites Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:52.286979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.646277Z digest=sha256:f77c2591866fcc9a27ef0d5ee513797e2f6e6bc566bab6a99e61a529bde5cbd2

Observation 46293429-1243-40a3-9b44-805d29587b3a · outbound

This paper cites Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.496675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.552005Z digest=sha256:58d25568cb69be2fb7f646b3b291bb0b2fa759464aa29b00328fe9e1d094fa03

Observation 2f177c02-655e-4476-ab67-569e5eed50a5 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.274896Z digest=sha256:813e4bf94b0e835f7cbb51d682da0824f742ba889248d7d5135a2f8a25f841a9

Observation c65c6e3a-f562-467d-81ce-1e2a83a8e978 · outbound

This paper cites Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4).

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.127133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.744428Z digest=sha256:dd613bda1a6b1e48f167752a4352d533042fa06dc9767562229ece35e1995783

Observation ff259256-d9db-4d11-929e-1baa37b58f5e · outbound

This paper cites In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode

Reference 8

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:51.897019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.835187Z digest=sha256:7e8d500c0080bdaff5ae8ad0f3011d062fc934dc60e612edbd5754877b5c716d

Observation 3df5a67f-bfa2-4c30-b26b-c66d3e4dc616 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:51.660143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:46.895063Z digest=sha256:8fdee7f8d228e11fb17cde1d233869a3d4f2c4555f3e3a966c6ac11980fa6c50

Observation c9700196-b019-4ea0-b376-8035a5b3ad6b · outbound

This paper cites Google usm: Scaling automatic speech recognition beyond 100 languages,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google usm: Scaling automatic speech recognition beyond 100 languages,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.963664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.963664Z digest=sha256:abd9061e26691c16818a9c8c578fc7b93a431083bb8ccce9d33b79a913a0d514

Observation ec2afa29-f28f-4268-8c15-77ae428a0d6f · outbound

This paper cites Self-supervised speech representation learning: A review,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised speech representation learning: A review,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.108528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.108528Z digest=sha256:d5d3b58500466f4e5dc045ddb40f64d36ffa6d7b49c163c4aea9f821ea9b30d9

Observation ac4fef8f-9040-40f2-8b58-b96e78ba18bf · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Robust Speech Recognition via Large-Scale Weak Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.158251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.158251Z digest=sha256:ee2f7ede2f0b11a680eb88dff2ceaf72bfe6e9f4f4c8a9079b30b574215a3d87

Observation 7b8ab252-ca3f-4248-ad46-fc0a3ba2ac4a · outbound

This paper cites Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.243261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.243261Z digest=sha256:c1bb055e150e2cf8321a0e4192148958ac1e4233ae2bdaffb6ae3a009e110f1c

Observation 81e526fa-4163-487f-9b74-bf211c4000bb · outbound

This paper cites Conformer with dual-mode chunked attention for joint online and offline ASR.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer with dual-mode chunked attention for joint online and offline ASR

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:50.153372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.355685Z digest=sha256:22423060b595839726b84b2314df2f0fd32e408c713f2285d4d89bbbb2dda1c0

Observation 4aebe4a4-8b17-490a-ae73-0d5d352b39a3 · outbound

This paper cites Multi- mode transformer transducer with stochastic future context,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi- mode transformer transducer with stochastic future context,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.451479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.466237Z digest=sha256:7f2789349fd7fd0f2540719120da69f75b14c87674e57058d294d0e40805e996

Observation 5b58ca95-c45e-4e83-8170-4a4517a1083b · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.072372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:48.584229Z digest=sha256:6942d18933a35adf83d4ce167509d10fb53da2ee3919dbed201897062740a04d

Observation 9acf3d35-c4ac-457a-8387-83306a72b0c6 · outbound

This paper cites Variable Attention Masking for Configurable Transformer Transducer Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Variable Attention Masking for Configurable Transformer Transducer Speech Recognition

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.920805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.679297Z digest=sha256:cc19739d0462ef14398b73601d0468070791fedab63c675e43b071a5bee03e00

Observation 862a48ce-cd21-4cf8-9386-2c06d30700c2 · outbound

This paper cites Sequence transduction with recurrent neural networks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Sequence transduction with recurrent neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.319686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.760230Z digest=sha256:ebca1cd8be902132aa0274cc77438bb1b13571a59587550559ca2687b0d66adf

Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · outbound

This paper cites SUPERB: Speech processing Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.829556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.829556Z digest=sha256:10d9efaf878b9a4b8f2f884a7be5eb432d0b60cad8bc10e2750d292c7cbd3171

Observation 534ac630-f760-4521-9d3a-06d5baefddae · outbound

This paper cites Self-supervised Learning with Random-projection Quantizer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised Learning with Random-projection Quantizer for Speech Recognition

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.761320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.911681Z digest=sha256:b511d7764439256217f05d2d699a3357db1c2a769ae242f182697437e991c154

Observation 157c9368-a37c-426c-ae3f-e7a5d50f0181 · outbound

This paper cites Universal-1: Robust and accurate multi- lingual speech-to-text,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Universal-1: Robust and accurate multi- lingual speech-to-text,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.197138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:48.032670Z digest=sha256:66d9c7bad9a92442662b028d87e68611f789349f6a6194904e53c7d31505affa

Observation b7da4468-3076-44dc-a9eb-cc06b1140fcc · outbound

This paper cites ML-SUPERB: Multilingual Speech Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation ML-SUPERB: Multilingual Speech Universal PERformance Benchmark

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.072445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.072445Z digest=sha256:dac4206d76d9f8d88ec9e75d520ce7a25bfc2c9b266e40792defee3694587990

Observation a09b4471-d3be-4e13-9274-7b338103663e · outbound

This paper cites Lib- rispeech: An asr corpus based on public domain audio books,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Lib- rispeech: An asr corpus based on public domain audio books,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.177512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.177512Z digest=sha256:8390444744bb5afe8b730dc20e778624ce4a93bee94321019153ef7a6580c994

Observation 7cca1858-305a-43de-a932-0019aa2902fb · outbound

This paper cites Mls: A large-scale multilingual dataset for speech research,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Mls: A large-scale multilingual dataset for speech research,

Reference 24

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:12:48.292894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.292894Z digest=sha256:fb58bb103059c5a606c3ddf641e6ff1e20f06f9cc014d16e132fa59a4b6fb301

Observation 8907619b-305c-480c-bbdf-d9ac1a0a2de7 · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Common Voice: A Massively-Multilingual Speech Corpus

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.384304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.384304Z digest=sha256:f4349207d259983e500ab110355e50dd5dac9d1edf9c91b6745f1d45220992a1

Observation 52b343b8-f89f-43d4-9dfa-e35a1fa2c00c · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.463964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.463964Z digest=sha256:828209e57af07b3f166dd35c11fe06e92ce73e99738159d23f46fb6ce2adca65

Observation 02ec53c7-2e9d-4eb7-ba40-45b105b4b1a3 · outbound

This paper cites On the utility of self-supervised mod- els for prosody-related tasks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation On the utility of self-supervised mod- els for prosody-related tasks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.352190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:49.357205Z digest=sha256:90402d70871c78dcf52526e72b4621c36815dce6d875b31bd85836b6845fe37f

Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · outbound

This paper cites VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.668479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.668479Z digest=sha256:74132d6dd8d20524771817e9595248f57a44cb071a8b2bf5ed47afc5bbb178f5

Observation 8c545bf0-d54b-4e6d-a7c6-dcd66f3bcc2b · outbound

This paper cites free public domain audiobooks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation free public domain audiobooks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.931336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:48.737679Z digest=sha256:e40414f7eb815ba6b18f949cd144fa0ccc46e6ff66b5f449eb46c51f4fcbddd5

Observation 3dc4be5b-c3cf-4eb8-bb71-2b8f6cede795 · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented transformer for speech recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.785451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.785451Z digest=sha256:914dffcb4ae0f93c3f9b705fd2e3a00188daf8d7bc97aa1a358a6ad5802b7ea4

Observation 91fb33ab-6728-4d4c-9828-3681a6d691d6 · outbound

This paper cites Long short-term memory,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Long short-term memory,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.979925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.979925Z digest=sha256:e75294dd7a3fdc5a6e94634f8e754558c193f05722a1092eb617f0754b5b7d79

Observation 4abe35ea-4d8e-4cca-845f-33de1adffd47 · outbound

This paper cites Specaugment: A simple data augmentation method for automatic speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Specaugment: A simple data augmentation method for automatic speech recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.723614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:49.032602Z digest=sha256:3e6cf1717977834fc7151f04f78176910036cab830bbfaed174b8d30499a6611

Observation 4690a66b-005f-4d52-9384-106eafc4bcdc · outbound

This paper cites Msp-podcast corpus: A large naturalistic speech emotional dataset,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Msp-podcast corpus: A large naturalistic speech emotional dataset,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.534911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:49.107373Z digest=sha256:5afbfa3e1a6029ea3bfce0ee81bd606a3c14d797f45814e9a928a9d9010d69af

Observation 98f3f17c-256a-433b-8ca4-77f34dc87fbf · outbound

This paper cites wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.162149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.162149Z digest=sha256:00f5484e98ec4a1401cae907a2d84ab5b2ba595ea4ed9679a1037d1c84ddbe05

Observation 2a313904-5856-45fa-90f8-15d6b4658691 · outbound

This paper cites mHuBERT-147: A Compact Multilingual HuBERT Model.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation mHuBERT-147: A Compact Multilingual HuBERT Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.231627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.231627Z digest=sha256:1f08e3bbe0c2b7c19e277e7d6a8c095d3d4e63833cb9b3325b5b9631c60b35a6

Observation f2af490b-06d7-4ba1-83c2-e87bfd0522b4 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.296912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.296912Z digest=sha256:c21ebbc0c841e9718dbd7dde602c90908c094813df41ae01c234e1a57ed7fbce

Observation 1e9dd8c2-89ec-41fb-9003-a684f8c85455 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.881846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.881846Z digest=sha256:6106104b543e70293b0fd426e16430e24d6ed2c73501ed545b6efc3ff9840a03

Observation 95b117fd-89c3-4d1b-b4e6-a7450b5d471a · outbound

This paper cites Multi-mode Transformer Transducer with Stochastic Future Context.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi-mode Transformer Transducer with Stochastic Future Context

Reference 2021

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:12:50.045674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-07T14:12:47.570835Z digest=sha256:59837db86bea687039a7daaece64c6ec11606786b72dfac3d3e36d0825e71260

Observation 1a5185df-5743-4224-b270-77f4de0e01cf · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.058460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.058460Z digest=sha256:3c8709c021a36dfd4e0e39b0acc72063d5df6db7155552f3a17a8fc447a89ec5

Pith citing papers

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · inbound

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation cites this paper.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:e8b44313237fbb5fba4580de7a0bcb1185b06cf9d36021d3f6e9f95c3c1f603a

Observation 4bbfd3a9-aead-4908-a0a2-b5d7e635a8a1 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.427323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.427323Z digest=sha256:802e1ec789f8c5a1215a47e4a1e34215dd1d3114aed0ba8cca4f67b36840bd48

Observation 582576fd-6365-410b-8a1b-82e0ab308188 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:39.509589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T05:32:27.371597Z digest=sha256:61f5df571a6799169a4164ce1569ef00b768dffaf144cf1ebfe635d4901414fc

Observation 8e31b141-02f7-419c-9087-a3b0b05247c2 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:51.419791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-29T05:01:44.202712Z digest=sha256:936b86adda0c1ef1e00d5aef63a1a8be5638a40b3e1d1b1f5e19a8f86a27c4d0