Pith. sign in

Paper Citation Record · LEDGER

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

As of 8 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 4 inbound Pith citation observations for arXiv:2505.19774.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.19774 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:49.357205Z

measured 43 of 43 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:12:46.355892Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-03T16:28:39.508010Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact3
  • verified fuzzy11
  • unresolved21
  • parse uncertain0
  • malformed identifier3
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 5f25726a-38b9-46a3-aa95-c855e7e20be6 · outbound

This paper cites Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Full-context encoders [1, 2], leveraging entire speech utterances, offer superior accuracy but higher latency, making them ideal for offline use

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:53.104993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.192370Z digest=sha256:0bf11194292a29f077606a5c74777904bdcce4c0c2588a5cb07775008b64e429

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · outbound

This paper cites DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:a53013496a6a9893486935452fde6dfc632b6c9e72525180d478e8d240666c7d

Observation a6ad7272-62b5-4be4-9a9a-02d1a43c21e6 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.685483Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.467017Z digest=sha256:43b933cfa452662ed46f23b0bbe35151eaa9cbead1b6a44409f7b39215623108

Observation fa9dcbb8-7061-43e1-b826-bdaec77172bb · outbound

This paper cites Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 1 presents a comparative analysis of the dual-mode en- coder and single-mode baseline encoders across full-context and streaming scenarios, in ASR tasks

Reference 4

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:52.286979Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.646277Z digest=sha256:e7f1d022d6eb47aaf22d55384e11671c35ba6b1336fea4cd5ecd4e3ae7ca6b76

Observation 46293429-1243-40a3-9b44-805d29587b3a · outbound

This paper cites Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Due to these variations, absolute WERs reported after our SUPERB bench- marking are higher than those reported in the original works

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.496675Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.552005Z digest=sha256:a91eebc8eef298f1f05ec109f216412199de940a869470926e1f8ecd2a33e313

Observation 2f177c02-655e-4476-ab67-569e5eed50a5 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:52.908471Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.274896Z digest=sha256:61f915d2b71d5a84678e41bbdacb8531fee5d5d8aef22b026ea20202ed562698

Observation c65c6e3a-f562-467d-81ce-1e2a83a8e978 · outbound

This paper cites Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4).

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Results Table 4 compares the performance of the DuRep-2B encoder in two inference modes with open-source encoders on ASR and non-ASR tasks, evaluated using SUPERB framework (3.4)

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:52.127133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.744428Z digest=sha256:fb24d7067da9d4d1e1e8e235d6381cdcbf20d1344382b8a39b787260282323ea

Observation ff259256-d9db-4d11-929e-1baa37b58f5e · outbound

This paper cites In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation In ASR-SUPERB evaluations, on average, the DuRep-200M encoder outperforms baselines by 13.06% in streaming and 10.42% in non-streaming mode

Reference 8

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T14:12:51.897019Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.835187Z digest=sha256:5d4e5e3af69b1b6b5a5af7b0655a7c19b6a31f3fa8806edc36fac659c870d33e

Observation 3df5a67f-bfa2-4c30-b26b-c66d3e4dc616 · outbound

This paper cites an unresolved cited work.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Unresolved cited work

Reference 9

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:12:51.660143Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:46.895063Z digest=sha256:376bf7e2829cee866375f7cd05bbded94ecb7c6794e0177ddc2a1343e6062f79

Observation c9700196-b019-4ea0-b376-8035a5b3ad6b · outbound

This paper cites Google usm: Scaling automatic speech recognition beyond 100 languages,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google usm: Scaling automatic speech recognition beyond 100 languages,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.963664Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.963664Z digest=sha256:df4e5883b45fcb86ddc6c6d1616d27551c0fa239ae792d2ed62ff8b785b8f7fc

Observation ec2afa29-f28f-4268-8c15-77ae428a0d6f · outbound

This paper cites Self-supervised speech representation learning: A review,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised speech representation learning: A review,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.108528Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.108528Z digest=sha256:4e66180e874bbac485309af3d79f75cb77663cc05275d63a6878445815e7e1d7

Observation ac4fef8f-9040-40f2-8b58-b96e78ba18bf · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Robust Speech Recognition via Large-Scale Weak Supervision

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.158251Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.158251Z digest=sha256:be6f81fafe301c9423131717e5b982ab214d246af5bd19fb9d5c291cc582ed44

Observation 7b8ab252-ca3f-4248-ad46-fc0a3ba2ac4a · outbound

This paper cites Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Dual-mode ASR: Unify and Improve Streaming ASR with Full-context Modeling

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.243261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.243261Z digest=sha256:6197e48cd4ed10e04ae55c8329bd4852d0f23296d7857642c8e6048a1d92e542

Observation 81e526fa-4163-487f-9b74-bf211c4000bb · outbound

This paper cites Conformer with dual-mode chunked attention for joint online and offline ASR.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer with dual-mode chunked attention for joint online and offline ASR

Reference 14

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:50.153372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.355685Z digest=sha256:7c9930c94189bca41073e19ba9ba2487f8324a92f1534e227c1023226058012d

Observation 4aebe4a4-8b17-490a-ae73-0d5d352b39a3 · outbound

This paper cites Multi- mode transformer transducer with stochastic future context,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi- mode transformer transducer with stochastic future context,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.451479Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.466237Z digest=sha256:cb75698be915606e4ca12afea5bff3ee7459cd9773230ac9a26c6b0e7e2244b4

Observation 5b58ca95-c45e-4e83-8170-4a4517a1083b · outbound

This paper cites Fleurs: Few-shot learning evaluation of universal representations of speech,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Fleurs: Few-shot learning evaluation of universal representations of speech,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.072372Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:48.584229Z digest=sha256:03043b550f80fb14c2c7bb879b1875a17cc0090354ce6f1475ba0fc7a676a7af

Observation 9acf3d35-c4ac-457a-8387-83306a72b0c6 · outbound

This paper cites Variable Attention Masking for Configurable Transformer Transducer Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Variable Attention Masking for Configurable Transformer Transducer Speech Recognition

Reference 17

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.920805Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.679297Z digest=sha256:76a8c1237d4075c9fdb4fca8a88739ac47b3d96fb371db876bd7d8c69acdfba5

Observation 862a48ce-cd21-4cf8-9386-2c06d30700c2 · outbound

This paper cites Sequence transduction with recurrent neural networks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Sequence transduction with recurrent neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.319686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.760230Z digest=sha256:4f765631faa61330a988053f05c1a2cf82b1c073637b3d73c9e8b2bd29d8b33b

Observation c822dc6b-1ea6-4b31-a86a-eea374679c94 · outbound

This paper cites SUPERB: Speech processing Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation SUPERB: Speech processing Universal PERformance Benchmark

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.829556Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.829556Z digest=sha256:445e5030257c4f0e9931aa712ef2deca8096113d76aced1a62edbfb11efbc84c

Observation 534ac630-f760-4521-9d3a-06d5baefddae · outbound

This paper cites Self-supervised Learning with Random-projection Quantizer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Self-supervised Learning with Random-projection Quantizer for Speech Recognition

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T14:12:49.761320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.911681Z digest=sha256:c255a07966bdfd5228abb76b0eb6cf92e6016bbab0507987b19088fe315d8d63

Observation 157c9368-a37c-426c-ae3f-e7a5d50f0181 · outbound

This paper cites Universal-1: Robust and accurate multi- lingual speech-to-text,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Universal-1: Robust and accurate multi- lingual speech-to-text,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:51.197138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:48.032670Z digest=sha256:d39138ea37e4df2c82072d1f148d8d66591dfce958c4b3008bc36f2f122ef6de

Observation b7da4468-3076-44dc-a9eb-cc06b1140fcc · outbound

This paper cites ML-SUPERB: Multilingual Speech Universal PERformance Benchmark.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation ML-SUPERB: Multilingual Speech Universal PERformance Benchmark

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.072445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.072445Z digest=sha256:b12c52e4a39e8236285492c8afcc3ba0be3fb649c382f42bc110746d5633cfa8

Observation a09b4471-d3be-4e13-9274-7b338103663e · outbound

This paper cites Lib- rispeech: An asr corpus based on public domain audio books,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Lib- rispeech: An asr corpus based on public domain audio books,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.177512Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.177512Z digest=sha256:1f5a7839450ee6bcb222f27960dbd504b4ddd90fc25ac2e6ec23f8bb407596db

Observation 7cca1858-305a-43de-a932-0019aa2902fb · outbound

This paper cites Mls: A large-scale multilingual dataset for speech research,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Mls: A large-scale multilingual dataset for speech research,

Reference 24

Resolution
malformed identifier
no resolver link, observed 2026-08-07T14:12:48.292894Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.292894Z digest=sha256:87aad9a87398395d96d3b1d95b188761f628681269b7b4db548cd884862814b9

Observation 8907619b-305c-480c-bbdf-d9ac1a0a2de7 · outbound

This paper cites Common Voice: A Massively-Multilingual Speech Corpus.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Common Voice: A Massively-Multilingual Speech Corpus

Reference 25

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.384304Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.384304Z digest=sha256:2a7d748448e85c13e197d4fa98a1a02060044721928fc663c747583d14004acf

Observation 52b343b8-f89f-43d4-9dfa-e35a1fa2c00c · outbound

This paper cites The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation The People's Speech: A Large-Scale Diverse English Speech Recognition Dataset for Commercial Usage

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.463964Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.463964Z digest=sha256:d7c966bea91c2b17751ba37bcfdf41809640dd91c1bec75e42449ffe50a97469

Observation 02ec53c7-2e9d-4eb7-ba40-45b105b4b1a3 · outbound

This paper cites On the utility of self-supervised mod- els for prosody-related tasks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation On the utility of self-supervised mod- els for prosody-related tasks,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.352190Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:49.357205Z digest=sha256:1c05c4b87d9d0d02a86aa1476d56d48568692cd894362aa0c38988f1f12cf79c

Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · outbound

This paper cites VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation

Reference 28

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.668479Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.668479Z digest=sha256:d2f63a7be88f3994fed910d8f9d36c245cd1b13d4ebf1805b3818c614978e9b0

Observation 8c545bf0-d54b-4e6d-a7c6-dcd66f3bcc2b · outbound

This paper cites free public domain audiobooks,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation free public domain audiobooks,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.931336Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:48.737679Z digest=sha256:13aa1000e52d2909b812ae4214bdadb800105abb4ad91baf70b62bb4a64d6a0e

Observation 3dc4be5b-c3cf-4eb8-bb71-2b8f6cede795 · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented transformer for speech recognition,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.785451Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.785451Z digest=sha256:dedcc45a5fb43ba70aa1629d8128bbf2b43bcfe92077abb9598757137a05093f

Observation 91fb33ab-6728-4d4c-9828-3681a6d691d6 · outbound

This paper cites Long short-term memory,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Long short-term memory,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.979925Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.979925Z digest=sha256:2c82c4a8283450e28d941bcecc295d569e38a75db3d8a8deb423b01825e4457a

Observation 4abe35ea-4d8e-4cca-845f-33de1adffd47 · outbound

This paper cites Specaugment: A simple data augmentation method for automatic speech recognition,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Specaugment: A simple data augmentation method for automatic speech recognition,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.723614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:49.032602Z digest=sha256:295f1efa65e2b41e497ed68fc83da4c3dd079b6353d306f64d1ec4d17482df92

Observation 4690a66b-005f-4d52-9384-106eafc4bcdc · outbound

This paper cites Msp-podcast corpus: A large naturalistic speech emotional dataset,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Msp-podcast corpus: A large naturalistic speech emotional dataset,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:12:50.534911Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:49.107373Z digest=sha256:534473ff26baf06306fac74ac9d85d39d720274d1f6296fb50d060905915c0e0

Observation 98f3f17c-256a-433b-8ca4-77f34dc87fbf · outbound

This paper cites wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation wav2vec 2.0: A Framework for Self-Supervised Learning of Speech Representations

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.162149Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.162149Z digest=sha256:252160e0aaf81e4cd782ce5c86bcc7cfa2d19bc5e53a7797180879b00541a808

Observation 2a313904-5856-45fa-90f8-15d6b4658691 · outbound

This paper cites mHuBERT-147: A Compact Multilingual HuBERT Model.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation mHuBERT-147: A Compact Multilingual HuBERT Model

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.231627Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.231627Z digest=sha256:41f791dbd4b765f32b15a079436ca44958a6fbfba3abb60d13dfcf1751108678

Observation f2af490b-06d7-4ba1-83c2-e87bfd0522b4 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:49.296912Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:49.296912Z digest=sha256:9682b2520d1abfa35262a96117692fdc98ec7b3f1049fcb74fd28b56fb0adc26

Observation 1e9dd8c2-89ec-41fb-9003-a684f8c85455 · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Conformer: Convolution-augmented Transformer for Speech Recognition

Reference 2020

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:48.881846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:48.881846Z digest=sha256:dafad3c69c2ddbaeb4e935163ac9be383a0294fd8d3edb256cfa1709f6162e89

Observation 95b117fd-89c3-4d1b-b4e6-a7450b5d471a · outbound

This paper cites Multi-mode Transformer Transducer with Stochastic Future Context.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Multi-mode Transformer Transducer with Stochastic Future Context

Reference 2021

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:12:50.045674Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T14:12:47.570835Z digest=sha256:9ee4a0a878e89ac04c2678274613784c1da43be804c9be6b51ba1490118f45f1

Observation 1a5185df-5743-4224-b270-77f4de0e01cf · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 2023

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:47.058460Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:47.058460Z digest=sha256:4831120a69a717e1a57c29fa6442928d45e1bb1eb9625e81552ce6bb1f0b91e9

Pith citing papers

Observation 9e950d8d-1520-491e-aa3a-3cb394e00e08 · inbound

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation cites this paper.

DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-07T14:12:46.355892Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:12:46.355892Z digest=sha256:a53013496a6a9893486935452fde6dfc632b6c9e72525180d478e8d240666c7d

Observation 4bbfd3a9-aead-4908-a0a2-b5d7e635a8a1 · inbound

Group Relative Policy Optimization for Speech Recognition cites this paper.

Group Relative Policy Optimization for Speech Recognition DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T12:07:21.427323Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T12:07:21.427323Z digest=sha256:5037a379c9eb0bb49a7c1ed541b5c9f2b816bac5d584be69d516cd5d02315b70

Observation 582576fd-6365-410b-8a1b-82e0ab308188 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-07-03T16:28:39.509589Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-27T05:32:27.371597Z digest=sha256:2dbaaf8b882b28cf6f9d9aece7ea3ffe1b97ea1805cd62dbe095962076e40f13

Observation 8e31b141-02f7-419c-9087-a3b0b05247c2 · inbound

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents cites this paper.

Adaptive Turn-Taking for Real-time Multi-Party Voice Agents DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation

Reference 37

Resolution
verified exact
arxiv_id, observed 2026-06-29T18:43:51.419791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-29T05:01:44.202712Z digest=sha256:9f5bd9d00a09c444de369b938d8945d9b7f10f174cc86d281948696211dd7217