Pith. sign in

Paper Citation Record · LEDGER

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems

As of 21 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2508.10456.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10456 v1

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:31:48.143276Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact11
  • verified fuzzy57
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 11db9d7c-8e3b-4198-a6bc-1b8e5fffbb68 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.040946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.040946Z digest=sha256:29a9e87232599f6d1e878794cdfe0592dec01e0d8259d5bbd4fca303f8621472

Observation f7f1ea8e-1476-4043-9d28-2b0aa8db8904 · outbound

This paper cites Hybrid ctc/attention architecture for end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hybrid ctc/attention architecture for end-to-end speech recognition,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.145089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.145089Z digest=sha256:15f7630c46cf733652dda10cd757a2f1565645232aa403d6f7a5d374771b3e96

Observation 2095d3b1-e8d6-49e8-afab-564009151f8f · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.228233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.228233Z digest=sha256:79d365a61831316ebeb681424ef8808cead4c8a4d46afba15b911b9d61be11b0

Observation b4c69068-b269-4da4-9bc8-7ac0012145a2 · outbound

This paper cites Attention is all you need,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attention is all you need,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.340737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.340737Z digest=sha256:e4326ae4079a0fa5ae0f2552ba5a3212b561d40a983bf2d1513fdb60074ff243

Observation ab75db89-276c-470b-9458-8c626c82fc0f · outbound

This paper cites Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.447669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.447669Z digest=sha256:a77405a43b1d69d8631d5f3320d514aeb651769a4f8a9233d2c6e1204840abab

Observation e1fd1b12-f44a-4d9d-bafc-2a70d90e6f47 · outbound

This paper cites A comparative study on transformer vs rnn in speech applications,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems A comparative study on transformer vs rnn in speech applications,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.537168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.537168Z digest=sha256:206c5ae11b1668904e496300e3cf1d8bbba02519c8cc7981c4be8a2f523810fe

Observation 4bee5d7f-62eb-48c2-82d2-5097c1aac52e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.633846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.633846Z digest=sha256:50b584a05547f4ecc9d8bd8dd07aa97c19f39044e085cdbf8def3c7abf16872c

Observation d2390ca1-8927-4b0a-acd4-f2140f2dfb33 · outbound

This paper cites Recent developments on espnet toolkit boosted by conformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent developments on espnet toolkit boosted by conformer,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.712245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.712245Z digest=sha256:ac3650eafea9133a6a14733ad0725326a2a6511b6f76711aa15fb853be7c6263

Observation e7c11225-302e-4ce8-bc5e-05b7baf6788f · outbound

This paper cites Sequence Transduction with Recurrent Neural Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Sequence Transduction with Recurrent Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.796983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.796983Z digest=sha256:3c4b3e4876762f8133c09a5fad3ea35c23d861d9e308b9260a1e9ed6e9f72f6d

Observation 047e0586-806e-42cd-81aa-7574d9fa98cd · outbound

This paper cites Recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recurrent neural networks,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.914416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.914416Z digest=sha256:c8c4885d24a65e4a729de4dceef2c14c9be7604e02c06cae33e0258b5f7740a8

Observation 3c41ce41-2d7b-4270-a270-7a0ebda4a6d8 · outbound

This paper cites Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.032849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.032849Z digest=sha256:82eee93f83f8b03d7c7da796f71ceec74e8ada89738f67881edf0db7ca2a58f4

Observation 6d4faef4-48f6-4892-a2c7-dc86704e821e · outbound

This paper cites Efficient training of neural transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Efficient training of neural transducer for speech recognition,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.112682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.112682Z digest=sha256:99b3668af4763c826f394ecf5568f8e1adc953bc7278122e647c7bf2b564b4f1

Observation 995b2255-cb3e-47ef-98d0-352f5904b3aa · outbound

This paper cites On the limit of english conversational speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems On the limit of english conversational speech recog- nition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.213383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.213383Z digest=sha256:dcae4e68b2dc772b0503da222c367b9620d878d35cfa30dc09ae0f68e6a5727d

Observation 7146709b-3bbc-4ffe-be3a-99477a6906f5 · outbound

This paper cites Improving the training recipe for a robust conformer-based hybrid model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving the training recipe for a robust conformer-based hybrid model,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.327114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.327114Z digest=sha256:41e58c73f039e640ffdac3007a66f53d9530bc8d5bb42aaf7e951fca5e5d4efd

Observation 9cd39ca8-bd36-4e63-9213-3cfcdf89c110 · outbound

This paper cites Confidence score based conformer speaker adaptation for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Confidence score based conformer speaker adaptation for speech recognition,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.415984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.415984Z digest=sha256:c9e9d1de03304ad376f31445a7ac268c840a0fa9bc1f59d7cf50b36b71ea3362

Observation 00bb68b6-2a8c-4174-b0bf-d35ecab8dbb3 · outbound

This paper cites Diagonal State Space Augmented Transformers for Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Diagonal State Space Augmented Transformers for Speech Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.243108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:40.515370Z digest=sha256:a072917e5cbf09f8fa1550abf24874d46305fb3d82842a9460772a351edb2547

Observation 0617d538-d4e0-42ce-b3bd-101d1af0fe27 · outbound

This paper cites Recent advances in end-to-end automatic speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent advances in end-to-end automatic speech recog- nition,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.601112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.601112Z digest=sha256:b79c6ec1b3b49fa7b8285a504d86cd5bbb254244437961ee6f7046313ab41d9f

Observation 0bcb61f5-291a-410e-b19c-38ee77418444 · outbound

This paper cites Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.028105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:40.689037Z digest=sha256:f45122f61c5447a9b89e37d2cff3d7dd78d39faad901c71c32ed705b190625b7

Observation ec1cfdea-1210-4095-8423-246400ab3e25 · outbound

This paper cites Improving asr contextual biasing with guided attention,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving asr contextual biasing with guided attention,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.773045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.773045Z digest=sha256:81a58ee9d6b9a27af93c2234641381034fffc2379a8ebcdb95f97d1c0665382d

Observation 0d3c780a-92b5-4b2b-8207-abaff055228d · outbound

This paper cites Optimizing byte-level representation for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Optimizing byte-level representation for end-to-end asr,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.851093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.851093Z digest=sha256:39d01529a21cc1a820f9762216029cc0e6c8522585922bde9a83fddc9b341a05

Observation d9907ec3-f172-4699-9634-6a4165bf1ed9 · outbound

This paper cites Contextual modeling for document-level asr error correction,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual modeling for document-level asr error correction,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.979654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.979654Z digest=sha256:d0acda7e4680d9adf64ff9b2f244207a9582ffeaf31d566b4063e2a15cd2a9a2

Observation 9c199e58-a953-4b7a-9239-740734c6c521 · outbound

This paper cites Promptasr for contextualized asr with controllable style,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Promptasr for contextualized asr with controllable style,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.065159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.065159Z digest=sha256:c8c19a0115833d4429babc82d7a9a6a5591bdfa65d032ee224b4f6e486223217

Observation 6f881a35-e3a6-4ae2-80d2-cd350f18f212 · outbound

This paper cites Conversational speech recognition by learning audio- textual cross-modal contextual representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conversational speech recognition by learning audio- textual cross-modal contextual representation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.143119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.143119Z digest=sha256:cc58a0e982a2c69f2502ce80f5aafa06ba57fb3db9b34605fa1a9eb4bcebf460

Observation 5efe2905-0318-4095-96cc-99e8f0f248b1 · outbound

This paper cites Dual-mode nam: Effective top-k context injection for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dual-mode nam: Effective top-k context injection for end-to-end asr,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.229606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.229606Z digest=sha256:062c54576b836dbef75a82eca49a549f6c9428c4b7b1b934177d09ac459400fa

Observation 3e8e1dbc-1089-4d79-ab86-e7db061ad785 · outbound

This paper cites CPPF: A contextual and post-processing-free model for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CPPF: A contextual and post-processing-free model for automatic speech recognition

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.766549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.333464Z digest=sha256:4911e67bf0d2c6120e9a7479d43f0bc55580da5a42b1ff7a984e7b15a6801509

Observation 71a14c27-013a-4ac2-9140-e92a55b71557 · outbound

This paper cites CASA-ASR: Context-Aware Speaker-Attributed ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CASA-ASR: Context-Aware Speaker-Attributed ASR

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.531977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.404760Z digest=sha256:0d0d4f64a7df07689be920f1b6d7582f2b0569275d259340f5f1b2514562d6b4

Observation 593b9995-e5c3-43dc-8322-202f0bcb9e72 · outbound

This paper cites Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.493267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.493267Z digest=sha256:fb60abfbcff012ea894f592af2fdf9663e59dea3dad4a31ae3bf95c364215c0e

Observation 461d666f-6e32-4906-83d4-166bd3dec097 · outbound

This paper cites Bring dialogue-context into rnn-t for streaming asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bring dialogue-context into rnn-t for streaming asr,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.257605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.598588Z digest=sha256:0dc304df58fa784b361c955bc26d9afd18aae277cf7a9b9dee36eba89cc11fdb

Observation 74b8dca5-c583-4fd1-acbd-7016962c5497 · outbound

This paper cites CopyNE: Better Contextual ASR by Copying Named Entities.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CopyNE: Better Contextual ASR by Copying Named Entities

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.264086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.707826Z digest=sha256:4f4612f7f305c5ec212b391ea377856afae330130ad5ef1405228d572c977e0e

Observation 08faf3b5-bde3-4fae-a1fe-1c067bc2b645 · outbound

This paper cites Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.050215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.805565Z digest=sha256:9ba5114fc39a3cb4530cf2e14e2ed90f215d3b73601c477a17d391d51a48c7c5

Observation 6976842c-c120-407c-8157-c58085e3a12f · outbound

This paper cites Training language models for long-span cross-sentence evaluation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Training language models for long-span cross-sentence evaluation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.036595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.884033Z digest=sha256:ddf6ed2e85a223deeacee50f132d91855dd68b264eec1c4120844f3ae3162c3c

Observation d274690e-fcbd-4c06-aefb-1596b24930c2 · outbound

This paper cites LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.827274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:41.984032Z digest=sha256:26d7ac71c82b9270ff06d5c6991ea13011faa7cc46077f3eb18974237596381f

Observation a44c5369-b7b4-4c4f-938f-b250fe31b8f0 · outbound

This paper cites Session-level language modeling for conversational speech,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Session-level language modeling for conversational speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.851580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.099735Z digest=sha256:41392b7494f540c032c033526019a80204c316e328f31b198925b235db2ee233

Observation 1a018f5b-d065-4fb8-9e11-3bc6c34ad837 · outbound

This paper cites Transformer-xl: attentive language models beyond a fixed-length context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-xl: attentive language models beyond a fixed-length context,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.636395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.175833Z digest=sha256:9ecd0ed8320f8ec82e3ae0e6e614065f696067ed16855ffdf851b5646670c6df

Observation de503af8-0a5f-4b5b-8e64-72b83282e881 · outbound

This paper cites End-to-end speech recognition on conversations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition on conversations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.225580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.289439Z digest=sha256:9484b111fc45fd27e554347407dcd19ee219d73e9aeae0c7d73434f5854cde70

Observation 20aad9f9-3f85-4409-9fd3-f433b7213289 · outbound

This paper cites Contextualizing ASR lattice rescoring with hybrid pointer network language model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualizing ASR lattice rescoring with hybrid pointer network language model,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.827745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.363244Z digest=sha256:6d77dbfe8a9ae6e91ba0565494810cb0ad9a61c1aca6add472030d2e0773fa81

Observation 893e0851-d763-4751-9ca6-b08b2789e449 · outbound

This paper cites Use of contexts in language model interpolation and adaptation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Use of contexts in language model interpolation and adaptation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.689635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.455433Z digest=sha256:b14beebaf33e8a6e6c8b438e8c7e41b9e8c7698613d8f728a460c1cab4a49698

Observation 4a711db4-3544-43ce-89b3-b878233fdb68 · outbound

This paper cites Longformer: The Long-Document Transformer.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longformer: The Long-Document Transformer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.516968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.516968Z digest=sha256:6f43ea878395083f975e2fb88ff74101496e05d3f421e79a5a96c7b3162afabf

Observation 29dfea9b-e96d-4063-8f0c-75fe30248ea4 · outbound

This paper cites Transformer language models with lstm-based cross- utterance information representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer language models with lstm-based cross- utterance information representation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.523332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.603756Z digest=sha256:bf34feffb93484606a73903838f583d79dad8ff383288f4d38fcfb852cf25ed2

Observation 9d5d7fc7-5001-4462-9ccb-d1aa6801088e · outbound

This paper cites Ctc-assisted llm-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Ctc-assisted llm-based contextual asr,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.365324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.670064Z digest=sha256:4eff32241f821c33ff2961c5f5dc1614ed7a48623714706db1943691159d1f97

Observation 49e04d82-0591-4a2b-bcfa-165cee1248a3 · outbound

This paper cites Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.749315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.749315Z digest=sha256:758420d42cd329289331eafae19c8670adce6cef30bd3ae8355a6557d1b61b74

Observation 40e7f0d3-ce42-4085-9876-cddf36846caf · outbound

This paper cites Contextual asr error handling with llms augmentation for goal-oriented conversational ai,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual asr error handling with llms augmentation for goal-oriented conversational ai,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.195690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.859181Z digest=sha256:d984177cae73bca71287e30f003f77f674543169b68441be69af39bd3eb643bc

Observation 091432e8-62b9-48a7-86de-4e49e62bfb26 · outbound

This paper cites Towards asr robust spoken language understanding through in-context learning with word confusion networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards asr robust spoken language understanding through in-context learning with word confusion networks,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.005515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:42.920140Z digest=sha256:1f39ec557ed76ec583b4ca60ac0c59588c11adaaebb3c7251a24772b26616a69

Observation 14d87e6b-ea9c-43bd-a846-96e9386cd530 · outbound

This paper cites Contextualization of asr with llm using phonetic retrieval- based augmentation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualization of asr with llm using phonetic retrieval- based augmentation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.856857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.000136Z digest=sha256:a749a7c598d581a917e128ce5c9e7637e33ed35974a66a7e748a8d4c3f9a6d52

Observation fd7b20e2-e5d2-4c48-8491-71acd464f171 · outbound

This paper cites End-to-end speech recognition contextualization with large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition contextualization with large language models,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.746831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.104408Z digest=sha256:6bb9933171c239edc5adfa0f4d38ee97963b19ea6e47f8799f7dec8882e6e65b

Observation 91402f0c-236e-4798-b02d-1be9e1a24467 · outbound

This paper cites Improving domain-specific asr with llm-generated contextual descriptions,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving domain-specific asr with llm-generated contextual descriptions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.599667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.219288Z digest=sha256:ef908c9c9c1cef9d84e9b0262a58ffe600fee274d7cc0a85f8dae7589eca0e50

Observation e5436655-70a5-4320-af73-a6360e8e52d1 · outbound

This paper cites MaLa-ASR: Multimedia-Assisted LLM-Based ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems MaLa-ASR: Multimedia-Assisted LLM-Based ASR

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.498400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.295375Z digest=sha256:9d376e0cbc5b9621c8b49eff3df4c18e3777a32acfe7ee41e60621e4f1883f8e

Observation fde25a80-3665-471e-971e-2ca0bdc51efe · outbound

This paper cites Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.403844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.373827Z digest=sha256:58e5dddd8ee8d5f06f38fd10cf1548482f926cf2af313d92897c05496ce5a1ff

Observation 4eecbe69-a9bd-41d0-9f06-79167b1a6412 · outbound

This paper cites Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.237641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.456893Z digest=sha256:3b16b325031d4f126ec9e56f865850872d6ea898d5466e146a1b51a99e40e6c6

Observation af0de628-c71a-4cc2-94db-f74b2e7f04d3 · outbound

This paper cites Using Large Language Model for End-to-End Chinese ASR and NER.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using Large Language Model for End-to-End Chinese ASR and NER

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.555162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.555162Z digest=sha256:643ee051482f23eb98339f073f1fe82e1330a386e3f695acef3ffd47507ce6da

Observation cdcaf6aa-a5b6-4586-b7a8-9ab090614695 · outbound

This paper cites Compressive Transformers for Long-Range Sequence Modelling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Compressive Transformers for Long-Range Sequence Modelling

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.616971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.616971Z digest=sha256:0bfb78c4758fd4130e75c41b2da1ba311aa5bc6ebef777fc02a093f2b310a035

Observation 4d49d666-82cd-4600-a11c-daaf11c69055 · outbound

This paper cites Speaker-aware Speech-transformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker-aware Speech-transformer,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.090493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.710776Z digest=sha256:4642f4de227a4ff643e259322a96d57eb1ac5b314475cb011d25dd4488b7423d

Observation a151e8a4-7ff0-434a-9d7a-4e0a24c7906e · outbound

This paper cites Transformer ASR with Contextual Block Process- ing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer ASR with Contextual Block Process- ing,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.946512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.789641Z digest=sha256:047cbc39f196ec717559c0526d65cefe6fd5c3a09421a0ddb4be78c0aee10a86

Observation c9307482-10d9-4ae7-bc12-dbcab7aa9904 · outbound

This paper cites Transformer-based long-context end-to-end speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-based long-context end-to-end speech recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.806972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.856177Z digest=sha256:742015b8b9a1d18787289e5ec282de3747c08e3dac5b440e58759c996b15036c

Observation 24bdc12b-946d-40df-a2f9-8cf21cc5a672 · outbound

This paper cites Advanced long-context end-to-end speech recognition using context-expanded transformers,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-context end-to-end speech recognition using context-expanded transformers,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.684237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:43.939035Z digest=sha256:34d547c42fc6296ba2e931c88a674f8ee2039d39c14d8fbe16f7378496dcd300

Observation 89b091be-dc99-4f00-a415-0b62fc89a353 · outbound

This paper cites Context-aware end-to-end asr using self-attentive embedding and tensor fusion,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware end-to-end asr using self-attentive embedding and tensor fusion,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.527917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.010039Z digest=sha256:5f7d400287aec0caaa1398af9c3bf5fff796eacc9bd3864a7dd35bb57fb7079a

Observation 97cf2b87-6708-4372-9ff2-d8c13f6f46ce · outbound

This paper cites Longfnt: Long-form speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longfnt: Long-form speech recognition with factorized neural transducer,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.374355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.107285Z digest=sha256:17e4227f4de8e71f486091a7dd693a5d9bff2b1a019d5db293a10f171f33638c

Observation 9bf31ccf-200c-406b-bb7e-aaf0ef13288c · outbound

This paper cites Advanced long-content speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-content speech recognition with factorized neural transducer,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.213167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.172636Z digest=sha256:3221dcfc3d8da3b355fc6e3ef56fea9dc2d3b9fdcbf8d7b350b5f8ee8e42ce0b

Observation 288a580c-37a1-4974-87da-c5a9a22e0703 · outbound

This paper cites Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.180696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.256170Z digest=sha256:86e43863193772949d9d513a3918ef22bf21863ca7f33d90688709c508db3e9a

Observation dace3c57-6b23-4716-b1f7-579469cae9fa · outbound

This paper cites Dialog-context aware end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dialog-context aware end-to-end speech recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.039618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.339765Z digest=sha256:23582fd11918d97800ba5101c972afc783ba799058b4db6f1ea433d412733f3d

Observation 500088fc-abbc-4dc7-970a-27a59f22f618 · outbound

This paper cites Improving RNN-T ASR Accuracy Using Context Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving RNN-T ASR Accuracy Using Context Audio

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.854090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.437262Z digest=sha256:4b938291543c288915df759825aeadfb738e38663f525bd6abeb002e38940e57

Observation 58ce565f-11bf-4f36-af09-e80c3f164fc3 · outbound

This paper cites Large-context automatic speech recognition based on rnn transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Large-context automatic speech recognition based on rnn transducer,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.894247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.553967Z digest=sha256:db5c42145c2468d6afab8305798b61b48ea4efaf98e12d7cf991b6731d04e14f

Observation 15b7a962-8d94-426f-92b8-22bd2eff6411 · outbound

This paper cites Phoneme-aware encoding for prefix-tree-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Phoneme-aware encoding for prefix-tree-based contextual asr,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.762269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.635612Z digest=sha256:6a31de7609b3c50b208597ca25756a94d0dbaaa3ae82635de7d11aea4ea8854c

Observation 35ac3e39-c498-464b-8d23-00baa8686d08 · outbound

This paper cites Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.551140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.729117Z digest=sha256:11b07d3b239531e1399fdf5f7c7fa99ccd1e7b58d07b97921d2f24b921bc377e

Observation 93f9e98f-fa5e-4f4c-a78a-0cc1a17b62a3 · outbound

This paper cites Contextualized end-to-end speech recognition with contextual phrase prediction network,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized end-to-end speech recognition with contextual phrase prediction network,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.611543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.809319Z digest=sha256:bfc53af01d32be103a8f93041f831d880efc38dad5975af63d56581bbaadb7bf

Observation 90f4525e-4987-4c89-9d6d-2f4f2e527839 · outbound

This paper cites Semi-autoregressive streaming asr with label context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Semi-autoregressive streaming asr with label context,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.459164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.893812Z digest=sha256:8322129a3987bef3f32bfdf5be4cb0759b71e28062d061c17ea7b2e5016b4540

Observation 8e0f3d92-8326-49b5-b0de-56fd29b0fb72 · outbound

This paper cites Context-aware transformer transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware transformer transducer for speech recognition,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.309374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:44.983848Z digest=sha256:5f667d4c2a3b9ca123fbd464f4a75b2b176a9d8522a52425330e9633d39b054b

Observation 10454637-34e0-4bc5-a807-76b2982a8845 · outbound

This paper cites Towards effective and compact contextual representation for conformer transducer speech recognition systems,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards effective and compact contextual representation for conformer transducer speech recognition systems,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.180065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.047116Z digest=sha256:f85de60b6192c0820cdd67b303631d6217140e587739d940d185281ada1820e1

Observation 3c8d4152-b711-4728-aacd-1b96368900e8 · outbound

This paper cites Wav2vec 2.0: a framework for self-supervised learning of speech representations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wav2vec 2.0: a framework for self-supervised learning of speech representations,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.044753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.119425Z digest=sha256:d531cae6b34380818fe84e7dfa928fd52c9bab83d2111a1844bf96db57339a26

Observation 5701ae8b-ddff-49f9-80a5-e7d78d2fd852 · outbound

This paper cites Wavlm: large-scale self-supervised pre-training for full stack speech processing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wavlm: large-scale self-supervised pre-training for full stack speech processing,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.898704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.187611Z digest=sha256:323790bb840074213b370e45374caea0d7bd6bc0852d6d83fedd7a15c2b675ce

Observation 3816819a-e1c4-41b2-91fc-8eda76b5cfc6 · outbound

This paper cites Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.727367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.269084Z digest=sha256:3875b9647e97cfd168242573cf3ab9f8b675d093eda1b43ee5b488afeeb2f158

Observation d25de274-8d78-40ef-8a6c-2c55cd28f302 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.559832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.364942Z digest=sha256:dacfa5b52b42840806f2f36bbf83758ff3b2151a9e6037516891b0b78648b7c8

Observation 1b8b7bd8-9bc8-49d5-8bcf-a14c1f119d77 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.366817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.431578Z digest=sha256:ff6c44fcbd62084af4daef242fd64b36cc2afe6342e120f6e83a799e36fc157d

Observation 6d1bce3b-7bed-4b95-a086-62e88057544f · outbound

This paper cites Data2vec: A general framework for self-supervised learning in speech, vision and language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Data2vec: A general framework for self-supervised learning in speech, vision and language,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.203752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.508204Z digest=sha256:188f1bdece5f98fdf4b62a2ba4df39fff6ce83a9f002f655a30112a268542225

Observation 2f732a3c-5b53-4682-ad90-78c29eb96669 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Robust speech recognition via large-scale weak supervision,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.018209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.576141Z digest=sha256:9ec04bd23a178246a3f858aa32bb3e61c1011f3aca45a0b6312ba9aff241dd04

Observation d1bf3b75-9894-4e2f-9d00-f8bb23dc60eb · outbound

This paper cites fairseq: A Fast, Extensible Toolkit for Sequence Modeling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems fairseq: A Fast, Extensible Toolkit for Sequence Modeling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.668112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.668112Z digest=sha256:ae82d5c08558482c4200737f6e6a7085f14974dd00ce532e11aa82dd773125b2

Observation e171fbec-8b85-4df6-ac32-35cbfffa24d1 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Zipformer: A faster and better encoder for automatic speech recognition

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.751156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.751156Z digest=sha256:eddfd8403b6e6c5861e3d312dcdf9aada92fc9cb205a33cd2b88338c0921041c

Observation d7e200e4-be8f-4994-bed1-132c5b6756fa · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems ESPnet: End-to-End Speech Processing Toolkit,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.878240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:45.863510Z digest=sha256:19b8feabc0b68d4657fb206026223e036f3e1d71db39f0351e337f3e8d364729

Observation 0a05ab93-c5b3-4b3c-9891-049835cbe534 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.946423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.946423Z digest=sha256:5b6bbfe030a3f3d26f1027e6f0546dcf03813b2b473ef46fc52807183eaf29eb

Observation bf08cde8-13d4-4769-93f2-da483918513b · outbound

This paper cites Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.678706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.004011Z digest=sha256:87560d2d4246963cd49377b4ead3c38dda2d33063e9c9a23c63f2e0d9c1bbf48

Observation 41321727-56bf-411a-a94d-3d2ea0de4481 · outbound

This paper cites The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.500000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.098424Z digest=sha256:c2872a0472ee838b6b7438cb47d83d53a769c6e02a01598a18fde221477f1f11

Observation 877244d9-6847-4c06-97dc-e304f44a1a21 · outbound

This paper cites Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.337853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.188228Z digest=sha256:75d6b18f99064f696433a6a868bd762574ebeb61e7c4570cac10e2550c7f7ee1

Observation f96dc9c6-d6f3-4266-81f9-4c88cbab1a89 · outbound

This paper cites Self-supervised asr models and features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Self-supervised asr models and features for dysarthric and elderly speech recognition,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.172237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.265388Z digest=sha256:9a0215c02d0942df806a1e1bec4180aae2ca7d6301b41377d25df7aac13eb555

Observation fffb07a8-d62b-4a14-b776-a01fae785d9e · outbound

This paper cites Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.979937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.350732Z digest=sha256:246e0e15a077b9ea134fd769c6a7f7cc1920441959feec8b55d6eeedef1a9f7d

Observation 6a90f7b1-9a5d-412f-bc2f-ffe669f11b70 · outbound

This paper cites Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.827935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.434543Z digest=sha256:7f08726c41c6a72f0e4f214f776bfd17930dfbde985abcc7b8cbb42d1a00b2c5

Observation 9a4deac2-3a99-4335-b266-c56ae1695911 · outbound

This paper cites Transformer-Transducer: End-to-End Speech Recognition with Self-Attention.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-Transducer: End-to-End Speech Recognition with Self-Attention

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.546745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.546745Z digest=sha256:39c1a0e101740f3e9ef4caa2f1bb4dea618b3de12fe5457bd366104504bdc47d

Observation 9cd3241e-12b5-4540-9b0d-224aa63924af · outbound

This paper cites Language modeling with gated convolutional networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Language modeling with gated convolutional networks,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.598121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.629844Z digest=sha256:b0b29b97b883979a803c7322a5d35a45e5b22037b6537484840d770ed41899fa

Observation ead4386f-fbc7-462c-87c3-4eb7b118fbc1 · outbound

This paper cites Attentive Pooling Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attentive Pooling Networks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.700394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.700394Z digest=sha256:04cba5add2e086ac5f3d1166e0b4539b3d62ec0c871a6c4e1e1039f574974025

Observation 0ae1fd88-4d05-4373-bcbd-1ce6a34afdb4 · outbound

This paper cites Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.437465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.810510Z digest=sha256:0e38e39ca05835a57708866f332c2ee82b78875b2151749c7e3d77525fc828b6

Observation bf71d3d1-869a-41fa-8a5f-4ea6c3ede40d · outbound

This paper cites Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.276825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:46.871585Z digest=sha256:81e8224204c6d808e8a0c845c968ed5e0b96dbad38e28233b8f1df43a0c2c82c

Observation 0d376c8b-bcb5-424a-b8d6-1e2a86371335 · outbound

This paper cites Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.119919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.037997Z digest=sha256:bf75baa3c97ab2cc9dd33bb58d31d8028d857c6acc770205ece96da10b5428fe

Observation d2f7a809-7897-40ff-8f05-8ffa777f06f6 · outbound

This paper cites Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.961044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.136299Z digest=sha256:25481743aba8f0300138113cd03d5d664432fa94c2acef7b52418b7194533a5f

Observation 6b58fa7c-73a7-4d23-bdf8-1f82f5f72748 · outbound

This paper cites Some statistical issues in the comparison of speech recognition algorithms,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Some statistical issues in the comparison of speech recognition algorithms,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.783057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.141110Z digest=sha256:ecf505cbbab0688b161bb67af5aba811f75fab375593966d16604a9dc1866d3f

Observation b9a8482f-8ffd-4bea-ac15-1162bb3c1154 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.623835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.221732Z digest=sha256:50b5f099b4a405a03ecfecb72b3cf22bfac8a02742a6030fd21967c88f555a47

Observation 4c969da5-5255-4ec0-93e0-c7c517cc5419 · outbound

This paper cites Tools for the analysis of benchmark speech recognition tests,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Tools for the analysis of benchmark speech recognition tests,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.403485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.380300Z digest=sha256:2e103a3491e7b878d5aca6ac850bdc189538547cc012a85e04303d6665d08782

Observation 44c0c64e-9953-4c8c-8a7e-eb64a6067c21 · outbound

This paper cites Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.219031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.521344Z digest=sha256:3afecce0ef41271b0d1f72b4adfb815021862acbd951af94b41bb30d8d6f7517

Observation ff45c5f5-5b0f-4e81-8a1a-ce0dfad35a06 · outbound

This paper cites Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.029574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.653179Z digest=sha256:ceaeb6fe4a9abff28eef3694b527e758d40315196bedeabfa15ea5d17e0e82b0

Observation bc9f7db8-6cba-4681-9c3c-de2bb60715e9 · outbound

This paper cites Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.813760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:47.808713Z digest=sha256:c901618102397b1a5a5ad2c978fef74dddea9794c989b9dbe75a32844e911c8a

Observation 65f23f34-7a71-4544-b5e4-7a1fe7f6ee55 · outbound

This paper cites Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.616251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:48.005043Z digest=sha256:81b89bf433ff00cdb280993407fd19b67c02fc816a837db26b94a4e795bbe6ec

Observation fc64bde5-cc7b-458a-adf0-b7ea595b99f4 · outbound

This paper cites Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.418413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-08-05T20:31:48.143276Z digest=sha256:534d485138410bfaad8a9cde4afcfde1d92a9acbdbf00887c68fb791a2fc3273

Pith citing papers

No inbound Pith citation observations are available.