Pith. sign in

Paper Citation Record · LEDGER

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems

As of 8 August 2026, this Paper Citation Record lists 100 of 100 outbound references and 0 inbound Pith citation observations for arXiv:2508.10456.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2508.10456 v1

Coverage vector

measured 100 of 100 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T20:31:48.143276Z

measured 100 of 100 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

100 of 100 outbound references displayed

  • verified exact11
  • verified fuzzy57
  • unresolved32
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 11db9d7c-8e3b-4198-a6bc-1b8e5fffbb68 · outbound

This paper cites Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Connectionist temporal classification: labelling unsegmented sequence data with recurrent neural networks,

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.040946Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.040946Z digest=sha256:fed7dd37de4e6db6368c590bca8c69908baba7e48041e0bc564fa5ed3be5b33f

Observation f7f1ea8e-1476-4043-9d28-2b0aa8db8904 · outbound

This paper cites Hybrid ctc/attention architecture for end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hybrid ctc/attention architecture for end-to-end speech recognition,

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.145089Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.145089Z digest=sha256:329e45bb682906808b86d973918a9759e21f9a03b9dcdb822285fb6e64e9ca15

Observation 2095d3b1-e8d6-49e8-afab-564009151f8f · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.228233Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.228233Z digest=sha256:5943d723b1efd3f1cc7916eb71451cfd0ad01d05cdc1dfe13f13d2aab2fdcf27

Observation b4c69068-b269-4da4-9bc8-7ac0012145a2 · outbound

This paper cites Attention is all you need,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attention is all you need,

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.340737Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.340737Z digest=sha256:ef4d016e85adf7b5c39b153d99d88ae60731eb921b2dd6c40266b7c079e05c51

Observation ab75db89-276c-470b-9458-8c626c82fc0f · outbound

This paper cites Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speech-transformer: a no-recurrence sequence-to- sequence model for speech recognition,

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.447669Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.447669Z digest=sha256:8ac23cff845bf3ee45ec656b2ab0e30acc53bbe78fe889121ba9ed7b24325de4

Observation e1fd1b12-f44a-4d9d-bafc-2a70d90e6f47 · outbound

This paper cites A comparative study on transformer vs rnn in speech applications,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems A comparative study on transformer vs rnn in speech applications,

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.537168Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.537168Z digest=sha256:38267c0d81bb93a04f4ab617ff86cdc75fdbce6453603b7a4769569af79d5923

Observation 4bee5d7f-62eb-48c2-82d2-5097c1aac52e · outbound

This paper cites Conformer: Convolution-augmented Transformer for Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer: Convolution-augmented Transformer for Speech Recognition,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.633846Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.633846Z digest=sha256:b73c8e2ad78d82aadb15d06e77f63344adb42fa149452fe6095d6202af533499

Observation d2390ca1-8927-4b0a-acd4-f2140f2dfb33 · outbound

This paper cites Recent developments on espnet toolkit boosted by conformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent developments on espnet toolkit boosted by conformer,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.712245Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.712245Z digest=sha256:cbd458f889deab6e6830c83075b3db6c33c5f1310dd3b4cb159e727f4eb324d0

Observation e7c11225-302e-4ce8-bc5e-05b7baf6788f · outbound

This paper cites Sequence Transduction with Recurrent Neural Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Sequence Transduction with Recurrent Neural Networks

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.796983Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.796983Z digest=sha256:ff2d34378175edddb82d16ae38651245734f9af776d964a65081e61a11edd13a

Observation 047e0586-806e-42cd-81aa-7574d9fa98cd · outbound

This paper cites Recurrent neural networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recurrent neural networks,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:39.914416Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:39.914416Z digest=sha256:3643e33f9329f50756ba783acdb880e21a26444b28368cd14fe0ba72eb3965a9

Observation 3c41ce41-2d7b-4270-a270-7a0ebda4a6d8 · outbound

This paper cites Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Long Short-Term Memory Based Recurrent Neural Network Architectures for Large Vocabulary Speech Recognition

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.032849Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.032849Z digest=sha256:bb7c2e97f778832afbeb2012efa4266baee9b8e40f2a6f80663d8e576beca1ff

Observation 6d4faef4-48f6-4892-a2c7-dc86704e821e · outbound

This paper cites Efficient training of neural transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Efficient training of neural transducer for speech recognition,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.112682Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.112682Z digest=sha256:f36b4b9190e53b92626dd928e83a386ff04a17cb557204f4835aadf55161be90

Observation 995b2255-cb3e-47ef-98d0-352f5904b3aa · outbound

This paper cites On the limit of english conversational speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems On the limit of english conversational speech recog- nition,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.213383Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.213383Z digest=sha256:6a6217b655297a53f0a4ba81d3624a3f7ab91239c7dd891de05e87fdfd38f7f0

Observation 7146709b-3bbc-4ffe-be3a-99477a6906f5 · outbound

This paper cites Improving the training recipe for a robust conformer-based hybrid model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving the training recipe for a robust conformer-based hybrid model,

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.327114Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.327114Z digest=sha256:cbab20cfb4bf6ecef5cff7ad89a04e60664a4d2191d77ecefe7f4ba59e5f44cb

Observation 9cd39ca8-bd36-4e63-9213-3cfcdf89c110 · outbound

This paper cites Confidence score based conformer speaker adaptation for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Confidence score based conformer speaker adaptation for speech recognition,

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.415984Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.415984Z digest=sha256:46ce3ee2c1f78436eee8d33cdfc8b6d9974ce8381124c66ab105e2d93321ae6e

Observation 00bb68b6-2a8c-4174-b0bf-d35ecab8dbb3 · outbound

This paper cites Diagonal State Space Augmented Transformers for Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Diagonal State Space Augmented Transformers for Speech Recognition

Reference 16

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.243108Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:40.515370Z digest=sha256:ba797bf83551e40e2fcbb7cbda2de745ca75e48168c8b0f190bff3cc05ecee56

Observation 0617d538-d4e0-42ce-b3bd-101d1af0fe27 · outbound

This paper cites Recent advances in end-to-end automatic speech recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Recent advances in end-to-end automatic speech recog- nition,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.601112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.601112Z digest=sha256:ce3d75e9d61b22c7fe7df27aa1442616d88428ad3fde9a6760d1538aa4b2597e

Observation 0bcb61f5-291a-410e-b19c-38ee77418444 · outbound

This paper cites Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Transformer-based Conversational ASR by Inter-Sentential Attention Mechanism

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:51.028105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:40.689037Z digest=sha256:8defbd9eb4f3a1fb0e2f732d93c18dea0a417d1f47aab98e6bdc0f765244b623

Observation ec1cfdea-1210-4095-8423-246400ab3e25 · outbound

This paper cites Improving asr contextual biasing with guided attention,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving asr contextual biasing with guided attention,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.773045Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.773045Z digest=sha256:d51243ebabe5846d41e780d9dc1143ba54e97fb853eba987954ea2183b55fd44

Observation 0d3c780a-92b5-4b2b-8207-abaff055228d · outbound

This paper cites Optimizing byte-level representation for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Optimizing byte-level representation for end-to-end asr,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.851093Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.851093Z digest=sha256:bef201ac69c7bc769f65c31909f240982ea07e050c92e6c9016f82bf18272542

Observation d9907ec3-f172-4699-9634-6a4165bf1ed9 · outbound

This paper cites Contextual modeling for document-level asr error correction,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual modeling for document-level asr error correction,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:40.979654Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:40.979654Z digest=sha256:660609cdf659e41a32139a104502ddc6cde4d0e207f8983b047563a3354f831b

Observation 9c199e58-a953-4b7a-9239-740734c6c521 · outbound

This paper cites Promptasr for contextualized asr with controllable style,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Promptasr for contextualized asr with controllable style,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.065159Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.065159Z digest=sha256:7a9e031b726a3c598f3ef671d3b5f71d16c9c505459af3ccd529de0a9d93841a

Observation 6f881a35-e3a6-4ae2-80d2-cd350f18f212 · outbound

This paper cites Conversational speech recognition by learning audio- textual cross-modal contextual representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conversational speech recognition by learning audio- textual cross-modal contextual representation,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.143119Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.143119Z digest=sha256:e4dce24e8537e7d22bf6898567d2db0fde2e2ab00874ad89c49fdfacfeb480c8

Observation 5efe2905-0318-4095-96cc-99e8f0f248b1 · outbound

This paper cites Dual-mode nam: Effective top-k context injection for end-to-end asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dual-mode nam: Effective top-k context injection for end-to-end asr,

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.229606Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.229606Z digest=sha256:6db312f09bb37afda3c1113c7953e3bef755b04e65b27580bd94702cacd71c1a

Observation 3e8e1dbc-1089-4d79-ab86-e7db061ad785 · outbound

This paper cites CPPF: A contextual and post-processing-free model for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CPPF: A contextual and post-processing-free model for automatic speech recognition

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.766549Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.333464Z digest=sha256:d5a18b64fd09cce213644eb742d9553d4cabf6fd09fcd247b7ecce02882ec6b1

Observation 71a14c27-013a-4ac2-9140-e92a55b71557 · outbound

This paper cites CASA-ASR: Context-Aware Speaker-Attributed ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CASA-ASR: Context-Aware Speaker-Attributed ASR

Reference 26

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.531977Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.404760Z digest=sha256:63330b1574b3c601ea0c8e9eee71aa57e8eb2cc704892835a9cfce0cef21e9e1

Observation 593b9995-e5c3-43dc-8322-202f0bcb9e72 · outbound

This paper cites Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:41.493267Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:41.493267Z digest=sha256:c2a2fc024e9d7805b588be0434890e2ed2785e960fb1a2d3f1ef1219256a41fd

Observation 461d666f-6e32-4906-83d4-166bd3dec097 · outbound

This paper cites Bring dialogue-context into rnn-t for streaming asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bring dialogue-context into rnn-t for streaming asr,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.257605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.598588Z digest=sha256:9ae58391bb4887ebb8941cdab491479ab36c5788d94f2b75b5f3ad5be6527615

Observation 74b8dca5-c583-4fd1-acbd-7016962c5497 · outbound

This paper cites CopyNE: Better Contextual ASR by Copying Named Entities.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems CopyNE: Better Contextual ASR by Copying Named Entities

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.264086Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.707826Z digest=sha256:31a6870eb99a892151dff889ee531db7dc56cabe5792a1c5e221d13b205121e6

Observation 08faf3b5-bde3-4fae-a1fe-1c067bc2b645 · outbound

This paper cites Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving Neural Biasing for Contextual Speech Recognition by Early Context Injection and Text Perturbation

Reference 30

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:50.050215Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.805565Z digest=sha256:241248cb6ab7a510df5082091d6e12c8c4c91820c1813999743e2a9296a3f8fb

Observation 6976842c-c120-407c-8157-c58085e3a12f · outbound

This paper cites Training language models for long-span cross-sentence evaluation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Training language models for long-span cross-sentence evaluation,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:01.036595Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.884033Z digest=sha256:16ee66dbc3f81e0e1c6e9c0e53650197b457ada074f470a2cf01fda5b24069ef

Observation d274690e-fcbd-4c06-aefb-1596b24930c2 · outbound

This paper cites LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems LSTM-LM with Long-Term History for First-Pass Decoding in Conversational Speech Recognition

Reference 32

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.827274Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:41.984032Z digest=sha256:bd1c6338193e67e68d3af916bbe73d62432b336da0636ff1f510b322a1832410

Observation a44c5369-b7b4-4c4f-938f-b250fe31b8f0 · outbound

This paper cites Session-level language modeling for conversational speech,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Session-level language modeling for conversational speech,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.851580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.099735Z digest=sha256:1ea6e1a5dc95d53f70e7c8118ef7bca7c0ea6f08936fac93fc68d824120d3477

Observation 1a018f5b-d065-4fb8-9e11-3bc6c34ad837 · outbound

This paper cites Transformer-xl: attentive language models beyond a fixed-length context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-xl: attentive language models beyond a fixed-length context,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.636395Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.175833Z digest=sha256:77a4346097ae8a49b791a4e6952ee6ff00434b5fbaf5f2d2079616dbed70039b

Observation de503af8-0a5f-4b5b-8e64-72b83282e881 · outbound

This paper cites End-to-end speech recognition on conversations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition on conversations,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:32:00.225580Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.289439Z digest=sha256:9e763605c5596656c0c36f8a1a972dd34ebc57d7d6c3a0c1d3cd74b7ae6c73dc

Observation 20aad9f9-3f85-4409-9fd3-f433b7213289 · outbound

This paper cites Contextualizing ASR lattice rescoring with hybrid pointer network language model,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualizing ASR lattice rescoring with hybrid pointer network language model,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.827745Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.363244Z digest=sha256:494c22c529097f9a0b07bd5ce2a4af86f1d2546026c5cd797762f141ddf34b3e

Observation 893e0851-d763-4751-9ca6-b08b2789e449 · outbound

This paper cites Use of contexts in language model interpolation and adaptation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Use of contexts in language model interpolation and adaptation,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.689635Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.455433Z digest=sha256:472808bfe134415570563a5ba3931ed83161be340b6b62d976e30e497cabb813

Observation 4a711db4-3544-43ce-89b3-b878233fdb68 · outbound

This paper cites Longformer: The Long-Document Transformer.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longformer: The Long-Document Transformer

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.516968Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.516968Z digest=sha256:480f2e4e1595893f2bfffd9cebb8fa19731ad35ead1283dc1198ae8b042edd5b

Observation 29dfea9b-e96d-4063-8f0c-75fe30248ea4 · outbound

This paper cites Transformer language models with lstm-based cross- utterance information representation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer language models with lstm-based cross- utterance information representation,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.523332Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.603756Z digest=sha256:2c4d15c4e57519418bd5d4add3fb98b31e919b9733ba8fa28454f108871ca89e

Observation 9d5d7fc7-5001-4462-9ccb-d1aa6801088e · outbound

This paper cites Ctc-assisted llm-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Ctc-assisted llm-based contextual asr,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.365324Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.670064Z digest=sha256:d8c24dcb686d4c697cdb7d9822de8725b21ebd331f527bf1dab580d38c608e80

Observation 49e04d82-0591-4a2b-bcfa-165cee1248a3 · outbound

This paper cites Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Seed-ASR: Understanding Diverse Speech and Contexts with LLM-based Speech Recognition

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:42.749315Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:42.749315Z digest=sha256:5c3e9aa0ce241c067c8c49057729d317919c7be0f67c9ecb65b0ea0c325fcf20

Observation 40e7f0d3-ce42-4085-9876-cddf36846caf · outbound

This paper cites Contextual asr error handling with llms augmentation for goal-oriented conversational ai,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextual asr error handling with llms augmentation for goal-oriented conversational ai,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.195690Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.859181Z digest=sha256:f3911ee85bcb7b10427888d5ee90b3a97273cbb00c3354ee5bac400934e378be

Observation 091432e8-62b9-48a7-86de-4e49e62bfb26 · outbound

This paper cites Towards asr robust spoken language understanding through in-context learning with word confusion networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards asr robust spoken language understanding through in-context learning with word confusion networks,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:59.005515Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:42.920140Z digest=sha256:6aae34b6a1d1a79669b64a0332e415dfa3bc34dcf7ecb8af0e2529e5e127d944

Observation 14d87e6b-ea9c-43bd-a846-96e9386cd530 · outbound

This paper cites Contextualization of asr with llm using phonetic retrieval- based augmentation,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualization of asr with llm using phonetic retrieval- based augmentation,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.856857Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.000136Z digest=sha256:39b9fafa8f08724e4df06280a1f6bd9209bbe1311044e0feae9041456d82ea2d

Observation fd7b20e2-e5d2-4c48-8491-71acd464f171 · outbound

This paper cites End-to-end speech recognition contextualization with large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems End-to-end speech recognition contextualization with large language models,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.746831Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.104408Z digest=sha256:c826443f1625f673fbdca072991319ea5eb12e6488f15e1e233a228dca2d2906

Observation 91402f0c-236e-4798-b02d-1be9e1a24467 · outbound

This paper cites Improving domain-specific asr with llm-generated contextual descriptions,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving domain-specific asr with llm-generated contextual descriptions,

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.599667Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.219288Z digest=sha256:a806048ea28be7c1e31db1beea119ad7ba89286fdbcbc1bd15f3accbfa0921df

Observation e5436655-70a5-4320-af73-a6360e8e52d1 · outbound

This paper cites MaLa-ASR: Multimedia-Assisted LLM-Based ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems MaLa-ASR: Multimedia-Assisted LLM-Based ASR

Reference 47

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.498400Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.295375Z digest=sha256:3d65c388f95e77b5b6c7c039cfa0de5145ab860f95224a0fa93f6e1d41ecd219

Observation fde25a80-3665-471e-971e-2ca0bdc51efe · outbound

This paper cites Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized speech recognition: rethinking second- pass rescoring with generative large language models,

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.403844Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.373827Z digest=sha256:62c8210afb146429e32bd208fa9c1c756214e93ab43edccda166844503b9b15b

Observation 4eecbe69-a9bd-41d0-9f06-79167b1a6412 · outbound

This paper cites Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving speech recognition with prompt-based contextualized asr and llm-based re-predictor,

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.237641Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.456893Z digest=sha256:0745ea836f960441cda6d6c7de125541ccb8fba2bb0c6909f52bd29f71f7a597

Observation af0de628-c71a-4cc2-94db-f74b2e7f04d3 · outbound

This paper cites Using Large Language Model for End-to-End Chinese ASR and NER.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Using Large Language Model for End-to-End Chinese ASR and NER

Reference 50

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.555162Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.555162Z digest=sha256:41c552b607df2b50fb927a530245e093ed66f4f8d9e81a3ffdbf5052e3bc7b07

Observation cdcaf6aa-a5b6-4586-b7a8-9ab090614695 · outbound

This paper cites Compressive Transformers for Long-Range Sequence Modelling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Compressive Transformers for Long-Range Sequence Modelling

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:43.616971Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:43.616971Z digest=sha256:99e727edaeef7244b0360faf19156c4503650ee147e069f130add81da44fd7bf

Observation 4d49d666-82cd-4600-a11c-daaf11c69055 · outbound

This paper cites Speaker-aware Speech-transformer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker-aware Speech-transformer,

Reference 52

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:58.090493Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.710776Z digest=sha256:2edc82e4867517148858c88db4ff1249be217ca863ee7539123e72312ed1fc97

Observation a151e8a4-7ff0-434a-9d7a-4e0a24c7906e · outbound

This paper cites Transformer ASR with Contextual Block Process- ing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer ASR with Contextual Block Process- ing,

Reference 53

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.946512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.789641Z digest=sha256:f0f1fb2ae3665c3f498a6341ddde784e542ba160e6c08ddfa33a864e77c24f92

Observation c9307482-10d9-4ae7-bc12-dbcab7aa9904 · outbound

This paper cites Transformer-based long-context end-to-end speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-based long-context end-to-end speech recognition

Reference 54

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.806972Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.856177Z digest=sha256:10ed38956337c1de634e08f903a95aa51ebcd278e9f406ee0bcd1ab4f61591b8

Observation 24bdc12b-946d-40df-a2f9-8cf21cc5a672 · outbound

This paper cites Advanced long-context end-to-end speech recognition using context-expanded transformers,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-context end-to-end speech recognition using context-expanded transformers,

Reference 55

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.684237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:43.939035Z digest=sha256:d26988a189f7d8d382e5dc4d010e4dfc746e9555f8c3453909ec07139f8208e9

Observation 89b091be-dc99-4f00-a415-0b62fc89a353 · outbound

This paper cites Context-aware end-to-end asr using self-attentive embedding and tensor fusion,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware end-to-end asr using self-attentive embedding and tensor fusion,

Reference 56

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.527917Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.010039Z digest=sha256:a400fb2edca84d486ea52eac543f2cfd67291d2e0f8a66c0a28699cc70d45999

Observation 97cf2b87-6708-4372-9ff2-d8c13f6f46ce · outbound

This paper cites Longfnt: Long-form speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Longfnt: Long-form speech recognition with factorized neural transducer,

Reference 57

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.374355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.107285Z digest=sha256:36db07b7026f4485c704f61f5c9e9580ebef054560f3cb22ebb4c6eecd60459a

Observation 9bf31ccf-200c-406b-bb7e-aaf0ef13288c · outbound

This paper cites Advanced long-content speech recognition with factorized neural transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Advanced long-content speech recognition with factorized neural transducer,

Reference 58

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.213167Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.172636Z digest=sha256:9d0c3f2c1fb0383c7f6c869770f7c3c663fc448979bf23931f2a94234159ac29

Observation 288a580c-37a1-4974-87da-c5a9a22e0703 · outbound

This paper cites Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Gated Embeddings in End-to-End Speech Recognition for Conversational-Context Fusion

Reference 59

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:49.180696Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.256170Z digest=sha256:c1f6eeb280dda921a01f4d7ed00c356444f24a7311b73879b420fc0fe29ab206

Observation dace3c57-6b23-4716-b1f7-579469cae9fa · outbound

This paper cites Dialog-context aware end-to-end speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Dialog-context aware end-to-end speech recognition,

Reference 60

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:57.039618Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.339765Z digest=sha256:9f94ccf7424d492ac106536804120d8e61d6f4f67d1e0d87408f7f475c853082

Observation 500088fc-abbc-4dc7-970a-27a59f22f618 · outbound

This paper cites Improving RNN-T ASR Accuracy Using Context Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Improving RNN-T ASR Accuracy Using Context Audio

Reference 61

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.854090Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.437262Z digest=sha256:a80ed4e6521b2f4c7533abd8963854bb50e1d0e339db2e08d22bf1a09c116442

Observation 58ce565f-11bf-4f36-af09-e80c3f164fc3 · outbound

This paper cites Large-context automatic speech recognition based on rnn transducer,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Large-context automatic speech recognition based on rnn transducer,

Reference 62

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.894247Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.553967Z digest=sha256:56e4415718cf5210eda2dcc90bfae398ca2dd48874b0967cc760a2f695aa610d

Observation 15b7a962-8d94-426f-92b8-22bd2eff6411 · outbound

This paper cites Phoneme-aware encoding for prefix-tree-based contextual asr,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Phoneme-aware encoding for prefix-tree-based contextual asr,

Reference 63

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.762269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.635612Z digest=sha256:decce6ed9fc627a52bcb898eab180dae92cd3708777e9e63fe1fe454cfed5dcb

Observation 35ac3e39-c498-464b-8d23-00baa8686d08 · outbound

This paper cites Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Deferred NAM: Low-latency Top-K Context Injection via Deferred Context Encoding for Non-Streaming ASR

Reference 64

Resolution
verified exact
local_arxiv, observed 2026-08-05T20:31:48.551140Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.729117Z digest=sha256:1bcd41d549e4dda4b87cf063cda0ac09477d9d5764ff201071d689fa73806db2

Observation 93f9e98f-fa5e-4f4c-a78a-0cc1a17b62a3 · outbound

This paper cites Contextualized end-to-end speech recognition with contextual phrase prediction network,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Contextualized end-to-end speech recognition with contextual phrase prediction network,

Reference 65

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.611543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.809319Z digest=sha256:383a7dfcc4520076ab774f5610b076bda6101ecf65bbf153839fdcbe8eea3e63

Observation 90f4525e-4987-4c89-9d6d-2f4f2e527839 · outbound

This paper cites Semi-autoregressive streaming asr with label context,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Semi-autoregressive streaming asr with label context,

Reference 66

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.459164Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.893812Z digest=sha256:df872e1d15f65e003c91513f38012608c2cb8443833478c2026a6e76b14a93e3

Observation 8e0f3d92-8326-49b5-b0de-56fd29b0fb72 · outbound

This paper cites Context-aware transformer transducer for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Context-aware transformer transducer for speech recognition,

Reference 67

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.309374Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:44.983848Z digest=sha256:d6a355b738e41151b3b8bf42442c89bd25d3a0f662a26741ff14b603859e4208

Observation 10454637-34e0-4bc5-a807-76b2982a8845 · outbound

This paper cites Towards effective and compact contextual representation for conformer transducer speech recognition systems,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Towards effective and compact contextual representation for conformer transducer speech recognition systems,

Reference 68

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.180065Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.047116Z digest=sha256:e1a73f258e466697e2e30f90bedd5b883cf0f533d9133c674ee8482cb794fb97

Observation 3c8d4152-b711-4728-aacd-1b96368900e8 · outbound

This paper cites Wav2vec 2.0: a framework for self-supervised learning of speech representations,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wav2vec 2.0: a framework for self-supervised learning of speech representations,

Reference 69

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:56.044753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.119425Z digest=sha256:173839be2ef4221e7ae00cfcafdabbba9ec832e14c34b06dd855ade61f16b9bc

Observation 5701ae8b-ddff-49f9-80a5-e7d78d2fd852 · outbound

This paper cites Wavlm: large-scale self-supervised pre-training for full stack speech processing,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wavlm: large-scale self-supervised pre-training for full stack speech processing,

Reference 70

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.898704Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.187611Z digest=sha256:9d466235819419a2a1b9aa1fc2307b018cdb6ab2af600623c3ec77958c9f3c3c

Observation 3816819a-e1c4-41b2-91fc-8eda76b5cfc6 · outbound

This paper cites Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hubert: Self-supervised speech representation learn- ing by masked prediction of hidden units,

Reference 71

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.727367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.269084Z digest=sha256:7ab623db446777bfa23a83ab33f559620d918385db3da4d17add44790df840a2

Observation d25de274-8d78-40ef-8a6c-2c55cd28f302 · outbound

This paper cites data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems data2vec: A General Framework for Self-supervised Learning in Speech, Vision and Language,

Reference 72

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.559832Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.364942Z digest=sha256:ad9c04db806064d8f5ec285b87bf8fcd16d0ebdbc4b3c26dcfe3b40db885e247

Observation 1b8b7bd8-9bc8-49d5-8bcf-a14c1f119d77 · outbound

This paper cites XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems XLS-R: Self-supervised Cross-lingual Speech Repre- sentation Learning at Scale,

Reference 73

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.366817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.431578Z digest=sha256:06fe70c69320f5e84d059118ed3cce51bb6b93618f3250c6212134e4ac024b90

Observation 6d1bce3b-7bed-4b95-a086-62e88057544f · outbound

This paper cites Data2vec: A general framework for self-supervised learning in speech, vision and language,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Data2vec: A general framework for self-supervised learning in speech, vision and language,

Reference 74

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.203752Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.508204Z digest=sha256:1a00c6c3b604d1e3988b2c3304091017f78be0b8c9228feeef39889a85e43bbb

Observation 2f732a3c-5b53-4682-ad90-78c29eb96669 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Robust speech recognition via large-scale weak supervision,

Reference 75

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:55.018209Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.576141Z digest=sha256:e6d89fcad500f64406321b4c087cf9204ff0c5315e5986d70532f3132c9c1888

Observation d1bf3b75-9894-4e2f-9d00-f8bb23dc60eb · outbound

This paper cites fairseq: A Fast, Extensible Toolkit for Sequence Modeling.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems fairseq: A Fast, Extensible Toolkit for Sequence Modeling

Reference 76

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.668112Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.668112Z digest=sha256:0970390d993dbe6799754aa9d6783b5919a0a5ff52085d1d07cc47e6c75b0513

Observation e171fbec-8b85-4df6-ac32-35cbfffa24d1 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Zipformer: A faster and better encoder for automatic speech recognition

Reference 77

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.751156Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.751156Z digest=sha256:bbb94c6d80cd8853ceb35f6bdbf3540afeda1fde830993d92f82e42ae48f4212

Observation d7e200e4-be8f-4994-bed1-132c5b6756fa · outbound

This paper cites ESPnet: End-to-End Speech Processing Toolkit,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems ESPnet: End-to-End Speech Processing Toolkit,

Reference 78

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.878240Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:45.863510Z digest=sha256:8313bdc4ee4be238bf7cfb591027c9782cc8841aa028886f51d1d5a74c68a6fe

Observation 0a05ab93-c5b3-4b3c-9891-049835cbe534 · outbound

This paper cites GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems GigaSpeech: An Evolving, Multi-domain ASR Corpus with 10,000 Hours of Transcribed Audio

Reference 79

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:45.946423Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:45.946423Z digest=sha256:23d5fdc8006e118c7ce4c97b1bbc9236c91e5429d1684d8a3ac916f69f692a7e

Observation bf08cde8-13d4-4769-93f2-da483918513b · outbound

This paper cites Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Wenetspeech: A 10000+ hours multi-domain man- darin corpus for speech recognition,

Reference 80

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.678706Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.004011Z digest=sha256:0765e38c0c92b5b63799d3d5d76d564466f4d0d65f3aec3346cf4ec0949bcd81

Observation 41321727-56bf-411a-a94d-3d2ea0de4481 · outbound

This paper cites The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems The natural history of alzheimer’s disease: description of study cohort and accuracy of diagnosis,

Reference 81

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.500000Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.098424Z digest=sha256:f27aa5cb1c8c57728fb3ea5d8622f76d04f96bc0a20a7784a8a82ac82db7d04d

Observation 877244d9-6847-4c06-97dc-e304f44a1a21 · outbound

This paper cites Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker turn aware similarity scoring for diarization of speech-based cognitive assessments,

Reference 82

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.337853Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.188228Z digest=sha256:e8fc8f3921c774b88b80726e03f5566f085cb84255659e12a52201d140d806b3

Observation f96dc9c6-d6f3-4266-81f9-4c88cbab1a89 · outbound

This paper cites Self-supervised asr models and features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Self-supervised asr models and features for dysarthric and elderly speech recognition,

Reference 83

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:54.172237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.265388Z digest=sha256:68db06095f5f713ce12cae479d31f2da234f499bc72e46f8f9917c34b9df3e32

Observation fffb07a8-d62b-4a14-b776-a01fae785d9e · outbound

This paper cites Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Developing real-time streaming transformer transducer for speech recognition on large-scale dataset,

Reference 84

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.979937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.350732Z digest=sha256:cdb298b00fb08e1547f91a1c646b8f1bf36ffa49fb8e0185737bf4a57b00604e

Observation 6a90f7b1-9a5d-412f-bc2f-ffe669f11b70 · outbound

This paper cites Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer transducer: a streamable speech recog- nition model with transformer encoders and RNN-T loss,

Reference 85

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.827935Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.434543Z digest=sha256:c2a876ff618e2ed6df22ab979d01729a62b3a85c4f84893444650382643aab27

Observation 9a4deac2-3a99-4335-b266-c56ae1695911 · outbound

This paper cites Transformer-Transducer: End-to-End Speech Recognition with Self-Attention.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Transformer-Transducer: End-to-End Speech Recognition with Self-Attention

Reference 86

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.546745Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.546745Z digest=sha256:05c3f69a099741302ca4b555d9a17abd01eef9453df35cfb76f08bdb838d8869

Observation 9cd3241e-12b5-4540-9b0d-224aa63924af · outbound

This paper cites Language modeling with gated convolutional networks,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Language modeling with gated convolutional networks,

Reference 87

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.598121Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.629844Z digest=sha256:895188f2b077f6269c2ca5396d301d6e84589aba5eefea89d5ea904290457d0c

Observation ead4386f-fbc7-462c-87c3-4eb7b118fbc1 · outbound

This paper cites Attentive Pooling Networks.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Attentive Pooling Networks

Reference 88

Resolution
unresolved
no resolver link, observed 2026-08-05T20:31:46.700394Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T20:31:46.700394Z digest=sha256:c2c9582e8b93ec4cbe6910a8d08950ae40242a1e2fb1b811140438251f9c8b8d

Observation 0ae1fd88-4d05-4373-bcbd-1ce6a34afdb4 · outbound

This paper cites Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Alzheimer’s Dementia Recognition Through Sponta- neous Speech: The ADReSS Challenge,

Reference 89

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.437465Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.810510Z digest=sha256:91cec4039a2667b7a97938fb01c604b150973125cca3355ffc3d347d74d5ae35

Observation bf71d3d1-869a-41fa-8a5f-4ea6c3ede40d · outbound

This paper cites Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Development of the cuhk elderly speech recognition system for neurocognitive disorder detection using the dementiabank corpus,

Reference 90

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.276825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:46.871585Z digest=sha256:bb076cb96e607f1aa5158b662c1502ba3d388824bf278fc3f73edb056619d27e

Observation 0d376c8b-bcb5-424a-b8d6-1e2a86371335 · outbound

This paper cites Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker Turn Aware Similarity Scoring for Diarization of Speech-Based Cognitive Assessments,

Reference 91

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:53.119919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.037997Z digest=sha256:a91ff83a0b021eacc0264e475540b458b252d2ba69514eabe92e2f8dbe047a2f

Observation d2f7a809-7897-40ff-8f05-8ffa777f06f6 · outbound

This paper cites Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Speaker adaptation using spectro-temporal deep features for dysarthric and elderly speech recognition,

Reference 92

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.961044Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.136299Z digest=sha256:ffe167e40e68310346cf03347a2afe9c8def674c837d5cc949772c4003a52c16

Observation 6b58fa7c-73a7-4d23-bdf8-1f82f5f72748 · outbound

This paper cites Some statistical issues in the comparison of speech recognition algorithms,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Some statistical issues in the comparison of speech recognition algorithms,

Reference 93

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.783057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.141110Z digest=sha256:e8786c8565322d965f43e013cf8b986a30243a5e36b5ff9dad12a1cf9c012168

Observation b9a8482f-8ffd-4bea-ac15-1162bb3c1154 · outbound

This paper cites SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems SpecAugment: A Simple Data Augmentation Method for Automatic Speech Recognition,

Reference 94

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.623835Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.221732Z digest=sha256:c371e07a6d9fdd4cc14f6ed2cabd345dbe20941b6380bbe8e2604fb982d0b1ae

Observation 4c969da5-5255-4ec0-93e0-c7c517cc5419 · outbound

This paper cites Tools for the analysis of benchmark speech recognition tests,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Tools for the analysis of benchmark speech recognition tests,

Reference 95

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.403485Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.380300Z digest=sha256:a733fe06106f1985e1cf2e027832ad3ae681d1ca211f343727488f9a7edd5bf1

Observation 44c0c64e-9953-4c8c-8a7e-eb64a6067c21 · outbound

This paper cites Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Exploiting Cross-domain And Cross-Lingual Ultrasound Tongue Imaging Features For Elderly And Dysarthric Speech Recog- nition,

Reference 96

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.219031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.521344Z digest=sha256:49d54cda6a5609fe165d95a3261aa80dca176d313a9ecf8a9e839b8c5ee44fd3

Observation ff45c5f5-5b0f-4e81-8a1a-ce0dfad35a06 · outbound

This paper cites Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Homogeneous speaker features for on-the-fly dysarthric and elderly speaker adaptation and speech recognition,

Reference 97

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:52.029574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.653179Z digest=sha256:7eaea3adc0b97aafb7f1e77febd795fa88b7a2fd87cd191a37a09b524c4db1d8

Observation bc9f7db8-6cba-4681-9c3c-de2bb60715e9 · outbound

This paper cites Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Bayesian Parametric and Architectural Domain Adap- tation of LF-MMI Trained TDNNs for Elderly and Dysarthric Speech Recognition,

Reference 98

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.813760Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:47.808713Z digest=sha256:98e262f078b6d9bd86d2eee11061fdc5a43d06156937bb1c4e91588d59a8d07d

Observation 65f23f34-7a71-4544-b5e4-7a1fe7f6ee55 · outbound

This paper cites Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Conformer Based Elderly Speech Recognition System for Alzheimer’s Disease Detection,

Reference 99

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.616251Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:48.005043Z digest=sha256:17de68a09ba9c3505e01a0e4ce188fac98c4c10b30a3359945f64a50c2746d10

Observation fc64bde5-cc7b-458a-adf0-b7ea595b99f4 · outbound

This paper cites Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,.

Exploring Cross-Utterance Speech Contexts for Conformer-Transducer Speech Recognition Systems Hyper-parameter Adaptation of Conformer ASR Sys- tems for Elderly and Dysarthric Speech Recognition,

Reference 100

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T20:31:51.418413Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-05T20:31:48.143276Z digest=sha256:b818ee7b8e620e7f02a056f66100caa7ed1688fa33c62213cd4c56c5a4523a53

Pith citing papers

No inbound Pith citation observations are available.