Pith. sign in

Paper Citation Record · LEDGER

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios

As of 10 August 2026, this Paper Citation Record lists 39 of 39 outbound references and 1 inbound Pith citation observation for arXiv:2506.14204.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.14204 v1

Coverage vector

measured 39 of 39 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:25:14.088343Z

measured 40 of 40 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-10T06:31:04.303077+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T00:25:11.153732Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T00:25:14.576014Z

Reference resolution

39 of 39 outbound references displayed

  • verified exact1
  • verified fuzzy31
  • unresolved4
  • parse uncertain0
  • malformed identifier2
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 17609435-dd88-4d87-b18a-80565b52b05d · outbound

This paper cites However, the challenge of recognizing overlapping speech in multi-talker sce- narios remains a critical area of research.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios However, the challenge of recognizing overlapping speech in multi-talker sce- narios remains a critical area of research

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.432925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.104804Z digest=sha256:3966a0f431b87fad076d1839835eeeeaa5917bdf239be3c0d1fed67b4e7fc9c0

Observation 6ac59759-3e08-4313-9f9c-91924f8bbea7 · outbound

This paper cites Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:25:14.796052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.153732Z digest=sha256:2c08d9030ebf2d29d074606bc984d4b983b4ae1f0e3bf807a3b66b0642630a43

Observation dc050b56-313c-4e70-9f5c-4ecdcaed6254 · outbound

This paper cites Avg. ” column. 0L and 0S are 0% overlap conditions with long and short inter-utterance silences. Column “CSS.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Avg. ” column. 0L and 0S are 0% overlap conditions with long and short inter-utterance silences. Column “CSS

Reference 3

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:25:19.413213Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.302499Z digest=sha256:4b3dc5da6a72a74bc5ee59831ae30cddee63eb0152bb11fe883e3869d58f843a

Observation b0ae8375-f4be-4ed1-bb36-82e558f03bfd · outbound

This paper cites First, we leverage speech separated signals by using two channel CSS en- coder in our ASR models.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios First, we leverage speech separated signals by using two channel CSS en- coder in our ASR models

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.394435Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.437698Z digest=sha256:efc23ca6e6bcdaf3bab326f176536e43f187e8c8817ad4f35745a7ba92d7aa2c

Observation 374f03fc-4644-414b-99dd-684c65d1c9aa · outbound

This paper cites Clearly, this model significantly outperforms the CT-tSOT model in row 2 across all scenarios.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Clearly, this model significantly outperforms the CT-tSOT model in row 2 across all scenarios

Reference 5

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T00:25:19.403813Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.371923Z digest=sha256:cb56e95551c442c824a63095994dabbd3097630e8a2d76143e2caf78f717d5f5

Observation 094828e0-e881-4699-ae19-aed8ae57c985 · outbound

This paper cites Continuous speech separation: dataset and analysis,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Continuous speech separation: dataset and analysis,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.334387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.922157Z digest=sha256:33146b3d724d24ff28fdd878942b41e904817348df0bb2e5200573336318ef08

Observation ff11b036-779b-4206-a88c-99cf3cf7c00f · outbound

This paper cites Recent advances in end-to-end automatic speech recogni- tion,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Recent advances in end-to-end automatic speech recogni- tion,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.384937Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.530032Z digest=sha256:39cdb7b34311fd6b87c44501ec54de5b5ece1952f509c24deb1195dc70c816f8

Observation 44ee7f40-c828-435e-9bfb-5a4d6a0507ba · outbound

This paper cites End-to-end speech recognition: A survey,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios End-to-end speech recognition: A survey,

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T00:25:11.612090Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:25:11.612090Z digest=sha256:44b9edfc08b31ceee86308ee773496607db99403888b945392e13bd286d1f3d3

Observation 53b917ff-5d68-4fbf-923e-7e8921d174ff · outbound

This paper cites A streaming on-device end-to-end model surpassing server-side conventional model quality and latency,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios A streaming on-device end-to-end model surpassing server-side conventional model quality and latency,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.369043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.695533Z digest=sha256:e73d68c762b3ae77fe1bfc93f14dea6d520074019e2c106fdc6c7865ef0198ba

Observation 90b34695-2ec8-4f44-9358-952d70846d45 · outbound

This paper cites Developing RNN- T models surpassing high-performance hybrid models with cus- tomization capability,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Developing RNN- T models surpassing high-performance hybrid models with cus- tomization capability,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.358880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.762507Z digest=sha256:2bca00d47b68910ea20792f46c80ba016e512d0c6d1ef491512febd4709c8cf1

Observation 5c7c17db-7681-4857-9843-743497f39d4a · outbound

This paper cites Robust speech recognition via large-scale weak su- pervision,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Robust speech recognition via large-scale weak su- pervision,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.345872Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.837004Z digest=sha256:1ab61dbcf437b3d8e8430ba7ee99e34b1051ac1333fad8ed57023f4e567c4c21

Observation 5dc1da24-8022-4cce-ab54-26ea26df6970 · outbound

This paper cites Streaming multi-talker ASR with token-level serialized output training,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Streaming multi-talker ASR with token-level serialized output training,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:18.351548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.301599Z digest=sha256:879c391e74f62e5286d1932b410192a4ec272211f0567b6755f7dcb11cf17bc0

Observation 58ea0287-25a2-4cd5-814e-4607eae42a51 · outbound

This paper cites Integration of speech separation, diarization, and recognition for multi-speaker meetings: System description, comparison, and analysis,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Integration of speech separation, diarization, and recognition for multi-speaker meetings: System description, comparison, and analysis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.247621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.007478Z digest=sha256:b95017ca63c9b8427199e6ca9d58c2eb653a53380176a21df0989327a936f73e

Observation 7641a0d2-6c85-49a2-999d-c6ebf8d76ca2 · outbound

This paper cites Permutation invari- ant training of deep models for speaker-independent multi-talker speech separation,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Permutation invari- ant training of deep models for speaker-independent multi-talker speech separation,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.163355Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.058115Z digest=sha256:77320ca39c422cdbe3d03c6afd87daaa26a401b235122d8a66e713df59dc1555

Observation b0b0ec89-5b07-4517-97c2-ee39ab7eba34 · outbound

This paper cites Recognizing multi-talker speech with permutation invariant training,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Recognizing multi-talker speech with permutation invariant training,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:18.836259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.138836Z digest=sha256:6e311a63530f4e10dc3738d803c1108e534de6fee1d05a149e49bcf78fe3f442

Observation 7e7bfa55-b636-470b-a7dd-3e883a10abcd · outbound

This paper cites Figure 2:Conformer Transducer with Multi-Talker Cascaded Encoder.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Figure 2:Conformer Transducer with Multi-Talker Cascaded Encoder

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:19.422910Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.219046Z digest=sha256:9a3881b848344de3ec049c344dd380d8f2e15764880dcbbb954d6d826103c08f

Observation e4a1df3e-3b8a-40c8-8934-94486baaf80e · outbound

This paper cites Seri- alized output training for end-to-end overlapped speech recogni- tion,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Seri- alized output training for end-to-end overlapped speech recogni- tion,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T00:25:12.182783Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:25:12.182783Z digest=sha256:45c354fd12b293aaf22277767ffa9906c9f689bf8e51bb1dae25a5ec96b9bb00

Observation a8d149a5-e6bb-4edc-9346-b6fcc1bf9619 · outbound

This paper cites Rec- ognizing overlapped speech in meetings: A multichannel separa- tion approach using neural networks,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Rec- ognizing overlapped speech in meetings: A multichannel separa- tion approach using neural networks,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:18.522978Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.206262Z digest=sha256:a816302dadb140bab97abe11ab9b94301477dda3afb4be38f16802cc89a238ec

Observation 8821ce48-5924-452b-a494-017c7448fae0 · outbound

This paper cites Speech separation with large-scale self-supervised learning,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Speech separation with large-scale self-supervised learning,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:18.207138Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.371398Z digest=sha256:3026ef801b3794bf15f599d5e4400bbcc8badb09784d0f5a75342ea6ee0ac1f0

Observation 41310359-94db-4e23-b846-8f46bb6ac63a · outbound

This paper cites MIMO-Speech: End-to-end multi-channel multi-speaker speech recognition,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios MIMO-Speech: End-to-end multi-channel multi-speaker speech recognition,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:18.017941Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.410546Z digest=sha256:04f38665df9582796bce00d7f26933e7bf74010b3e27f14be30c6bd95f36ff00

Observation 85eaf911-914c-4e36-bb23-f6e0a365c14f · outbound

This paper cites Directional ASR: A new paradigm for E2E multi-speaker speech recognition with source localization,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Directional ASR: A new paradigm for E2E multi-speaker speech recognition with source localization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.836005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.542403Z digest=sha256:61ca8b3ee601b1cfa44673e54dd208bd353cf40eab779c790d150c516ad1fff7

Observation e831f744-112f-4105-a331-d4b0176edca2 · outbound

This paper cites VarArray meets t-SOT: Advancing the state of the art of stream- ing distant conversational speech recognition,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios VarArray meets t-SOT: Advancing the state of the art of stream- ing distant conversational speech recognition,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.651858Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.651288Z digest=sha256:c0d1f53e1cfbbbb9847274e64bc5b5936b57f1951ec7a39d9bfce745985536c5

Observation dd45a35d-d79b-4119-9662-38c566091302 · outbound

This paper cites End-to-end multi-speaker speech recognition with transformer,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios End-to-end multi-speaker speech recognition with transformer,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.507989Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.703582Z digest=sha256:1b891bf231140ff3c40082193dfc04b3b05312d13c52400dd41aba671434cc29

Observation bfd09a9b-d49e-4d32-b0fc-2a0fc3c8789a · outbound

This paper cites Conformer: Convolution-augmented transformer for speech recognition,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Conformer: Convolution-augmented transformer for speech recognition,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.356621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.808464Z digest=sha256:c6d1bd717155c72aaa1960a10dbccb16dedcadbaef144c295e7c63e52916bd0d

Observation bea94d92-8c63-4f02-8bbe-cdf215b9f84e · outbound

This paper cites Developing real-time streaming transformer transducer for speech recognition on large- scale dataset,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Developing real-time streaming transformer transducer for speech recognition on large- scale dataset,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.206613Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:12.888888Z digest=sha256:3ae24ad6060d29987290eda8a727e2db832732d0022bace2e7ff11ab2465adbe

Observation f4629428-c691-4067-815c-3d959b1c824a · outbound

This paper cites End-to-end Continuous Speech Recognition using Attention-based Recurrent NN: First Results.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios End-to-end Continuous Speech Recognition using Attention-based Recurrent NN: First Results

Reference 26

Resolution
unresolved
no resolver link, observed 2026-08-07T00:25:13.014632Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:25:13.014632Z digest=sha256:e2e6c3b169bfa75fe109095cb6522e0539ab0e14719b38938a7feb6324381524

Observation 7a2c18b8-95df-4fc7-b461-51197c454355 · outbound

This paper cites Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Listen, attend and spell: A neural network for large vocabulary conversational speech recognition,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:17.059275Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.071856Z digest=sha256:15fedb43be03cb1ee1093d5e114603fcc94ea889caecd54f35e991728180719d

Observation 9de94e1e-ff3d-4423-8f1d-ff0cd18b7326 · outbound

This paper cites Cascaded encoders for unifying streaming and non-streaming ASR,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Cascaded encoders for unifying streaming and non-streaming ASR,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.876718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.144245Z digest=sha256:2e9deca5870598b280652e43e92a1ef8a5505be7187b6a0e2bdbab36448dab1d

Observation 4d5f32ed-e341-4bcd-ad64-e3fb68ca3c44 · outbound

This paper cites Continuous speech separation with con- former,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Continuous speech separation with con- former,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.694122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.216206Z digest=sha256:da60b34bcadb408c464a26744f03484ff23ec216ef0c3e68506f70a0977d90fb

Observation f92627fb-4000-42c5-9777-7f6a30fdff3a · outbound

This paper cites Joint speaker counting, speech recognition, and speaker identification for overlapped speech of any number of speakers,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Joint speaker counting, speech recognition, and speaker identification for overlapped speech of any number of speakers,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.500799Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.325547Z digest=sha256:0247803fdd0bbd7feea78b344e5bdd0c126f47dc84ec15a5af284dc708d47416

Observation 9541bc28-c7a5-457c-8a7a-1aba80653a23 · outbound

This paper cites Investigation of end-to-end speaker-attributed ASR for continuous multi-talker recordings,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Investigation of end-to-end speaker-attributed ASR for continuous multi-talker recordings,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.341036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.387053Z digest=sha256:25a536c3d433000b8f6274648bd209fff9f968737bb64043077001e24e10ca77

Observation 09f0485f-d3e8-4147-9a37-7c038e489e56 · outbound

This paper cites Joint CTC-attention based end-to-end speech recognition using multi-task learning,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Joint CTC-attention based end-to-end speech recognition using multi-task learning,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.189953Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.457586Z digest=sha256:54a99f7469198e8d919d8c5d9ec95ac1b90df362d397d7add9948674382eb9c0

Observation ced51003-9cca-423f-8e83-2fc40c78685b · outbound

This paper cites High-accuracy and low-latency speech recognition with two-head contextual layer trajectory LSTM model,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios High-accuracy and low-latency speech recognition with two-head contextual layer trajectory LSTM model,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:16.008079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.540695Z digest=sha256:5001cf60df70548e66a1583db475af84ff13b1cf9341d966048b95c1eb774627

Observation 7d6260b2-bf6c-44c3-a4b4-21f3d1b4747d · outbound

This paper cites The AMI meeting corpus: A pre- announcement,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios The AMI meeting corpus: A pre- announcement,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:15.804992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.615086Z digest=sha256:addd05c63d5e13e6abc7f0d9973c2039144610c3717bb23f1e6dfa322664c14b

Observation 11d42dff-65b2-4f51-85a2-71cfa4df692c · outbound

This paper cites The ICSI meeting corpus,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios The ICSI meeting corpus,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:15.576909Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.711094Z digest=sha256:ec3c39484c1f329b3c7f0f81b8b53c56445e81bb31298dd7e6b18eb4036c2312

Observation b0e05795-c22a-42df-8950-ffe751988e63 · outbound

This paper cites The NIST Scoring Toolkit (SCTK),.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios The NIST Scoring Toolkit (SCTK),

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:15.339241Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.795693Z digest=sha256:933a829136288c93001ae4b23184df46198d5dc37d9fc2c70656f53003121851

Observation 6436fc04-8ad3-4c5d-92c3-d361bb3c3fac · outbound

This paper cites WavLM: Large-scale self- supervised pre-training for full stack speech processing,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios WavLM: Large-scale self- supervised pre-training for full stack speech processing,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-08-07T00:25:13.906564Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T00:25:13.906564Z digest=sha256:a1a08350a81f765f79ab1c8ad0e3e621d7aeeaa29c03c622d860ef5ec3f884aa

Observation 0a2e3f95-6411-428b-8f0f-65808b4eae26 · outbound

This paper cites Handling Trade-Offs in Speech Separation with Sparsely-Gated Mixture of Experts.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Handling Trade-Offs in Speech Separation with Sparsely-Gated Mixture of Experts

Reference 38

Resolution
verified exact
local_arxiv, observed 2026-08-07T00:25:14.387880Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:13.997966Z digest=sha256:20e609022ea0c4b4a4ee03d7bdf2b0c1b13f53574d07a0635d089236d769dd01

Observation 520496b5-07e4-401c-adb4-f49cf4309d22 · outbound

This paper cites Improving wideband speech recognition using mixed-bandwidth training data in CD- DNN-HMM,.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Improving wideband speech recognition using mixed-bandwidth training data in CD- DNN-HMM,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T00:25:15.063030Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:14.088343Z digest=sha256:ed402522b80d6c13396bc433aa3f4ce7553a679bc5b959399551426cf3884df4

Pith citing papers

Observation 6ac59759-3e08-4313-9f9c-91924f8bbea7 · inbound

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios cites this paper.

Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios Improving Practical Aspects of End-to-End Multi-Talker Speech Recognition for Online and Offline Scenarios

Reference 2

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T00:25:14.796052Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-10T06:31:04.303077+00:00.

source=pdf_text observed=2026-08-07T00:25:11.153732Z digest=sha256:2c08d9030ebf2d29d074606bc984d4b983b4ae1f0e3bf807a3b66b0642630a43