Pith. sign in

Paper Citation Record · LEDGER

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary

As of 8 August 2026, this Paper Citation Record lists 47 of 47 outbound references and 2 inbound Pith citation observations for arXiv:2506.09448.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.09448 v1

Coverage vector

measured 47 of 47 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:54:02.600982Z

measured 49 of 49 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:53:56.912097Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T02:49:24.371678Z

Reference resolution

47 of 47 outbound references displayed

  • verified exact2
  • verified fuzzy35
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 59881a6c-55b3-489f-83f4-52933688c0a6 · outbound

This paper cites OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:53:56.912097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:53:56.912097Z digest=sha256:5092905560e137825b58a772250dd21a488ebc46c44ef5242bcbe3d2c768d215

Observation ab025761-f535-4e25-930f-77b30ac6510c · outbound

This paper cites an unresolved cited work.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:54:13.042152Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:56.999349Z digest=sha256:9546f92e45d08d4e2e542e975d849009584548eaaaa7c767180e2814380a782c

Observation 370cefc1-66d0-4671-8372-5328ff287ba1 · outbound

This paper cites all”, “ig.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary all”, “ig

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:12.778202Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.117525Z digest=sha256:4ef96b9fee8243b20eff56b916d62e4e6c6e0605cb85f38f055557edf2852498

Observation a3d24052-a4c2-4205-9a00-a273612b572c · outbound

This paper cites We train the embed- ding and output layers of OWSM with vocabulary sizeKof 5,000.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary We train the embed- ding and output layers of OWSM with vocabulary sizeKof 5,000

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:12.463826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.219539Z digest=sha256:22a9628a4ec3cc45cde8cb44cd22e0cc319668d9b1a1c27e791e7471cce7195c

Observation 3a697e3a-9b89-44c3-a633-5fc9add54175 · outbound

This paper cites the” and “and.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary the” and “and

Reference 5

Resolution
malformed identifier
raw_fallback, observed 2026-08-07T04:54:12.309387Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.322370Z digest=sha256:b0cb7d613a3e7e918e0b6e8c069378c8a989e9b629b0ccbedda2f009292db7df

Observation f86f73b0-7092-45d9-b67f-f7f5cd637c26 · outbound

This paper cites an unresolved cited work.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Unresolved cited work

Reference 6

Resolution
unresolved
raw_fallback, observed 2026-08-07T04:54:12.161965Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.433861Z digest=sha256:42c170ad8c259d32e91c5e38ad8f0057e4e8820d7c58a17e7a8f68ff416955e2

Observation b957b6df-0b3a-4e35-8947-8d3e39650e26 · outbound

This paper cites Ro- bust speech recognition via large-scale weak supervision.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Ro- bust speech recognition via large-scale weak supervision

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:12.039913Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.544368Z digest=sha256:e779922036b7b10cd8ef84cafa46537efdbd527f898c690b60d40d7cd0bf7a40

Observation 235a55a1-b154-43cf-9697-fe726a6a0d7a · outbound

This paper cites Repro- ducing Whisper-style training using an open-source toolkit and publicly available data,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Repro- ducing Whisper-style training using an open-source toolkit and publicly available data,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:11.812242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.673027Z digest=sha256:724aee4c74f36bfd8d937639ff9fc558cdfd9bdcb43a86fe5015136f4fbdc16c

Observation 4f187dca-064d-4cd8-ab66-bb2d60499661 · outbound

This paper cites Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T04:53:57.801970Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:53:57.801970Z digest=sha256:34902a2be6da8ebc4bcd221b4c7e7f1ea0fe3c9d59bf4cfad2fbd645ebf39882

Observation 826a9a5a-a245-4b2f-92e7-ac1fa68efe40 · outbound

This paper cites Scal- ing speech technology to 1,000+ languages,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Scal- ing speech technology to 1,000+ languages,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:11.523089Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:57.953563Z digest=sha256:fc17259864584c9966134ca3245ec34fe09af214f40fe4a6e55935967f290309

Observation febaa2cc-4d75-417e-889f-2ff813c68001 · outbound

This paper cites Less is more: Accurate speech recognition & translation without web-scale data,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Less is more: Accurate speech recognition & translation without web-scale data,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:11.178196Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.111749Z digest=sha256:51666fe518323c57052f92961dceee7ad9f86bf57edb50ad2c42d9766792912d

Observation 6f02bc4a-7529-44ea-86d5-32f0227246d8 · outbound

This paper cites OWSM-CTC: An open encoder-only speech foundation model for speech recog- nition, translation, and language identification,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary OWSM-CTC: An open encoder-only speech foundation model for speech recog- nition, translation, and language identification,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:10.945478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.207904Z digest=sha256:b024870617377b6da756c3c647c40e33e8403dd98c8a15b1e9216a2227c8512c

Observation ff1dc0f1-86b6-4acd-9290-ec937acad892 · outbound

This paper cites OWSM v3.1: Better and faster open Whisper-style speech models based on e- branchformer,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary OWSM v3.1: Better and faster open Whisper-style speech models based on e- branchformer,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:10.671293Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.261774Z digest=sha256:19753868ca2d29123bf24656444020fbf19bad71ac9739e938303728d3dcb1b2

Observation 62cd81ba-b439-44bc-ab9a-d1ebe2345837 · outbound

This paper cites CB-Whisper: Contex- tual biasing Whisper using open-vocabulary keyword-spotting,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary CB-Whisper: Contex- tual biasing Whisper using open-vocabulary keyword-spotting,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:10.364686Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.304434Z digest=sha256:4c9512551d726c7462f2fe662684abe92ffe2c214f41bec052bf3d8452bb0db5

Observation 843c8724-e1e8-4f57-bbae-ecc6ca9e3b9e · outbound

This paper cites Can contex- tual biasing remain effective with Whisper and GPT-2?.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Can contex- tual biasing remain effective with Whisper and GPT-2?

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:10.084534Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.418507Z digest=sha256:f678ea2faeca15492f5b2736c7a4defc1b931ea67a981ab6a2bcc1f61d16984f

Observation 93f0d36d-402a-4e37-913e-6d13d6176877 · outbound

This paper cites Adding user feedback to enhance CB-Whisper,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Adding user feedback to enhance CB-Whisper,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:09.839604Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.545879Z digest=sha256:db48a40aed5b8569654b243b027e65c8532878401a1cd54ef31d6fd39b336573

Observation 380983ad-aa40-40d7-bd5f-24398ac29a06 · outbound

This paper cites Contextual biasing to improve domain- specific custom vocabulary audio transcription without explicit fine-tuning of Whisper model,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextual biasing to improve domain- specific custom vocabulary audio transcription without explicit fine-tuning of Whisper model,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:09.587036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.696074Z digest=sha256:9e9fd0c3e3cf1872263714b7e68a5790e41a404bb2464084d827be9ef07a1cba

Observation 361d0375-8662-4284-a6b8-6da2f19ee3dc · outbound

This paper cites A multitask train- ing approach to enhance Whisper with open-vocabulary keyword spotting,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary A multitask train- ing approach to enhance Whisper with open-vocabulary keyword spotting,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:09.323740Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:58.811210Z digest=sha256:f265b76b5784846693c30dcc5181abd04dc3dbf74afd7e8d9facac951292feeb

Observation dff02807-cf82-40a2-93cf-ae1d78d5bc25 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T04:53:58.965895Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:53:58.965895Z digest=sha256:09ab1d4f6e1f3f6541389e32d4dfe359ef0b3d7b3e10c88cad2a119cc7fe7158

Observation ee4f9c0d-aac5-4dab-be09-7221277096ac · outbound

This paper cites SALMONN: Towards generic hearing abilities for large language models,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary SALMONN: Towards generic hearing abilities for large language models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:09.069261Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.096100Z digest=sha256:8bb79724ca38fa5b95bb70ded635656b8a74bba0a56dbf8cd805ed41ff6d4ec4

Observation 56c9aa8c-2be0-4566-b5f8-4c3940d3a695 · outbound

This paper cites An Embarrassingly Simple Approach for LLM with Strong ASR Capacity.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary An Embarrassingly Simple Approach for LLM with Strong ASR Capacity

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-07T04:53:59.219105Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:53:59.219105Z digest=sha256:57e878e94e58d072486b2d55148b34e5bc392442e85a6d3eb7d7d90eba60d7c1

Observation 44cf86a3-558d-4733-aef5-f4892ad12fe8 · outbound

This paper cites VoiceTextBlender: Augmenting Large Language Models with Speech Capabilities via Single-Stage Joint Speech-Text Supervised Fine-Tuning.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary VoiceTextBlender: Augmenting Large Language Models with Speech Capabilities via Single-Stage Joint Speech-Text Supervised Fine-Tuning

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:54:03.157702Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.331661Z digest=sha256:d6c3c4e1bf809203219db6f898ce53bac47329ff308165ca3e51168dfd1ed058

Observation a7bc907d-74be-42b0-b516-e1d1c7ac1e93 · outbound

This paper cites Contextual biasing speech recognition in speech-enhanced large language model,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextual biasing speech recognition in speech-enhanced large language model,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:08.766360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.442949Z digest=sha256:92c52cd51d1e213ab98a147595f85c5554089af7d7ada592e9878d8bd0536617

Observation df0ac8f8-8eb2-4b95-b881-1dfecf2c26bd · outbound

This paper cites Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Harnessing the Zero-Shot Power of Instruction-Tuned Large Language Model in End-to-End Speech Recognition

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-08-07T04:54:02.883695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.566215Z digest=sha256:58d67209ece19b5a74a9b1da9a2141cd8a9938f4d36caaff9e049730aa1c87b5

Observation 8b6031f6-cf8a-45cc-ad22-f8b11d582077 · outbound

This paper cites Deep context: End-to-end contextual speech recogni- tion,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Deep context: End-to-end contextual speech recogni- tion,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:08.477112Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.721060Z digest=sha256:f3d74c69335d631caf2ea4ac5a629af1cb5dde037fc880be189a919b1d21153b

Observation e94d55c1-8e35-49a9-8f1c-a62c3d694fd2 · outbound

This paper cites Contextual RNN-T for open domain asr,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextual RNN-T for open domain asr,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:08.161085Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.870856Z digest=sha256:7c7e66e545abf91dec3fcdb3f58570d803b7e506d19b178661601bdb9cac41d7

Observation f026e41b-e611-4ccc-9daa-898825d39f90 · outbound

This paper cites Instant one-shot word-learning for context-specific neural sequence-to-sequence speech recognition,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Instant one-shot word-learning for context-specific neural sequence-to-sequence speech recognition,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:07.904348Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:53:59.983951Z digest=sha256:9391f74693a86f5d8df00604068d0a8a35494a4c92a3a7dab1503063ec0d6293

Observation 79c5e661-87dc-4fd3-9dd0-1c01cddb1e0c · outbound

This paper cites Retraining-free customized ASR for enharmonic words based on a named-entity-aware model and phoneme similarity estimation,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Retraining-free customized ASR for enharmonic words based on a named-entity-aware model and phoneme similarity estimation,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:07.689145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.110123Z digest=sha256:2e41db34fe90700b1c7f10bf8cd62dfd23512f747d50b2516bb238f9c8481f70

Observation 6fad5b74-ad37-4e31-8134-c10af9bf3c82 · outbound

This paper cites Contex- tualized End-to-End Speech Recognition with Contextual Phrase Prediction Network,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contex- tualized End-to-End Speech Recognition with Contextual Phrase Prediction Network,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:07.397919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.189159Z digest=sha256:ee9725523b394fc1c428ec12d2848b11e640515fdb6b451dced4c98f237485c7

Observation a25e54b9-02b2-42c0-8376-dcdf0c7a252a · outbound

This paper cites CopyNE: Better Contextual ASR by Copying Named Entities.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary CopyNE: Better Contextual ASR by Copying Named Entities

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-07T04:54:00.318717Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:54:00.318717Z digest=sha256:3b923c11d5139e7f6ae1214d37fd5cff86950bdc0fc0f21ffc1946648379decd

Observation 6129b83c-cf28-481e-859c-77766f5690ab · outbound

This paper cites Lib- rispeech: an ASR corpus based on public domain audio books,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Lib- rispeech: an ASR corpus based on public domain audio books,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:07.121938Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.439366Z digest=sha256:629afb744a885a074f3c3ff0e006025107031904d673abb6937585b310c0b088

Observation fd400bce-61db-4a71-bea1-30b938aec5f9 · outbound

This paper cites Contextualized end-to-end automatic speech recognition with intermediate bias- ing loss,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextualized end-to-end automatic speech recognition with intermediate bias- ing loss,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:06.868816Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.566901Z digest=sha256:7885fdecdd2443b8d47708c8a16ab9be53b166bbbe698cdce045aa7978e7052b

Observation 9e5933a4-62e5-4efa-bf61-9bc2c3e9d4ca · outbound

This paper cites Corpus of spontaneous Japanese: Its design and evaluation,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Corpus of spontaneous Japanese: Its design and evaluation,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:06.575918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.677729Z digest=sha256:9d2171f3cc46698911c835ba7927cdaf79e21ac048e5bb5e8a72bd97e95ab632

Observation bc1cc5f7-7aaf-41d7-b84f-94305fcf2fc8 · outbound

This paper cites Phoneme-aware encoding for prefix-tree-based contextual ASR,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Phoneme-aware encoding for prefix-tree-based contextual ASR,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:06.328980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.835123Z digest=sha256:e628d38cde63ee9607901381eca4479825fa1884c449f4ad73b4e5578e0ac2ca

Observation 67bf9620-fd0a-4c2b-af4d-f73a07c5b631 · outbound

This paper cites Interbiasing: Boost unseen word recognition through biasing intermediate predictions,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Interbiasing: Boost unseen word recognition through biasing intermediate predictions,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:06.033611Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:00.924739Z digest=sha256:e01dcff5176f9ad32f5190022c46b50b51a20a46512cab625b240794f7ef835f

Observation 187c9787-d8a7-4b87-aa0e-9b54b00c8276 · outbound

This paper cites PromptASR for contextualized ASR with controllable style,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary PromptASR for contextualized ASR with controllable style,

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:05.777584Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.044115Z digest=sha256:9d992e8bf80c4b44f8d8fb3d058ba98b2028ff9eb67e6116ec31be45555c2c79

Observation 13922ba9-ed43-418a-8f7b-85dbc1098280 · outbound

This paper cites Contextualized automatic speech recognition with attention- based bias phrase boosted beam search,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextualized automatic speech recognition with attention- based bias phrase boosted beam search,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:05.524736Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.159181Z digest=sha256:442cf35fb99063da469ae01a30490fb9c13cd6aa51e98bd0233e0cb73975c774

Observation 9127bea1-2010-45b8-9a09-a9546f3f13bb · outbound

This paper cites Improving large-scale deep biasing with phoneme features and text-only data in streaming transducer,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Improving large-scale deep biasing with phoneme features and text-only data in streaming transducer,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:05.220461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.251377Z digest=sha256:bc37353d1004ed122956fea06c7b3f26d1755a06776365d0faa4c84285027cee

Observation 5b540885-d7e6-47a5-88a0-ab7c183e5c66 · outbound

This paper cites Contextualized streaming end-to-end speech recognition with trie-based deep biasing and shallow fusion,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextualized streaming end-to-end speech recognition with trie-based deep biasing and shallow fusion,

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:04.996151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.398030Z digest=sha256:942464bb24b3ae631836db63d1f5ff1df488566189d89df2d76543a535a9864e

Observation 58b4ed50-3064-4d4b-8f8e-902c41579286 · outbound

This paper cites Contextualized automatic speech recognition with dynamic vo- cabulary,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Contextualized automatic speech recognition with dynamic vo- cabulary,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:04.694432Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.600693Z digest=sha256:ac181b3ded5670c2f676f67db4ba5d34d8c0a99ff33409fcda4a937b18450ba7

Observation 9256327f-826d-476e-94b3-f7c4eedec0c9 · outbound

This paper cites Hy- brid ctc/attention architecture for end-to-end speech recognition,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Hy- brid ctc/attention architecture for end-to-end speech recognition,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-07T04:54:01.741283Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:54:01.741283Z digest=sha256:a7e292a87c16bed41d71114378e87930dc7605621f9284679713eb6cdc485fc4

Observation d013093b-5684-4afd-b11d-159cf326c5da · outbound

This paper cites Time- synchronous one-pass beam search for parallel online and offline transducers with dynamic block training,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Time- synchronous one-pass beam search for parallel online and offline transducers with dynamic block training,

Reference 42

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:04.448388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:01.877231Z digest=sha256:fb1f5f7f8c391e6445db2ee48833047c0ad71cc0747ff8536aa2ed0ed010ba15

Observation af9f1496-d3b5-44dd-943e-f797cc60d2a4 · outbound

This paper cites Joint beam search integrating ctc, attention, and transducer decoders,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Joint beam search integrating ctc, attention, and transducer decoders,

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:04.152176Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:02.007470Z digest=sha256:6c56a948878716855219fac558b2c7c8ae33d11f1e4ed531599430ae3da470dc

Observation cd5e6539-0cde-4a56-a05d-a37e60ffa3c0 · outbound

This paper cites E-Branchformer: Branchformer with enhanced merging for speech recognition,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary E-Branchformer: Branchformer with enhanced merging for speech recognition,

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:03.930476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:02.141841Z digest=sha256:eb70414b30cc566384ebc936a55144e7f3c3573c302cc9bcd98e40bbfd4dc482

Observation 7e0949b4-2891-443d-9a4a-7c8a1e5c4c43 · outbound

This paper cites Attention is all you need,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Attention is all you need,

Reference 45

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:03.710031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:02.297436Z digest=sha256:1c1faa21ee1fe4f91ea933808cff2b28b2d367854832b1a62a02a77c839bbf46

Observation b281a407-237d-4dab-b7b8-218de9fb78e3 · outbound

This paper cites Adam: A method for stochastic opti- mization,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary Adam: A method for stochastic opti- mization,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T04:54:02.462065Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:54:02.462065Z digest=sha256:937e8ffb1cb2c0a5ca641d3020459facbab9134ac84b8668d43c8916249dd1e0

Observation 56f76f15-2638-4ca3-927f-e1c7d519c1a1 · outbound

This paper cites ESPnet: End-to-end speech processing toolkit,.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary ESPnet: End-to-end speech processing toolkit,

Reference 47

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T04:54:03.491927Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T04:54:02.600982Z digest=sha256:022207dcde051cfa5989bd49978bd18ed789f8fd8c83ec3c2ccfb1f53dffc483

Pith citing papers

Observation 59881a6c-55b3-489f-83f4-52933688c0a6 · inbound

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary cites this paper.

OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T04:53:56.912097Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T04:53:56.912097Z digest=sha256:5092905560e137825b58a772250dd21a488ebc46c44ef5242bcbe3d2c768d215

Observation 30b23701-2ee9-4692-bb78-16ef1438cc55 · inbound

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages cites this paper.

IndicContextEval: A Benchmark for Evaluating Context Utilisation in Audio Large Language Models Across 8 Indic Languages OWSM-Biasing: Contextualizing Open Whisper-Style Speech Models for Automatic Speech Recognition with Dynamic Vocabulary

Reference 29

Resolution
verified exact
arxiv_id, observed 2026-07-04T02:49:24.373156Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-06-26T19:14:26.452851Z digest=sha256:8235c96b27e7cd7c55bb29247e7422bbeeb9e53464404a9908d4535565518b02