Pith. sign in

Paper Citation Record · LEDGER

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

As of 5 August 2026, this Paper Citation Record lists 50 of 50 outbound references and 1 inbound Pith citation observation for arXiv:2605.02782.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2605.02782 v1

Coverage vector

measured 50 of 50 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-08T18:06:47.859998Z

measured 51 of 51 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-05T06:32:48.257954+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T16:57:11.349967Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

50 of 50 outbound references displayed

  • verified exact10
  • verified fuzzy34
  • unresolved2
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch4

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48dd6110-208d-40dd-822c-182dc72dcd16 · outbound

This paper cites Attention is All you Need , url =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Attention is All you Need , url =

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.854601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:2c554415136c3d310e12f231cba15dbd45f1f4444c7986ca073873289e38e3c5

Observation abfbf2bf-15ea-4e16-bd20-bb0be9f09530 · outbound

This paper cites Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models

Reference 2

Resolution
metadata mismatch
arxiv_id, observed 2026-05-12T18:57:28.886773Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:75dd8a2114e7b9e12c46d913ba63360cbe1648ee2e120ef7741624e9421f2890

Observation 25f871f0-514e-4a1d-b2f9-725f1bf5a208 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.843242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:59a8d6ad29960caad5d8509e33cea45d1acc7717b37e731293908a0cffaff970

Observation 36730354-62dc-4052-bb4d-7221d63a623e · outbound

This paper cites 2022 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2022 , eprint=

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.846544Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:1cac37b39a57e10c9aed8fae7907a85703f880c2159489a4db05b9aabe569dc9

Observation 382eb0ea-8551-47ea-aca8-ec0f3a8e9594 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.835612Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:db707b92068f56d1b81a0940cec2b329c6745a0b4830182f465613d832eec10c

Observation ef4927ab-c9a5-46fd-925a-068d7e2ed7d7 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.839378Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:62ed8bf9dda55da5e00140a28b18d2c1d48c401c83857644dc7754d931955b19

Observation a8570081-f8b5-49f8-a4c0-77433598c369 · outbound

This paper cites Huang and Kenneth Watkin and Simone Frame , year =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Huang and Kenneth Watkin and Simone Frame , year =

Reference 7

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.429906Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:1b31438fbd8e40dbc4a8b882e522c40773c009830a5c578a7624cb82b093af30

Observation 04cb6bed-f215-41a4-911b-333432346655 · outbound

This paper cites The TORGO database of acoustic and articulatory speech from speakers with dysarthria , volume =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition The TORGO database of acoustic and articulatory speech from speakers with dysarthria , volume =

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.850173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:4c7c147f59d1cbec8b139a2ddab6dff0892b51127026d3670b3752b24adcc8b0

Observation 9eee9828-d0d9-42c1-b5a3-2afd7d957b40 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.859162Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:1b8c3e9110bfacecaec792dab965328a5d34e61bfe17ed464003f6f8cd17f9f4

Observation 00b9c6d1-0fe8-46fc-8ab1-067e2e28de5c · outbound

This paper cites Robust Speech Recognition via Large-Scale Weak Supervision.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Robust Speech Recognition via Large-Scale Weak Supervision

Reference 10

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.438480Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:df8103f1d89736861ef0b82eeaf618eca3fa055e00372fbbdb2b7abd306db185

Observation 1e592c9d-a193-465e-ac09-909bf9770593 · outbound

This paper cites Advances in Neural Information Processing Systems , volume=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Advances in Neural Information Processing Systems , volume=

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.875307Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:4240cbcdb91e09d71ec6abc27173877423fd5e5d846c4246f52ddb15805ea52c

Observation 5f66ad36-d3be-400e-82e4-f5463f295cbe · outbound

This paper cites Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Robust wav2vec 2.0: Analyzing Domain Shift in Self-Supervised Pre-Training

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-09T06:45:44.665734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:0dd5c41903b00ccb125d5a5258c63703d0e0a8511a1276ac8b824cbaa01954c1

Observation cfffa3e2-59ab-4fe8-8c71-e4ff17edcb33 · outbound

This paper cites Qwen2-Audio Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen2-Audio Technical Report

Reference 13

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T02:14:45.520303Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:9ec8de0c4c59fce1c51f2a7d9f42f836489b25a0b813d7048090bb5e6e47354b

Observation cc47f93b-d4b6-46d3-98cc-71af1ef3e089 · outbound

This paper cites Qwen3-Omni Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen3-Omni Technical Report

Reference 14

Resolution
metadata mismatch
arxiv_id, observed 2026-05-11T00:20:38.220353Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:21bcbaed9503a25beef2d7ca5a562aa633d175027a10b338d436a0581f541e04

Observation b09c780d-ef32-4bd3-835d-1fb4239e4ed6 · outbound

This paper cites Qwen3-ASR Technical Report.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Qwen3-ASR Technical Report

Reference 15

Resolution
metadata mismatch
arxiv_id, observed 2026-05-13T13:59:09.287362Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:c4572bb676dca5329171c0daae85bc1b3ad76fa282693bfcda7790e000a4ed78

Observation 1f382364-325a-42ce-b87c-474eeecb3c28 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 16

Resolution
unresolved
raw_fallback, observed 2026-05-26T05:51:46.867064Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:ee01c28b61997f21e1454089c309618f49379ee987e49a2ef9ef9770070c7833

Observation afc141da-f4c4-4e2a-b833-c8411c26ee84 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.870905Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:f2a6aeb1bccf7414f51de0a870e20a62c02a6e31165b02be2e52598941b98e43

Observation 759082b8-cc24-4121-bb31-9339ad7052b4 · outbound

This paper cites 2026 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2026 , howpublished=

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.808050Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:db600f3a2d6838d0a6fd9679623df466d228deff9af0bc6e3d5e19f6dc0894aa

Observation 815abc3d-fa9b-4526-a1cb-69c235784603 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.820714Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:31d2eec0dd3637a042ca285abbaf70d8cb75a8bee3676f3968b6b10d9b8d7877

Observation 9e2681ed-8247-4c3e-9b71-7ab2d1ea266f · outbound

This paper cites MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition MiniCPM-V 4.5: Cooking Efficient MLLMs via Architecture, Data, and Training Recipe

Reference 20

Resolution
verified exact
arxiv_id, observed 2026-05-15T17:07:27.685083Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:d953765c091079cc3486707c13ddccd064bbe48476f2b61e6219c3df24c2ba65

Observation 0606d955-6152-4192-92aa-d46fb9b50092 · outbound

This paper cites 2026 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2026 , eprint=

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.827925Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:5bb580437ac4ef2b0a90abfce5b339acc5a395e64fe164339f2731ee5ed9fcd8

Observation a69bc087-7e50-403c-b22f-74aaddf284c1 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.831903Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:9e53b2659bb822f4f47f2788fd81cdb581f3d5df5b1829f3cc52dfcf45100e50

Observation 17fb27cc-669d-4ee4-8bf8-e2a22b41866e · outbound

This paper cites 2023 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2023 , eprint=

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.800656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:b91a8a4439207163774185ff9b662b7189dd4ec2676d788b4f81d18a050fb56e

Observation 8494eeaa-22d6-4b7f-a522-ee7c528f3cf9 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 24

Resolution
unresolved
raw_fallback, observed 2026-05-26T05:51:46.789546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:a79267d060f920f8c02a19fbcc8f20cba20f9b0fd2133bda76f7006f5a5ba27f

Observation a3f07603-28fe-4459-8fc2-99edfeb344cc · outbound

This paper cites 2020 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2020 , eprint=

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.781921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:f82873436e23cde769eca66d1c6ba36665d447d305fc6f491b5d58aa8cb528cd

Observation 3138e4ef-d2fa-4e0f-affc-18745287b4e0 · outbound

This paper cites 2019 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2019 , eprint=

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.778031Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:0fa4c9df86f01d763efcc75c0124982ade7cac3ae4e9ed187425ca9a315a1214

Observation f819bf0d-ead0-479e-af11-d0ace543000a · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.769333Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:98ea70960b9f35c2d2294e8ad9d00f7e1a4f9ee9f720e93425b906e6fdedc956

Observation 82b25f50-6f04-4bd9-a567-13d95ccd4b42 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.773634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:58075066b4d05e073e1f8d1917b41b8552be7eb4deff35eb5490fcb5c9729494

Observation 47434211-2451-46ad-b537-05a2aad74924 · outbound

This paper cites 2020 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2020 , eprint=

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.785680Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:c73dc45bdf6527510e11571dc7cb1d92b065e05d384e5e7c4e00bfbdf281a454

Observation dae98243-70fb-40e5-a150-56d58d85fa51 · outbound

This paper cites State-transition interpolation and map adaptation for HMM-based dysarthric speech recognition.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition State-transition interpolation and map adaptation for HMM-based dysarthric speech recognition

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.793600Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:611f6448068122783cf99b162ca70be6a47d0b7f15ec1a5fb2e47a4ad0d7c2e2

Observation d8a3375b-1393-40c3-995f-9075ae3be6bd · outbound

This paper cites Estimation of Phoneme-Specific HMM Topologies for the Automatic Recognition of Dysarthric Speech , volume =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Estimation of Phoneme-Specific HMM Topologies for the Automatic Recognition of Dysarthric Speech , volume =

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.797546Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:ec4a42a7b79f190b313dce5fad8eaf74770cd02ad5d62d6b3d179d9634d2d3b8

Observation f9fb7c86-9838-4fd6-9a3e-fac60b4e4ccd · outbound

This paper cites Proceedings of Interspeech 2016 , pages =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Proceedings of Interspeech 2016 , pages =

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.765101Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:4e0554127e7d5c83b93dc15a519d4077029c6ec2533cbe6d0c3337f3698bd388

Observation 461654aa-6da6-4048-92e6-80c647603d53 · outbound

This paper cites Proceedings of Interspeech 2023 , pages =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Proceedings of Interspeech 2023 , pages =

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.804302Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:7767baec5d601b09a7c6286daef34a18c967519922c2569eb3f5ab708806f751

Observation eaf12820-967b-4bd4-b1bc-4aa1b86833ce · outbound

This paper cites Two-stage data augmentation for improved ASR performance for dysarthric speech , journal =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Two-stage data augmentation for improved ASR performance for dysarthric speech , journal =

Reference 34

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.426444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:1425c0111868da40adcbc0ff7fdacc3319a4c7f0f0a7b94916c277e185e33f75

Observation 43211658-e092-48a1-a61e-a4f84dfb50f8 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.824345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:9638332a4814d8c9bdb80841bd1d665d8f69862c40055ee1cc4b77afe1b547e4

Observation cd90e2c1-92c2-4c4b-b8ec-92aea2305c8f · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.879111Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:d893378077cc163c2b52ff66f0f0cf0e0a271673fbc1a1c5cda39dbfe65cbacb

Observation bdbdef6e-27b8-4b95-9824-665614cd2726 · outbound

This paper cites 2024 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , eprint=

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.816987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:268f442349e045173437ebe8ae44359ca28114249dee7785636c0c81828ef5fa

Observation bb9911bf-b8dd-4e10-a4bc-14ad09125cdd · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.812244Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:99df4c118e6365d99051358cff75793e46f79da71aeefe0697f29712e7b5483f

Observation 89fc1373-6409-4701-b7ee-bfb6bc4d0f5d · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.748914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:8470280534a8f9fb3ecc5e2d2327848a5aebf702f10ae252636c7748da16838d

Observation 3521aa13-a46a-4964-af84-f56b5b91586a · outbound

This paper cites Urbina and Peter D.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Urbina and Peter D

Reference 40

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.433304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:2606666a0f499892838e28a89cde8b888062bca66339aa7512e894deec466af6

Observation a0f455ef-ccc3-45eb-9e62-687c2a082f53 · outbound

This paper cites 2025 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , howpublished=

Reference 41

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.416248Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:4c7222a496e2ad94b93ab54ff78ae1648551c84ce42bc5cb2497584ad2b26853

Observation 9e1483b4-ded0-4f9d-9019-0125eab270ec · outbound

This paper cites 2024 , howpublished=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2024 , howpublished=

Reference 42

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.412609Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:4ce1eac5a5a0bd542cd291f01f17024b5295c2012395fa7a44497d3e82338647

Observation b81bdc27-f711-4033-b93e-b126e8981e4f · outbound

This paper cites 2021 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2021 , eprint=

Reference 43

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.753016Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:dcd9c323b72579510ef100d4481c8d095121241792f2b1a8d56c5d027dcc22ec

Observation 0ec50091-831a-430d-b2c4-7fef8c9c57a1 · outbound

This paper cites 2025 , eprint=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2025 , eprint=

Reference 44

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.756634Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:5b3d35b054a4496738f03031e03ed0645a588e9c4069c8b01e690426b0a32171

Observation aa6951c3-96c7-4afd-b77e-7d82356f5c13 · outbound

This paper cites an unresolved cited work.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Unresolved cited work

Reference 45

Resolution
verified exact
arxiv_id, observed 2026-05-08T18:08:53.421100Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:473b98198f8216ecae35d3921770bacd46bfe5f216c7d917a5e687cc1961c764

Observation 1a996b29-27d5-42a3-8a09-cd3a1776eaad · outbound

This paper cites title =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition title =

Reference 46

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.760948Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:861cb1d99c12f12eaf7c63bd3c893b76cf529f11e37d7801ee81c985e910e1a4

Observation 6b7df4a9-0a82-41b1-b131-09227a6b6037 · outbound

This paper cites and Nelson, P.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition and Nelson, P

Reference 47

Resolution
verified exact
doi, observed 2026-05-08T18:08:53.408648Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:de961e3d59f9b01542212cf62e4df3d8dfbe81929658dc213ade18af0881c519

Observation effc897b-06b1-4270-bd68-ecdd91200da2 · outbound

This paper cites Population: The 2012 National Health Interview Survey (NHIS) , author =.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Population: The 2012 National Health Interview Survey (NHIS) , author =

Reference 48

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.745153Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:579584ea815fa41e154a7ab91dc0b9d0bb16a6473645a756b1fcd6c6187a4215

Observation b951dcaf-82a2-4d5c-8d66-6fcbb76bfa96 · outbound

This paper cites Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition , year=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition Variational Low-Rank Adaptation for Personalized Impaired Speech Recognition , year=

Reference 49

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.741151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:a94e6ef862645c5a69fc8d2138b7bb612ca5ea133a4cc43137397ad7381c08ad

Observation b6eb8d56-71e8-43f0-8c4d-c461aa698559 · outbound

This paper cites 2012 , publisher=.

When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition 2012 , publisher=

Reference 50

Resolution
verified fuzzy
raw_fallback, observed 2026-05-26T05:51:46.863170Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-05T06:32:48.257954+00:00.

source=arxiv_source observed=2026-05-08T18:06:47.859998Z digest=sha256:affba7d8e85c9c09dc9e3fc4b9de5c422155e57c99759eb3e40c4fca6f6388a0

Pith citing papers

Observation 475bb423-9c3c-4c9b-aff9-f39783c5f9f6 · inbound

ESCUCHA: A Spanish Speech Benchmark for Heterogeneous Acoustic Conditions cites this paper.

ESCUCHA: A Spanish Speech Benchmark for Heterogeneous Acoustic Conditions When Audio-Language Models Fail to Leverage Multimodal Context for Dysarthric Speech Recognition

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-01T16:57:11.349967Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T16:57:11.349967Z digest=sha256:73314421eaea6f9989b46fbac5afb9027de1c5006da4bb50c8b199fab99bf575