Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:52:49.886751Z
Paper Citation Record · LEDGER
As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2507.20169.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-15T17:52:49.886751Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-05T12:53:15.251788Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T12:53:15.386672Z
35 of 35 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation d8ba3bc5-c9d3-489c-88c0-baeb12cb0dba · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Language mod- els are few-shot learners,
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19912d0a-9069-4a27-96cd-5287d68f9453 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Conformer: Convolution-augmented Transformer for Speech Recognition
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f04e8eeb-fc82-4505-9780-14b678bd09ea · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Qwen-Audio: Advancing Universal Audio Understanding via Unified Large-Scale Audio-Language Models
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0fdf5d1f-b198-4d7e-9318-0b7fc8e144fa · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech WavLLM: Towards Robust and Adaptive Speech Large Language Model
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 658df609-86f8-45fd-87ae-c90c2f2b8575 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech SALMONN: Towards Generic Hearing Abilities for Large Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 65ca8625-4bc6-453f-9764-de5bb852e6a2 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech SpeechGPT-Gen: Scaling Chain-of-Information Speech Generation
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1219bc16-6922-4ebd-8784-4eb23686b590 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Prompting large language models with speech recognition abilities,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 81147adf-56c8-4d61-b3a5-5ea661c2d6cf · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Noise-robust speech recognition with 10 minutes unparalleled in-domain data,
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5794534e-f06c-42e7-932f-d2f7d5d7aac1 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Large Language Models Can Self-Improve
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9013bd3f-4e44-4b89-bcd1-2904e07f5243 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Feature alignment by uncertainty and self-training for source-free unsupervised domain adaptation,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 2cbc84a5-b221-4e60-a4f3-b4b2d8607589 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Self- critical sequence training for image captioning,
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 34f7504f-351a-4da9-bfb4-f939643c9c2b · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Mutual Enhancement of Large Language and Reinforcement Learning Models through Bi-Directional Feedback Mechanisms: A Planning Case Study
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8094ff38-7222-4227-93b7-cbf51fa21de8 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Opera: Alleviating hallucination in multi- modal large language models via over-trust penalty and retrospection- allocation,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 64a0015b-5c18-4c2e-be9f-28daba75ade6 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Unsupervised domain adaptation by back- propagation,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e10d1aec-24b9-4892-8f1e-a91f7a926993 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech An unsupervised deep domain adaptation approach for robust speech recognition,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 94cdb1b0-2bf3-4418-80bf-8f17281f608f · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Self-Taught Recognizer: Toward Unsupervised Adaptation for Speech Foundation Models
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 44bf5459-0adf-41be-bcfe-26e2ca95ef5d · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Speech Translation with Large Language Models: An Industrial Practice
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 591d0237-da1a-4c95-b6e5-550bf8b901c5 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Zero- shot spoken language understanding via large language models: A preliminary study,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 8c41e6ba-b84d-4f4b-93a4-d03b4893ffbb · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech SpeechVerse: A Large-scale Generalizable Audio Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fb4d206-dbcd-4ba8-91e6-3560f48928d8 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech AudioPaLM: A Large Language Model That Can Speak and Listen
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bafe0a09-d20e-494d-8a8f-e12641a163f3 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a44c946-da3f-405b-ba78-b0482755dbee · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech SpeechGPT: Empowering Large Language Models with Intrinsic Cross-Modal Conversational Abilities
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 577b1a68-00e5-4988-9c31-2ae4a8858268 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Spirit-lm: Interleaved spoken and written language model,
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21895f18-f986-4ed7-aacf-4a66dc443483 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech mSLAM: Massively multilingual joint pre-training for speech and text
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7958f443-af49-4acc-bd5f-abbb6e366599 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Qwen2-Audio Technical Report
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a17b7ed3-37c1-430a-b49d-28a6b760fbdc · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Are sixteen heads really better than one?
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6f3e2be-e38f-4b8c-9fc4-8a3d36fca789 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Deep Inside Convolutional Networks: Visualising Image Classification Models and Saliency Maps
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4d819a0f-2ea7-4f90-bc1a-1309283c94f4 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Minimum word error rate training for attention-based sequence-to-sequence models,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e9f75467-0827-4393-a651-17da6ff18cb6 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech The 4th chime speech separation and recognition challenge,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 76c84e1b-1577-44f2-97b1-c5252b77da6b · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Freesound technical demo,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation e151c0ef-d702-4bcd-b13d-3202f9d82018 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Common Voice: A Massively-Multilingual Speech Corpus
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c10d56c-2816-4b63-aa7c-8b5431bc1f92 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Single headed attention based sequence-to-sequence model for state-of-the-art results on Switchboard
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c2e1cfa8-193a-445b-ab81-46d0b1ddb873 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech CoVoST: A Diverse Multilingual Speech-To-Text Translation Corpus
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 546db52e-72d2-405e-989e-956518b8a9e8 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Librisqa: A novel dataset and framework for spoken question answering with large language models,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation 5efdfbc9-758d-4bb9-8a82-9b38fb63dd20 · outbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech Direct preference optimization: Your language model is secretly a reward model,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.
Observation c6ffad54-5afe-4e4e-ae33-44ea1d23afc8 · inbound
A Multimodal Deep Learning Framework for Early Diagnosis of Liver Cancer via Optimized BiLSTM-AM-VMD Architecture Self-Improvement for Audio Large Language Model using Unlabeled Speech
Reference 98
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.