Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:04:44.626476Z
Paper Citation Record · LEDGER
As of 13 August 2026, this Paper Citation Record lists 38 of 38 outbound references and 1 inbound Pith citation observation for arXiv:2411.15685.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-12T14:04:44.626476Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-13T06:32:02.005865+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T14:23:41.773642Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-07T14:23:47.775279Z
38 of 38 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation 272374f9-250f-4829-9815-c453ed1d6647 · outbound
State-Space Large Audio Language Models BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d5f0bb7-7496-45b7-be5a-7523f573eaad · outbound
State-Space Large Audio Language Models Explor- ing the limits of transfer learning with a unified text-to-text transformer,
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 864a61d4-b592-49bb-b4b2-2299d9643a2c · outbound
State-Space Large Audio Language Models Language Models are Few-Shot Learners
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f87d954-3b97-4147-9c20-6e53c97a56b6 · outbound
State-Space Large Audio Language Models Training language models to follow instructions with human feedback,
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation adb16fc8-8405-42da-84ad-213454a4f5f2 · outbound
State-Space Large Audio Language Models OPT: Open Pre-trained Transformer Language Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c472a50-9126-40e1-9686-67e1a22ab773 · outbound
State-Space Large Audio Language Models A Survey of Large Language Models
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1bf8d7be-1b4c-40ca-9bac-75bff6459c25 · outbound
State-Space Large Audio Language Models Pengi: An audio language model for audio tasks,
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 575f365d-278e-45c3-bf71-ebeef7723e4d · outbound
State-Space Large Audio Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2782f4f9-4361-4c18-841b-6148e704c862 · outbound
State-Space Large Audio Language Models Listen, Think, and Understand
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab25b993-9f24-45eb-8ad2-071d055f12f7 · outbound
State-Space Large Audio Language Models Joint audio and speech understanding,
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 63878ce0-6cab-4f08-8cd3-b2f90b15d257 · outbound
State-Space Large Audio Language Models Audiogpt: Understanding and generating speech, music, sound, and talking head,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation c4a2494b-f4ab-415f-a79a-70998b6a3c6c · outbound
State-Space Large Audio Language Models GAMA: A Large Audio-Language Model with Advanced Audio Understanding and Complex Reasoning Abilities
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3be49039-91f5-44ff-a143-203dfa4262f9 · outbound
State-Space Large Audio Language Models Attention is all you need,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe57f6c0-6794-462a-a70f-84cefef2c630 · outbound
State-Space Large Audio Language Models Linformer: Self-Attention with Linear Complexity
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation da4603c1-f7ac-42b9-bfd3-90272c465648 · outbound
State-Space Large Audio Language Models Swin transformer: Hierarchical vision transformer using shifted windows,
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12a3a627-b817-4887-bc50-3e3cf5fb3fa1 · outbound
State-Space Large Audio Language Models Efficiently Modeling Long Sequences with Structured State Spaces
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15d6765f-ea8b-4f66-9863-202b4e879905 · outbound
State-Space Large Audio Language Models Mamba: Linear-Time Sequence Modeling with Selective State Spaces
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6f88bef-02c5-4700-bcdb-670359c23866 · outbound
State-Space Large Audio Language Models VMamba: Visual State Space Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 19792666-4b0c-4e0b-8a24-13f6be1bcc6b · outbound
State-Space Large Audio Language Models DASS: Distilled Audio State Space Models Are Stronger and More Duration-Scalable Learners
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f32d957d-4fcd-49e5-9f4a-290ad292baa6 · outbound
State-Space Large Audio Language Models Long Range Language Modeling via Gated State Spaces
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d9d7d63-f346-40e3-88e0-b3be26b8a6f8 · outbound
State-Space Large Audio Language Models Vision Mamba: Efficient Visual Representation Learning with Bidirectional State Space Model
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 22042d7d-d742-40b4-aa5e-a3139f004b8e · outbound
State-Space Large Audio Language Models Audio mamba: Bidirectional state space model for audio representation learning,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 145a8cfe-1e78-430d-adc2-0ef047f59f1e · outbound
State-Space Large Audio Language Models Audio Mamba: Pretrained Audio State Space Model For Audio Tagging
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eff6762c-db0a-441f-bf51-e294705e2604 · outbound
State-Space Large Audio Language Models SSAMBA: Self-Supervised Audio Representation Learning with Mamba State Space Model
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc80da73-6746-44d4-87b5-ac17c579e499 · outbound
State-Space Large Audio Language Models Instruction tuning for large language models: A survey,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 432ba048-6687-4f1c-8b2c-c24609d87cd2 · outbound
State-Space Large Audio Language Models Hts-at: A hierarchical token-semantic audio transformer for sound classification and detection,
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 623af4e4-d30b-4153-8d37-2fd7dc0db7d5 · outbound
State-Space Large Audio Language Models Language models are unsupervised multitask learners,
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 75116d44-7e9a-4b3b-862a-5c7cf73ea329 · outbound
State-Space Large Audio Language Models Robust speech recognition via large- scale weak supervision,
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c20f9bb-4d80-486e-b72f-37a54b7a28be · outbound
State-Space Large Audio Language Models Conv-tasnet: Surpassing ideal time– frequency magnitude masking for speech separation,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c82c3ff4-178a-45ab-a3cd-dad617685b13 · outbound
State-Space Large Audio Language Models BEATs: Audio Pre-Training with Acoustic Tokenizers
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d096bba1-b40e-4e6f-9f69-2b1d370c861e · outbound
State-Space Large Audio Language Models AST: Audio Spectrogram Transformer
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6c6ccf2-faea-4832-b5e1-2754b0ed7c20 · outbound
State-Space Large Audio Language Models Au- dioclip: Extending clip to image, text and audio,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation 2b3af54f-4da7-43b9-a855-324b50c73b6f · outbound
State-Space Large Audio Language Models Large-scale contrastive language- audio pretraining with feature fusion and keyword-to-caption augmen- tation,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8ea84b7d-1e8a-43bc-acf4-2c35414e883b · outbound
State-Space Large Audio Language Models Clap learning audio concepts from natural language supervision,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9cfac192-3e15-495a-aa66-26643f4f1669 · outbound
State-Space Large Audio Language Models Audio set: An ontology and human-labeled dataset for audio events,
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16d6cc27-6e1d-49a1-959d-c279a508e147 · outbound
State-Space Large Audio Language Models LLaMA: Open and Efficient Foundation Language Models
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0aa2be76-7806-450d-bf01-bce034983a57 · outbound
State-Space Large Audio Language Models Vicuna: An open-source chatbot impressing gpt-4 with 90%* chatgpt quality,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.
Observation fca4f650-c47d-414d-8871-93ec79217e4d · outbound
State-Space Large Audio Language Models LoRA: Low-Rank Adaptation of Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32074b85-7582-4e86-b335-79e93b4ede4e · inbound
Towards Reliable Large Audio Language Model State-Space Large Audio Language Models
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-13T06:32:02.005865+00:00.