Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 11 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 16 inbound Pith citation observations for arXiv:2308.12792.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-11T06:34:44.6726+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T14:59:02.087260Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-05-22T20:45:08.153364Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation d0ea79f4-c8d4-4ab4-843d-bd20c128ea11 · inbound
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System? Sparks of Large Audio Models: A Survey and Outlook
Reference 2008
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3b48b721-800e-443b-ad35-198a71042cd6 · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey Sparks of Large Audio Models: A Survey and Outlook
Reference 217
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8799987d-4c64-45a9-be1c-410a45cd3d9c · inbound
Prepending or Cross-Attention for Speech-to-Text? An Empirical Comparison Sparks of Large Audio Models: A Survey and Outlook
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5586769-663d-4e7a-8a58-853b186df7fa · inbound
From Screens to Scenes: A Survey of Embodied AI in Healthcare Sparks of Large Audio Models: A Survey and Outlook
Reference 135
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 348d5f4f-9348-4e04-bbce-ceabc736624b · inbound
On The Landscape of Spoken Language Models: A Comprehensive Survey Sparks of Large Audio Models: A Survey and Outlook
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 25ba0e21-4489-4a94-aef1-40529f76f3ca · inbound
Probing the Robustness Properties of Neural Speech Codecs Sparks of Large Audio Models: A Survey and Outlook
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b81e8bff-6915-42ea-a251-eb248a06c311 · inbound
XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs Sparks of Large Audio Models: A Survey and Outlook
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0d309b6-211e-4ad6-974a-af2093367958 · inbound
The Man Behind the Sound: Demystifying Audio Private Attribute Profiling via Multimodal Large Language Model Agents Sparks of Large Audio Models: A Survey and Outlook
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 52fed0d3-eb66-4725-b249-4535076df960 · inbound
Multi-TW: Benchmarking Multimodal Models on Traditional Chinese Question Answering in Taiwan Sparks of Large Audio Models: A Survey and Outlook
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e2b93cb-934e-4adc-9cc4-32e48ebf169d · inbound
Via Score to Performance: Efficient Human-Controllable Long Song Generation with Bar-Level Symbolic Notation Sparks of Large Audio Models: A Survey and Outlook
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1e380dfc-05f5-4624-8b86-ba2db684e2f1 · inbound
Game-Time: Evaluating Temporal Dynamics in Spoken Language Models Sparks of Large Audio Models: A Survey and Outlook
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1b5d64a3-e72e-4c99-85af-ec6faa9844d4 · inbound
Hearing to Translate: The Effectiveness of Speech Modality Integration into LLMs Sparks of Large Audio Models: A Survey and Outlook
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 4a84b826-a364-475a-8450-b2583259cb58 · inbound
Generative AI in Signal Processing Education: An Audio Foundation Model Based Approach Sparks of Large Audio Models: A Survey and Outlook
Reference 8
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 1c77c32c-4c42-4db6-93c0-c645629de473 · inbound
Heterogeneity-Aware Dataset Scheduling for Efficient Audio Large Language Model Training Sparks of Large Audio Models: A Survey and Outlook
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation 6c6d05f5-03a7-48d3-b83c-8c7ed597307e · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Sparks of Large Audio Models: A Survey and Outlook
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.
Observation ca5800a4-c99a-4c78-94eb-dd7fa1bf053a · inbound
A Survey of Audio Reasoning in Multimodal Foundation Models Sparks of Large Audio Models: A Survey and Outlook
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-11T06:34:44.6726+00:00.