Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 19 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 38 inbound Pith citation observations for arXiv:2310.04673.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-19T06:32:44.657259+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-16T06:05:11.370627Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T12:19:49.665222Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation 1bbd4f02-3816-436a-89c4-61a94360561b · inbound
DASB - Discrete Audio and Speech Benchmark LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 6ff3ebd0-44db-421a-9894-2979d38956c8 · inbound
A Comparative Study of Discrete Speech Tokens for Semantic-Related Tasks with Large Language Models LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a316dd0-6089-4e12-b772-fe85c4600f4e · inbound
WavChat: A Survey of Spoken Dialogue Models LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fedfc8f9-0c10-4bc1-80ab-eae0ffdfc593 · inbound
AlignFormer: Modality Matching Can Achieve Better Zero-shot Instruction-Following Speech-LLM LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4174d551-b364-4fdd-b9c9-6ac3e7c1a9f5 · inbound
Comprehensive Audio Query Handling System with Integrated Expert Models and Contextual Understanding LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee79a058-2e8c-4acf-9ec8-70da2858fa0a · inbound
TouchTTS: An Embarrassingly Simple TTS Framework that Everyone Can Touch LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2bc23168-9306-4c21-9b9c-417df8767ec7 · inbound
YingSound: Video-Guided Sound Effects Generation with Multi-modal Chain-of-Thought Controls LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa87ad16-16f0-4646-883c-dcc999720988 · inbound
SilVar: Speech Driven Multimodal Model for Reasoning Visual Question Answering and Object Localization LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5a9409a4-187f-48ff-9a6d-3cffe028e505 · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 102
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 918f8e07-742d-4f2e-902c-5b6441582155 · inbound
CodecFake+: Codec-Based Resynthesized Data as a Proxy for Detecting CodecFake Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 66
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ce8b6928-7725-4f9a-ad72-f4fcab158215 · inbound
A Non-autoregressive Model for Joint STT and TTS LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9377f019-0df1-4fc9-84dd-54616c730186 · inbound
Muyan-TTS: A Trainable Text-to-Speech Model Optimized for Podcast Scenarios with a $50K Budget LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5e1c955-edc6-462f-bfaf-0e48780e59b4 · inbound
Contextual Paralinguistic Data Creation for Multi-Modal Speech-LLM: Data Condensation and Spoken QA Generation LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 454251e3-9fde-4541-94bc-08c0f9dd0653 · inbound
Improving Noise Robustness of LLM-based Zero-shot TTS via Discrete Acoustic Token Denoising LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6afe8582-3ffc-4312-992a-edafba17ca50 · inbound
Towards Reliable Large Audio Language Model LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6279db10-5903-4d71-adf5-822e051483e3 · inbound
VoiceStar: Robust Zero-Shot Autoregressive TTS with Duration Control and Extrapolation LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5ffb8d62-ec18-4574-8387-0ea3b3aaf67b · inbound
Leveraging LLM for Stuttering Speech: A Unified Architecture Bridging Recognition and Event Detection LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ed51a5e-b3e9-4b28-9310-5c372fc528a5 · inbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 76296d15-6b1d-4f77-a815-ea0fe44a5d52 · inbound
A Variational Framework for Improving Naturalness in Generative Spoken Language Models LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1dc0d152-f8f7-4e86-bc7e-a8fbaa3b0b19 · inbound
BoSS: Beyond-Semantic Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bafe0a09-d20e-494d-8a8f-e12641a163f3 · inbound
Self-Improvement for Audio Large Language Model using Unlabeled Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e60264e3-2f83-4ec8-8afa-5a7ab94d894e · inbound
MECAT: A Multi-Experts Constructed Benchmark for Fine-Grained Audio Understanding Tasks LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation f7bf4d08-aa16-43e5-963e-0f2bc451e6a3 · inbound
LLaSO: A Foundational Framework for Reproducible Research in Large Language and Speech Model LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c6a3ded-de9b-4d0d-8255-efc9e2389adb · inbound
Speech Discrete Tokens or Continuous Features? A Comparative Analysis for Spoken Language Understanding in SpeechLLMs LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f7c8791c-cb2f-4c63-b8b0-0cad294cf213 · inbound
Group Relative Policy Optimization for Speech Recognition LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4ac6042f-557b-4108-91ec-10fe944b8558 · inbound
FireRedChat: A Pluggable, Full-Duplex Voice Interaction System with Cascaded and Semi-Cascaded Implementations LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ed13a00-3143-4aea-8a16-0418fe098b38 · inbound
CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation ccb6452f-930e-411b-940d-9326f1634eef · inbound
CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 16ccd22c-94b2-4b0e-b467-4436db4b9cf0 · inbound
UniVoice: Unifying Autoregressive ASR and Flow-Matching based TTS with Large Language Models LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3bee28f1-6bc6-4620-9ba0-da05ee3d0f15 · inbound
Discriminative-Generative Target Speaker Extraction with Decoder-Only Language Models LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 77
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 4c30ce28-6c63-4bcb-afdb-7e8437470196 · inbound
StarTSE: Towards Streaming Target Speaker Extraction via Chunk-wise Interleaved Splicing of Autoregressive Language Model LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 0b217a86-fdfc-4ba8-9606-2efad0dde764 · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation d48d1aa7-664a-4398-a2e2-60701b5e7f1d · inbound
PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 645583c1-f047-49f7-a436-2bf7b10d3686 · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 85f3c95e-ff73-4f5e-84ec-0de7f1e6f0c3 · inbound
From Prompts to Context: An Ontology-Driven Framework for Human-Generative AI Collaboration LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 99f2961a-f65e-42c2-aab2-68a017066e01 · inbound
IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 744ba060-03ab-4eb7-9ab3-ecaf55f849e7 · inbound
FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-19T06:32:44.657259+00:00.
Observation 7256d16a-5279-4882-a7d4-39165b006d82 · inbound
FlowTTS-GRPO: Online Reinforcement Learning with Multi-Objective Reward Optimization for Flow-Matching Based Text-to-Speech LauraGPT: Listen, Attend, Understand, and Regenerate Audio with GPT
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.