Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 9 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 27 inbound Pith citation observations for arXiv:2502.17239.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T15:38:37.251024Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 7ac78b78-52a7-4d26-af96-90e2f944f2cf · inbound
DualToken: Towards Unifying Visual Understanding and Generation with Dual Visual Vocabularies Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 93c544ca-6fdf-4cf3-9a36-01f9219931fa · inbound
Kimi-Audio Technical Report Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3fe741c6-d071-41b1-9000-f47f5998c79a · inbound
S2SBench: A Benchmark for Quantifying Intelligence Degradation in Speech-to-Speech Large Language Models Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a861450a-14e7-4ae6-b631-4ebdec81cd59 · inbound
Breaking the Barriers of Text-Hungry and Audio-Deficient AI Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 69
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation be26c935-a7ac-46cc-8640-094e2ff6b65b · inbound
AURA: Agent for Understanding, Reasoning, and Automated Tool Use in Voice-Driven Tasks Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b0168cde-54d0-420f-92bd-c76a6392a6e2 · inbound
XY-Tokenizer: Mitigating the Semantic-Acoustic Conflict in Low-Bitrate Speech Codecs Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c5ec8a30-318c-426e-93a7-7007b19df62e · inbound
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 40b1610e-2ed8-40ec-b53c-be989790466a · inbound
CodecBench: A Comprehensive Benchmark for Acoustic and Semantic Evaluation Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 2019
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6eb69524-e602-42c6-9335-744ef968d22f · inbound
VoxRole: A Comprehensive Benchmark for Evaluating Speech-Based Role-Playing Agents Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6926b130-d033-4f16-badd-5e1c9729d192 · inbound
Towards Building Speech Large Language Models for Multitask Understanding in Low-Resource Languages Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 89d9be4e-f9a5-4f7b-a26b-0bc092ce3677 · inbound
VCB Bench: An Evaluation Benchmark for Audio-Grounded Large Language Model Conversational Agents Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 11e27bd1-6153-4f61-8b74-5684833c4c51 · inbound
MultiAPI Spoof: A Multi-API Dataset and Local-Attention Network for Speech Anti-spoofing Detection Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94e8db46-1223-4201-b70b-90ca2fa41cf8 · inbound
EchoingPixels: Aliasing-Resistant Joint Token Reduction for Audio-Visual LLMs Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d6683298-010e-43a2-9606-620f3e4996e6 · inbound
The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1619f4df-1231-4214-b6e6-ac818e748e8a · inbound
The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ecce569e-f6b6-4ab5-a71e-af81032600c7 · inbound
The Silent Thought: Modeling Internal Cognition in Full-Duplex Spoken Dialogue Models via Latent Reasoning Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ae31ae8d-13e0-4ce0-a579-01c2dac98406 · inbound
TiCo: Time-Controllable Spoken Dialogue Model Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7baf8c86-9613-4ee9-9ec7-cfa1175ab18d · inbound
MiniMind-O Technical Report: An Open Small-Scale Speech-Native Omni Model Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 61026f86-cd55-4732-8fe4-d45b601d3ce4 · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 104
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ec77828-2a15-4d3a-83d3-590e8adb6a76 · inbound
How Should LLMs Listen While Speaking? A Study of User-Stream Routing in Full-Duplex Spoken Dialogue Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6b2082f1-4432-4c22-9f48-8dafc6f09f5d · inbound
A Survey of Large Audio Language Models: Generalization, Trustworthiness, and Outlook Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 128
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation bbbc490e-cd64-4314-b29c-33519ad62833 · inbound
Raon-Speech Technical Report Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 99d7e8bb-82e4-4a41-85f5-5c04d7ac008a · inbound
PolySpeech-100: A Large-Scale Benchmark for Speech Understanding Across 100+ Languages and Dialects Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 181163d2-d0d8-453f-8e9a-015a68694417 · inbound
MOSS-Audio Technical Report Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 54
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4fc400e4-c217-41ba-8839-4e5e8e252bb1 · inbound
FlexiSLM: A Dynamic and Controllable Frame Rate Spoken Language Model Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 218
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation f7baa26c-ec12-481b-9480-c87ea1f3954c · inbound
Efficient Chain-of-Modality Reasoning via Progressive Compression for Spoken Language Models Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e51c39d-defd-4a26-86bf-4cf89f94e98d · inbound
X$^3$-OPD: Distilling Reasoning into Large Audio-Language Models via On-Policy Alignment Baichuan-Audio: A Unified Framework for End-to-End Speech Interaction
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.