Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 17 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 40 inbound Pith citation observations for arXiv:2101.00390.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-15T21:52:53.410512Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-04T08:59:42.926035Z
0 of 0 outbound references displayed
External citation measurements
No source-named external measurement is stored.
No outbound reference observations are available for this paper version.
Observation b5f067a1-5f74-48e5-bc36-15b8f732aaea · inbound
Gemini: A Family of Highly Capable Multimodal Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 114
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation eee5a22d-e625-4e09-90e3-1b350bbe1090 · inbound
WavChat: A Survey of Spoken Dialogue Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 208
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0416c9cf-68ee-4316-8720-57bb6deee458 · inbound
Scaling Speech-Text Pre-training with Synthetic Interleaved Data VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3d464ebc-e6c3-4f03-9102-50a4f111c769 · inbound
CA-SSLR: Condition-Aware Self-Supervised Learning Representation for Generalized Speech Processing VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b3ea748a-b7de-4c56-8c13-5ab7b14e3b4f · inbound
A Survey on Spoken Italian Datasets and Corpora VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 365fdb58-83d0-4756-99da-81ddc14db160 · inbound
When End-to-End is Overkill: Rethinking Cascaded Speech-to-Text Translation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation df8b7a56-ea67-465b-8f34-2edfe29b78dc · inbound
XAttnMark: Learning Robust Audio Watermarking with Cross-Attention VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 8c466509-5f86-4db4-80f3-df8a3dd8b06d · inbound
Synergistic Effects of Knowledge Distillation and Structured Pruning for Self-Supervised Speech Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b4ca6cc9-34e4-439c-b240-26c98d0a3493 · inbound
On the use of Performer and Agent Attention for Spoken Language Identification VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1c453f19-e2a0-4c06-8853-114f6e5ab342 · inbound
Speech to Speech Translation with Translatotron: A State of the Art Review VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b8dc698b-945b-4e0b-bc4d-6653c2e744d6 · inbound
Evaluation of Deep Audio Representations for Hearables VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 444d5cec-fedb-4c44-87c4-6ac9730a923b · inbound
Kimi-Audio Technical Report VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 4dbce808-0594-4b3e-8ae8-6d608a832352 · inbound
Granite-speech: open-source speech-aware LLMs with strong English ASR capabilities VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 339a71f8-d533-4f2b-8879-123ffeacb751 · inbound
Inclusivity of AI Speech in Healthcare: A Decade Look Back VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3cfacb16-af02-45e6-9163-d43e839a6056 · inbound
HPP-Voice: A Large-Scale Evaluation of Speech Embeddings for Multi-Phenotypic Classification VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ee506292-b81c-45b4-8312-f35493fea5df · inbound
EZ-VC: Easy Zero-shot Any-to-Any Voice Conversion VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b12b32fd-de54-4556-817b-b2c42d11dca1 · inbound
From Tens of Hours to Tens of Thousands: Scaling Back-Translation for Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fc9a8ab5-343b-4ba1-ae6c-dedcb0b7b62b · inbound
DuRep: Dual-Mode Speech Representation Learning via ASR-Aware Distillation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62cf6209-dec3-431b-98bc-12855fdd7f48 · inbound
MFLA: Monotonic Finite Look-ahead Attention for Streaming Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 51c5bdbc-2ebc-4811-8a2f-a6b72e0b28e1 · inbound
Unified Semi-Supervised Pipeline for Automatic Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2ae6d9e7-e12c-4cd5-98cb-8ce49c00d530 · inbound
Improved Intelligibility of Dysarthric Speech using Conditional Flow Matching VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9d90560f-7ae2-47ac-a5cf-b99fa37eaee7 · inbound
Instituto de Telecomunica\c{c}\~oes at IWSLT 2025: Aligning Small-Scale Speech and Language Models for Speech-to-Text Learning VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2e9d6ea3-8a1f-4ed8-8865-22e14f8480c2 · inbound
Edge-ASR: Towards Low-Bit Quantization of Automatic Speech Recognition Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3ece12cd-30a7-43f6-b941-35aafc10cc84 · inbound
Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 107
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 6d9537f6-57df-417c-a1d3-c98e3510fb15 · inbound
On Barriers to Archival Audio Processing VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6404bd70-32fa-466e-b654-406b91370c41 · inbound
An approach to measuring the performance of Automatic Speech Recognition (ASR) models in the context of Large Language Model (LLM) powered applications VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bf6f6e49-c096-41c8-86db-03287d382039 · inbound
Group Relative Policy Optimization for Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1d292078-0891-4d23-92c2-4d347df25a3b · inbound
SpeechLLM: Unified Speech and Language Model for Enhanced Multi-Task Understanding in Low Resource Settings VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation caa0a545-4f09-45b7-a604-0ee6e3a1e41a · inbound
From perception to production: how acoustic invariance facilitates articulatory learning in a self-supervised vocal imitation model VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cda344ad-35fd-40b0-8bde-87d3c6790e7e · inbound
StableToken: A Noise-Robust Semantic Speech Tokenizer for Resilient SpeechLLMs VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 72
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 512e650d-05d0-4566-8218-f216c51451e2 · inbound
ParsVoice: A Large-Scale Multi-Speaker Persian Speech Corpus for Text-to-Speech Synthesis VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9f150e90-e3a9-48ec-90b8-7a5747de2ada · inbound
FastSLM: Hierarchical Temporal Abstraction for Efficient Long-Form Speech Adaptation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 21093ffe-41f4-445d-8d76-c11f59339e69 · inbound
A Semi-spontaneous Dutch Speech Dataset for Speech Enhancement and Speech Recognition VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 75
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fa3ec273-1c9f-4c28-a4d4-a5f4b4cada9b · inbound
In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation c1ff7800-3715-43f2-8791-515dfcb4e63f · inbound
VITA-QinYu: Expressive Spoken Language Model for Role-Playing and Singing VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 126
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation d8f15238-338e-4b4a-bcd5-70bc8f4a5c2d · inbound
Raon-OpenTTS: Open Models and Data for Robust Text-to-Speech VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation afe61cb0-722c-4aac-b777-01fe72d86fb1 · inbound
A Unified and Reproducible Experimentation Framework for Speech Understanding VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 50d2ac01-4d92-4ee5-8c71-936d4a543597 · inbound
NaturalFlow: Reducing Disruptive Pauses for Natural Speech Flow in Simultaneous Speech-to-Speech Translation VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 796dd2b9-b815-40d1-9b73-63d8316d7d75 · inbound
Interleaved Speech Language Models Latently Work In Text VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.
Observation 2611f189-4158-47dd-b79a-961d6fd9cc4a · inbound
SimulS2ST-Omni: Data-Efficient Streaming Speech-to-Speech Translation via Explicit Trajectory Supervision VoxPopuli: A Large-Scale Multilingual Speech Corpus for Representation Learning, Semi-Supervised Learning and Interpretation
Reference 169
Source-reported events for the cited work
Unavailable: canonical work link unavailable.