Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:11:17.649998Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 41 of 41 outbound references and 1 inbound Pith citation observation for arXiv:2608.09288.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:11:17.649998Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-11T20:11:17.344737Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-11T20:11:18.008189Z
41 of 41 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation fc445400-b253-469c-9d90-86a84a66ae2c · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Unresolved cited work
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 25cdecb6-faa7-4f86-8dc5-9fec91c89c82 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a91461eb-53fb-4c65-a36f-5604f41ea4e0 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation As illustrated in Figure 1, the framework comprises three components
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 66e2904c-72a7-44cf-942e-09160a56e2a8 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Each GPU processes a batch of 2 three-second segments
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 75eea7ea-8709-4388-8bb1-1e3ec166f98e · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 83323de6-c4ea-426e-952e-d990f5845e29 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Conv-TasNet: Surpassing ideal time- frequency magnitude masking for speech separation,
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 115e74d8-e6ce-45f5-a227-59b676efcbc3 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Permutation invari- ant training of deep models for speaker-independent multi-talker speech separation,
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e8a1fcb8-2e8a-470c-b1f9-6c4cf83f3fd5 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Deep clus- tering: Discriminative embeddings for segmentation and separa- tion,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 57262555-fa9a-4c87-9042-8815d639a69f · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Ps4: Proxy-supervised joint training for real target speaker extraction,
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6bb8351a-2688-4b3f-a87f-a37375ae6694 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Attention is all you need in speech separation,
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5e71a295-8a89-476d-bb03-6856ba402afb · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Looking to listen at the cock- tail party: A speaker-independent audio-visual model for speech separation,
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 96b59652-e80d-4c7c-8076-0f3dc98bea56 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation The conversation: Deep audio-visual speech enhancement,
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 992753b2-b15a-4bfb-b1e8-9c9bbfe2381a · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation The fifth CHiME speech separation and recogni- tion challenge: Dataset, task and baselines,
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3f8a515a-c3a2-4c16-84e0-966d748376b7 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Time domain audio visual speech separation,
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d71aea48-a37b-45d7-8ae9-f0a815ffa552 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Dual-path RNN: Efficient long sequence modeling for time-domain single-channel speech sepa- ration,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 253453a5-10ff-4d60-a783-d704dc62d06e · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Tiger: Time-frequency interleaved gain extraction and reconstruction for efficient speech separation,
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d2608980-c4e3-496e-b4f6-25dfff961332 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation The AliMeeting corpus: A multi-modal meeting corpus,
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fd809e29-c043-4c39-9c87-f7b69a1723b6 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation MISP: A multi-modal interactive speech process- ing system,
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3550d6bb-89c4-4abc-b4b9-12293e97788f · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation AISHELL-4: An open source dataset for speech separation, diarization and recognition,
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7d840653-aa0e-466d-96f8-e7feaa5005db · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Pyroomacoustics: A Python package for audio room simulation and array processing algorithms,
Reference 20
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 493c7620-d86e-44ec-b7c9-abc3c55dfb05 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation A tutorial on hidden markov models and selected applications in speech recognition,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 17111b39-8046-4a0f-80e2-243022073d20 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation WeSpeaker: A research and production oriented speaker embedding learning toolkit,
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 0091574d-79e4-4869-bf45-747464705955 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation FunASR: A fundamental end-to-end speech recog- nition toolkit,
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 10891722-f7dd-4c05-974e-b8140a655b55 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Room impulse response generator,
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4c61cefb-b562-47f4-adbc-271e34d9e0f3 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation MUSAN: A Music, Speech, and Noise Corpus
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d49fe0c5-3e18-47d9-bfa1-bbf262df550c · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Overcoming catastrophic forgetting in neural networks,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cd12754a-dbc8-4030-b37a-449c3f8bbe28 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation SDR— half-baked or well done?
Reference 27
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 80506bb9-9d1d-4262-9005-0f181893d814 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Out of time: automated lip sync in the wild,
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4621ba01-de89-496c-b84a-3f695fe93e4c · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation An al- gorithm for intelligibility prediction of time-frequency weighted noisy speech,
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 37610fde-13b3-4199-b35a-cf321ca2121e · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Per- ceptual evaluation of speech quality (PESQ)—a new method for speech quality assessment of telephone networks and codecs,
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c499a89d-fa14-4693-a02d-f02d1cfabeca · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation UTMOS: UTokyo-SaruLab system for V oiceMOS challenge 2022,
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 38e85c96-5376-4b67-a3d7-376e9a4f4a18 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation CN-Celeb: A challenging Chinese speaker recognition dataset,
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bdaae359-ff6e-41a6-a41f-d9f99bc321c9 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation LatentSync: Taming Audio-Conditioned Latent Diffusion Models for Lip Sync with SyncNet Supervision
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98347166-25ef-4caa-be48-4d0cea703896 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation DNSMOS: A non-intrusive perceptual objective speech quality metric to evaluate noise sup- pressors,
Reference 34
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation bf281f55-d30e-403c-91c5-025d59aa647a · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Perfect match: Improved cross-modal embeddings for audio-visual synchronisa- tion,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 09a28137-d12b-41e7-b2de-3002d75bf121 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation How far are we from solving the 2D & 3D face alignment problem? (and a dataset of 230,000 3D facial landmarks),
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 756fc988-95c4-4891-bd19-25e58cd4a8ae · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation SEGAN: Speech en- hancement generative adversarial network,
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 2e55f7fb-f54a-4281-959d-9e2b2fe8e520 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation MossFormer2: Combining transformer and RNN-free recurrent network for enhanced time-domain monaural speech separation,
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 37ccb782-15f1-44b6-aa2c-b3cf76884711 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Recommendation ITU- R BS.1770-4: Algorithms to measure audio programme loudness and true-peak audio level,
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 58be2944-db1e-452e-a1ce-174982bea642 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation Adam: A method for stochastic opti- mization,
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3a2c279c-6dfb-463f-8a7b-c2b2b003c788 · outbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation PS4: Proxy-Supervised Joint Training for Real Target Speaker Extraction
Reference 2026
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 25cdecb6-faa7-4f86-8dc5-9fec91c89c82 · inbound
DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation DAVE: A Decoupled Audio-Visual Enhancement Framework for Real-World Speech Separation
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.