Pith. sign in

Paper Citation Record · LEDGER

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement

As of 18 August 2026, this Paper Citation Record lists 49 of 49 outbound references and 1 inbound Pith citation observation for arXiv:2606.17806.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2606.17806 v1

Coverage vector

measured 49 of 49 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-06-26T23:05:08.894347Z

measured 50 of 50 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-06-26T23:05:08.894347Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-03T23:09:00.904043Z

Reference resolution

49 of 49 outbound references displayed

  • verified exact7
  • verified fuzzy0
  • unresolved40
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 259dec7b-05ee-4cca-acb3-f4c8fcd9a10e · outbound

This paper cites PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T23:09:00.905894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:261ff49956b48bb9e0d1b1fb1db157f55ee092768150565e08b84060d6d62a67

Observation 4577b35a-f850-46ad-907d-c850c7f55ba4 · outbound

This paper cites an unresolved cited work.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Unresolved cited work

Reference 2

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:1889ec0a2766e0fc802066cb86a2d5f8d2a72a79ea01a318a12f219ac5e932be

Observation d2e8ad4e-7196-4f04-ac6c-09cebba5c73c · outbound

This paper cites Datasets The clean speech corpus comprises publicly available data from the DNS5 LibriV ox subset [23], VCTK [24], EARS [25], and LibriSpeech [26].

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Datasets The clean speech corpus comprises publicly available data from the DNS5 LibriV ox subset [23], VCTK [24], EARS [25], and LibriSpeech [26]

Reference 3

Resolution
malformed identifier
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:ac1de5bd57f41ba74adc3df0528d15e738e9c55734c8d69f4b38e7ce1cd4f2a8

Observation fee8362f-f0db-42cf-ad74-8b6147cf5c2e · outbound

This paper cites Our work demonstrates that this approach offers a robust and well-structured alternative to conventional spectral- domain methods.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Our work demonstrates that this approach offers a robust and well-structured alternative to conventional spectral- domain methods

Reference 4

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:9e5051f162e33e43e6809412b917db71723f0532f1bf52559f8f0c51ccd219d4

Observation 94c05127-af73-4338-90fd-5966e3daab26 · outbound

This paper cites 12274221) and the Yangtze River Delta Science and Technology Innovation Community Joint Re- search Project (Grant No.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement 12274221) and the Yangtze River Delta Science and Technology Innovation Community Joint Re- search Project (Grant No

Reference 5

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:274fef2e485581154a2263c1ee405eaad5be362132658b8c7951f9fcdae6e52a

Observation 12175251-4ca9-4688-bdfd-0d97930dfae2 · outbound

This paper cites Generative AI was employed exclusively for minor language editing and polishing to enhance clarity and readabil- ity.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Generative AI was employed exclusively for minor language editing and polishing to enhance clarity and readabil- ity

Reference 6

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:1e9ec5af51abeff279d59a86114862425ba74d633b71fc090e75a15704fe1a45

Observation 87cb9647-ebe9-4f9f-9408-ee60b30dd00e · outbound

This paper cites FlowSE: Efficient and High-Quality Speech Enhancement via Flow Matching,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement FlowSE: Efficient and High-Quality Speech Enhancement via Flow Matching,

Reference 7

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:5eaa0b07780f0a60c05a4baec86d98716fc2e14a782cad3a51d1099ed6481da3

Observation 3c600709-3401-47f8-9701-bf3527d31245 · outbound

This paper cites SenSE: Semantic-Aware High-Fidelity Universal Speech Enhancement.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement SenSE: Semantic-Aware High-Fidelity Universal Speech Enhancement

Reference 8

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:00.912655Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:ec5fa772cabd4e542d8103d117d885837e00dfc2b9340aa602f998a05014a3b8

Observation 7eb99388-ec05-4c9d-95df-610863ce6813 · outbound

This paper cites Efficient Speech Enhancement via Embeddings from Pre-trained Generative Audioencoders,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Efficient Speech Enhancement via Embeddings from Pre-trained Generative Audioencoders,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f4224269839575eaed5f642b052fe47b7ba52289fe36db6d1bd3b5147709c489

Observation f01fc13e-b723-4da6-beca-2b215662cc96 · outbound

This paper cites Speech enhancement and dereverberation with diffusion-based generative models,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Speech enhancement and dereverberation with diffusion-based generative models,

Reference 10

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:8e8bdba3bed44b6da918989de972fbb8eabea56a7664bad44a169b4e4f998ee2

Observation cea15566-4cdb-4f05-b468-0a9c24e61bef · outbound

This paper cites Storm: A diffusion-based stochastic regeneration model for speech en- hancement and dereverberation,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Storm: A diffusion-based stochastic regeneration model for speech en- hancement and dereverberation,

Reference 11

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:fdaabc701cae6b28373f0b89e9bd5e8510e7f6e715808267165aa73789164f86

Observation 62033308-71f4-40ab-9f61-695deda6b4d9 · outbound

This paper cites Selm: Speech enhancement using discrete tokens and language models,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Selm: Speech enhancement using discrete tokens and language models,

Reference 12

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:3c6f5e5d4fb9118355163fe7a78b8ff2275d14aecb83e02a5db22e5f33526fa7

Observation 01307d61-731b-41f5-95c6-46fc3c0fde9f · outbound

This paper cites Genhancer: High-Fidelity Speech Enhancement via Generative Modeling on Discrete Codec Tokens,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Genhancer: High-Fidelity Speech Enhancement via Generative Modeling on Discrete Codec Tokens,

Reference 13

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:5fea290b3ce2858dbc773b7eda4cc39723513449ff0a9092eb42ad227bf57c10

Observation 07ddf42f-fe45-4414-84ac-cb29c134cc43 · outbound

This paper cites PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement PASE: Leveraging the Phonological Prior of WavLM for Low-Hallucination Generative Speech Enhancement

Reference 14

Resolution
verified exact
arxiv_id, observed 2026-08-05T02:51:45.666161Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:30cf1249fd23443bcb9bc0e67b96e6f46754c7672bfbfae3e613e9e668006c47

Observation c144dbde-166a-4ab4-8d2a-28d6f856d7f1 · outbound

This paper cites Rethinking flow and diffusion bridge models for speech enhancement,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Rethinking flow and diffusion bridge models for speech enhancement,

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:09:00.925718Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f1ca53f564574054365141e287859a77fb761da3e84ae70c9eb886ef34dcef89

Observation 7eec4a31-d7a8-40c8-9616-eb3b31a9c6ad · outbound

This paper cites Generative speech foundation model pretraining for high-quality speech extraction and restoration,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Generative speech foundation model pretraining for high-quality speech extraction and restoration,

Reference 16

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:415adcd30d7d794e5e3b411f2f7c318d4a66b005415abd1bf0f73892ccae167a

Observation 02de860e-0402-4405-a533-0feac0096a62 · outbound

This paper cites Empirical distributions of dft- domain speech coefficients based on estimated speech variances,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Empirical distributions of dft- domain speech coefficients based on estimated speech variances,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:1df89e79a9035b25d129f06bfc3a673164cb78aca2612cf5e64bf0e953dca7aa

Observation 3353b5e5-4275-45eb-af59-16ed113714b5 · outbound

This paper cites Layer-wise analysis of a self-supervised speech representation model,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Layer-wise analysis of a self-supervised speech representation model,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:b3da469cc7e9ff41701d555b93be9f4eacbfe438497910e3d77d42fc084db8b0

Observation 0c174b1b-ca10-46a6-b346-ac8560c4b81a · outbound

This paper cites Investigating self-supervised learning for speech enhancement and separation,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Investigating self-supervised learning for speech enhancement and separation,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:ba0e44415ee9e3105dec0c404132b99b73c44dbb1bfc6843451460af38c94ff9

Observation e6978272-4121-4c05-9de6-d4a092c51a06 · outbound

This paper cites Boosting self-supervised embeddings for speech en- hancement,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Boosting self-supervised embeddings for speech en- hancement,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:12da36086a13903273ea1daa26f97f47ce16c26ff8700ca01608121f4988aa13

Observation 9d5ff2ab-6bb9-4a1d-bf49-1ad74b033919 · outbound

This paper cites Wavlm: Large-scale self-supervised pre-training for full stack speech processing,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Wavlm: Large-scale self-supervised pre-training for full stack speech processing,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:45b8ad7b78dac4379f410ae719b69fa9eb9a424994172bccc9e01ead9775af9e

Observation 25b02c04-45a1-4815-9e66-d2409a610c3c · outbound

This paper cites Scalable diffusion models with transform- ers,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Scalable diffusion models with transform- ers,

Reference 22

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:39e4f8cccfd83fce47cd0e97c1a67278fca0fd1d581d43f9615ec7ad258bfb92

Observation af39c964-42e2-489c-9f58-45714be7c78a · outbound

This paper cites F5-tts: A fairytaler that fakes fluent and faithful speech with flow matching,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement F5-tts: A fairytaler that fakes fluent and faithful speech with flow matching,

Reference 23

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:db670947c99add8b47aa1b47ae0794e9cf892f3181c07414ffcd10e53636dbf9

Observation ac060926-3e34-4452-aac7-60dfc31010b9 · outbound

This paper cites Back to Basics: Let Denoising Generative Models Denoise.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Back to Basics: Let Denoising Generative Models Denoise

Reference 24

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:00.923183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:d93515577b26baeebef44884c95d02e3cd285a194fcb79426ed2685a55c04ba4

Observation d3198330-4fbd-4cc6-8afc-792e86e9788b · outbound

This paper cites Flow matching for generative modeling,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Flow matching for generative modeling,

Reference 25

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:ee604a0da36b0f395bd97cc898a5c655e944d69a449100775fd4146f5bd32892

Observation 36c5eea7-1f1d-4987-b0d5-ea62cc4ac12c · outbound

This paper cites Available: https://openreview.net/forum?id= PqvMRDCJT9t.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Available: https://openreview.net/forum?id= PqvMRDCJT9t

Reference 26

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:bb5d9a81e4d28bf1f8390b82b50b8714a64e1559a2711b4512f1bbfed62fc1c8

Observation 8dec23c5-03d7-4633-a10e-3d92de0d2bad · outbound

This paper cites Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Vocos: Closing the gap between time-domain and Fourier-based neural vocoders for high-quality audio synthesis

Reference 27

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:09:00.912005Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:bde0f6ae4ff66bc9e114f7e8a67950466918efeacfd6496531dc042e2364965b

Observation 6616a914-6e72-4476-83ef-58c4d914daa1 · outbound

This paper cites Wavtokenizer: an efficient acoustic discrete codec tokenizer for audio language modeling,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Wavtokenizer: an efficient acoustic discrete codec tokenizer for audio language modeling,

Reference 28

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:4f0d603b069c0b09d22c67bb05c8f92621649f640b9b39f18ef78a05905e08d1

Observation e0efc43b-7d92-46e0-b1b8-17b4d8ec5eb0 · outbound

This paper cites High-fidelity audio compression with improved rvqgan,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement High-fidelity audio compression with improved rvqgan,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f0f88e337415d7187e5fa4de1e97c55e3d9928004a6f65548c0f8122955f988d

Observation 1bee1b0e-b621-4dea-9eed-8eb204d79165 · outbound

This paper cites Icassp 2023 deep noise suppression challenge,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Icassp 2023 deep noise suppression challenge,

Reference 30

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:10fc6f71bfc05ce224ddfb93e01aca67cf2daa12b10d938e9187faed7537d076

Observation 97140995-448b-4c80-8f31-1a818b269d4f · outbound

This paper cites The voice bank corpus: De- sign, collection and data analysis of a large regional accent speech database,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement The voice bank corpus: De- sign, collection and data analysis of a large regional accent speech database,

Reference 31

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:1dd55803f74a0ebb437321db3235b22aa05ec514e911b3b76926ea17ff763c46

Observation 60215fe5-8cf8-4a7f-8ca4-e17c5685df5c · outbound

This paper cites EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dere- verberation,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement EARS: An Anechoic Fullband Speech Dataset Benchmarked for Speech Enhancement and Dere- verberation,

Reference 32

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:cb4f015051c7d0a06e2ec8bf568bed974dad9ff356f9e2f9f1374c53740d86de

Observation 1365b103-3fa2-4e4a-9fb2-87e87cdbf947 · outbound

This paper cites Lib- rispeech: An asr corpus based on public domain audio books,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Lib- rispeech: An asr corpus based on public domain audio books,

Reference 33

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:c47c5c5a29d474caa6f4d6d50a6a9ea7a37d7913fb9621c2fbcd44a062c9c99e

Observation e94d2750-f6e1-42eb-a439-896d90b0b30f · outbound

This paper cites WHAM!: Extending speech separation to noisy environments,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement WHAM!: Extending speech separation to noisy environments,

Reference 34

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:526d813bcd79e96b6c43c5803fffedf631d7925b46fba584229e36e415e4b305

Observation 03c0940a-7a67-4a5d-936b-b06a87de4340 · outbound

This paper cites FSD50K: an open dataset of human-labeled sound events,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement FSD50K: an open dataset of human-labeled sound events,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:782a0cfeccab6f9b9ce8797d29c008db5448f65c44abd97b5a37d9183b571b1f

Observation 708dd38c-60df-4604-a8a5-a31ad1aea4b6 · outbound

This paper cites FMA: A Dataset For Music Analysis.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement FMA: A Dataset For Music Analysis

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:00.928538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:1021e8bbbf89e71c75c930bc7f598f3511f5749f4f31bb2ce739b155261f308f

Observation 3a4bb543-e066-4e09-919a-49d1961f0b0d · outbound

This paper cites A study on data augmentation of reverberant speech for robust speech recognition,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement A study on data augmentation of reverberant speech for robust speech recognition,

Reference 37

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:0f557616ce9fc75e7853a806d2af89764de8f1f2c5bb2ffae4d7ff469252d650

Observation f6c75420-7f9f-4073-a33b-bfc8270a4773 · outbound

This paper cites The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results,

Reference 38

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f827e75260dd7a840fe393bf18e8b8bd4a68e2c36c14eecfb1596c85e2450453

Observation 4b260545-f907-41a2-92dd-a34197d35d45 · outbound

This paper cites Dnsmos p.835: A non-intrusive perceptual objective speech quality metric to eval- uate noise suppressors,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Dnsmos p.835: A non-intrusive perceptual objective speech quality metric to eval- uate noise suppressors,

Reference 39

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f99ce730ed6e381c928b900ad8bee34d4414d04d58325c98e8277aa90e13ddb8

Observation 98664676-f6b0-40f1-92d4-43e97790940a · outbound

This paper cites UTMOS: UTokyo-SaruLab System for V oice- MOS Challenge 2022,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement UTMOS: UTokyo-SaruLab System for V oice- MOS Challenge 2022,

Reference 40

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:36b82a3a96de39972e4edca9c9a7946d459c847960519df3f988e7c4cb186511

Observation 0a46c703-fb82-428b-b55f-abfdfd608cbc · outbound

This paper cites SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement SpeechBERTScore: Reference-Aware Automatic Evaluation of Speech Generation Leveraging NLP Evaluation Metrics,

Reference 41

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:918b6bdc8f0c8304573e94f985fa864c4e605963b64b9b0363cbdac49550887a

Observation 1b5b4ae7-cdaa-4fb6-a1a0-65a863a0e608 · outbound

This paper cites mHuBERT-147: A Compact Multilingual HuBERT Model,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement mHuBERT-147: A Compact Multilingual HuBERT Model,

Reference 42

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:3952f89fc6350904568f630fd22d0dd6524ec9328077d85146b1209a8255a0f9

Observation 72b6552a-2d1a-4b93-ada8-3bc7a5fa41e2 · outbound

This paper cites Evaluation metrics for generative speech en- hancement methods: Issues and perspectives,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Evaluation metrics for generative speech en- hancement methods: Issues and perspectives,

Reference 43

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:f7913f9e87639243b1769ed0ae257158b783d2a4d2bc4b5bf8e400c16ea08fa1

Observation 455794c6-15cf-43dc-bba3-412eafe3a319 · outbound

This paper cites Simple and Effective Zero-shot Cross-lingual Phoneme Recognition,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Simple and Effective Zero-shot Cross-lingual Phoneme Recognition,

Reference 44

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:59a49d8540e8192925243089ae5e0913103aa16e7c39bdac89951f56544fbbe4

Observation d82fda8b-0ae7-4aa4-8271-cf6581724098 · outbound

This paper cites Robust speech recognition via large-scale weak supervision,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Robust speech recognition via large-scale weak supervision,

Reference 45

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:23c392f2e8e3cfe38d3597fc9ca6e94a68d6e3d3d86869d512eff6c03308d77b

Observation 89fe1262-c2ac-41b2-9737-7d2a85f76c9e · outbound

This paper cites Tf-gridnet: Integrating full- and sub-band modeling for speech separation,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Tf-gridnet: Integrating full- and sub-band modeling for speech separation,

Reference 46

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:914c2e3eef0d460ca55d3f40be184cd897d3cec9462642b86aebd1261def9d9f

Observation 204a1e36-8fbc-47a4-a68c-7aa745e0c694 · outbound

This paper cites LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement LLaSE-G1: Incentivizing Generalization Capability for LLaMA-based Speech Enhancement

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-07-03T23:09:00.925960Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:51843466c2eccf96b604d514dbab7538ad55009952fa8f9d5fa37bab5a70aba3

Observation 0aec4b56-5a05-40f4-bd10-1c3346a9d1ce · outbound

This paper cites Anyenhance: A unified generative model with prompt-guidance and self-critic for voice enhancement,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement Anyenhance: A unified generative model with prompt-guidance and self-critic for voice enhancement,

Reference 48

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:5fc9a23b4a2ab95c07677afb91f3732a8423f308bc686f9df10523e3bbb0c7ef

Observation c83168b7-2e0f-46b6-aa60-d63c4f822ffb · outbound

This paper cites A convnet for the 2020s,.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement A convnet for the 2020s,

Reference 49

Resolution
unresolved
no resolver link, observed 2026-06-26T23:05:08.894347Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:8db09be33b46ef105cded8d532b98f503b3b43e322f67b1b1b0507719488a06d

Pith citing papers

Observation 259dec7b-05ee-4cca-acb3-f4c8fcd9a10e · inbound

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement cites this paper.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-07-03T23:09:00.905894Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:261ff49956b48bb9e0d1b1fb1db157f55ee092768150565e08b84060d6d62a67