Pith. sign in

Paper Citation Record · LEDGER

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

As of 9 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 2 inbound Pith citation observations for arXiv:2506.02499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02499 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:10.917974Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:08.307554Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T10:01:29.067292Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · outbound

This paper cites DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation 477113ed-8c3a-4570-9f59-53ac54726830 · outbound

This paper cites an unresolved cited work.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:26:13.691594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.397233Z digest=sha256:23733309a6d8a5969b21446fce929fab26e51f6fdac3948fa65adf0620b463c3

Observation 240f02a4-00a8-467e-85ed-205cb400a7a6 · outbound

This paper cites Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.633367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.494147Z digest=sha256:a89195935dbc510e51b38cae73f75d2a02e763c750ad0fa8911cb274e352a825

Observation 863e12a4-1327-4f5c-9934-ab957207a5b1 · outbound

This paper cites Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.612512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.559016Z digest=sha256:cd8a3b0b29ad84b3a97f40f6452aeebf984d01735ef4462470fc73101b9d4c60

Observation 3b355aea-c3fd-4d79-9704-b58a086f8e51 · outbound

This paper cites To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.596545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.675340Z digest=sha256:d7bcbf92a0a8a4d95623fd74c0f03d989da32e463136ddcff23d68fe625b46a0

Observation 5260f1eb-e6ad-44bd-9c9d-4ddd636f6c02 · outbound

This paper cites The sound demixing challenge 2023-cinematic demixing track,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The sound demixing challenge 2023-cinematic demixing track,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.579211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.754633Z digest=sha256:5ca91d08894b371b9a912ba285a60e4d090aac0eec59262fe5e18a3a18bf83d4

Observation 853dc9ba-605b-4674-a8fc-b52f466a530d · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Supervised speech separation based on deep learning: An overview,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.561698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.834619Z digest=sha256:d0080c723387033ca4a4dd4884cae1975bbfb1ccb20a2d9b8f0b8044332c6122

Observation a6eadb68-fd61-42e2-9745-bdf983223694 · outbound

This paper cites Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.542876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:08.966726Z digest=sha256:19e64904875f287e32cf81fafc69548552d9a20c5bd1f007d426a6e13de70dc0

Observation bbb4d0b9-de4e-40fc-88e8-c020c634f3c7 · outbound

This paper cites TF-GridNet: Integrating full- and sub-band modeling for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds TF-GridNet: Integrating full- and sub-band modeling for speech separation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.054261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.054261Z digest=sha256:7d024604d781476be354863fb762789b1282a4b3838775607b05a768dbb2ceca

Observation c314b9b3-8fe4-4a57-b994-a22666d5099c · outbound

This paper cites The 2018 signal separation eval- uation campaign,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The 2018 signal separation eval- uation campaign,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.513919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.130709Z digest=sha256:fe1e8e336dc4752d51ef870ec9026ad72b196c3e8e50642ee164293678925334

Observation 08e9fa2b-41cb-4588-8611-2d27d872c9a3 · outbound

This paper cites Open- Unmix - a reference implementation for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Open- Unmix - a reference implementation for music source separation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.492320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.206282Z digest=sha256:ffafc92d4019ff7a2b75a24bceb80e3c0ad9b0c811a2df6a83e08a9c1282b671

Observation 94ae7cea-4d47-4906-bd38-e3e254de2160 · outbound

This paper cites Hybrid transformers for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hybrid transformers for music source separation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.470963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.309694Z digest=sha256:0ab8cf75389de09c80c87e4b46748f66edca3baa201d6b9adb5d7484e886bd64

Observation 452d209e-b91f-4382-aa50-02423d4e3b70 · outbound

This paper cites Universal sound separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Universal sound separation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.451787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.391215Z digest=sha256:7f9f923a4114b6fd6d3476f82e6289c3bd7f45ae7f93cb3e879b021a9ab77fbe

Observation 8841dde0-c76c-43ad-a949-2bfb315154a6 · outbound

This paper cites The cocktail fork problem: Three-stem audio separation for real- world soundtracks,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The cocktail fork problem: Three-stem audio separation for real- world soundtracks,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.432325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.462321Z digest=sha256:0df944d1fb8e0f9ed877795db8c0af2cf2525ed4cbf2d78a1bafa03a85b2e39b

Observation 21b927c6-b541-4d7e-a111-79d74a476add · outbound

This paper cites A general- ized bandsplit neural network for cinematic audio source separa- tion,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds A general- ized bandsplit neural network for cinematic audio source separa- tion,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.413048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.546046Z digest=sha256:3c516a920f7a822fe24e72422b9869ece7f42abecd592277491755a1f849741d

Observation b136c3a8-04ec-4a30-ac97-fbd73ea42424 · outbound

This paper cites Music source separation with band-split RNN,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Music source separation with band-split RNN,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.392521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.655749Z digest=sha256:69361e9c25d2b55d01ae0b17ab469dfec4a18725af9de04896945244180e02e5

Observation ec275346-37c0-4174-9872-8a9592af11af · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.369763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.733574Z digest=sha256:db926393e9471e9e8846d9bf1ed33028d5c47c539c0b2705052c17238a56d6f6

Observation 9b6ec578-7f61-4172-9148-91bed59b0c0b · outbound

This paper cites FMA: A dataset for music analysis,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FMA: A dataset for music analysis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.322807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:09.819652Z digest=sha256:19f93ba757fd4446bf0a3c756fea655df32babc0db15f442b6800a23cbd3313e

Observation 67de3525-a519-43b2-88af-a073a98efd70 · outbound

This paper cites FSD50K: An Open Dataset of Human-Labeled Sound Events.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.913604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.913604Z digest=sha256:6af201b0abce6b21551ca1f4da8489825325bb8de8cfd1927b9708d0045fc913

Observation 188cfe9e-2344-4090-b6b5-6b75da1901db · outbound

This paper cites Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.168035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.009245Z digest=sha256:f4f58ad605d7d42b56c30c956d017889c34dcac052abe2cd192a89c670212780

Observation 7fff145e-f462-4997-a7c8-7e4cb70e19e4 · outbound

This paper cites Hy- perbolic audio source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hy- perbolic audio source separation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.918502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.138699Z digest=sha256:f953907772064f93a8dea7cafc4e9960422953e4c879286b09eedc8c94ca99ce

Observation e4b34b3d-ccec-44a3-a532-c59dc6312802 · outbound

This paper cites PodcastMix: A dataset for separating music and speech in podcasts,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds PodcastMix: A dataset for separating music and speech in podcasts,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.830047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.209784Z digest=sha256:09756eab0eaac9ad58df84f5f575d37ff5ee4ba305003949430bb2907f9cc2e8

Observation bc6ca3bc-edb9-4db9-81d4-183c49c8bdc7 · outbound

This paper cites Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:26:11.116952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.280208Z digest=sha256:af2e037bf290f5c4be229a00af9e4abab6216a5a4ac882fff74ee880b3ff3640

Observation cfd91ca2-5a4d-4c17-a42d-6061da0cf752 · outbound

This paper cites CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.496001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.372449Z digest=sha256:a1bb825d43494f4fae12ef07fd60ad2b6ce7322a06820cba791eb7279e564217

Observation ba438656-0a9a-4d98-b8b3-3e3e0131593e · outbound

This paper cites AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.158601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.443563Z digest=sha256:6e505f677d6f8af9a6b0ece91dcabe9302480264d96355446e47c137c9564bf2

Observation da66c525-b791-4f70-b319-f46780d29c71 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Audio set: An ontology and human-labeled dataset for audio events,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.749914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.537215Z digest=sha256:5c4a9b1d934c55b8e048b3554f969a0bcc757dbbcd2170b3504bf6de67236926

Observation 2f38cb89-90c3-4efc-a684-0e9f230f34ef · outbound

This paper cites GPT-4o System Card.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.608433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.608433Z digest=sha256:0b197733c69dfdb59fc3b452bceff0b5d311876f26ac506e0d6f9f9410e50c07

Observation 2b12572a-98b7-40e1-982f-7e4a3623e13e · outbound

This paper cites Toward a recommendation for a European standard of peak and LKFS loudness levels,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Toward a recommendation for a European standard of peak and LKFS loudness levels,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.515817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.684202Z digest=sha256:2ba7bd44f4f0e63d68669b7fbcb2e50df2e1d21d1bca3a08eb7afdf0e0291ad6

Observation 59bec6dc-e9bc-4d7a-9212-5e66caa09e6b · outbound

This paper cites Long short-term memory,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Long short-term memory,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.770747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.770747Z digest=sha256:30759cb96cacb7ecdd3a4dd65070c5ea64b086e591f23281394a41bf9fae590b

Observation b1910e17-cd90-41d9-b1cc-604b71c8a438 · outbound

This paper cites Adam: A method for stochastic opti- mization.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Adam: A method for stochastic opti- mization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.421775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.865536Z digest=sha256:e9021b471639ece1bd4507f6be242204c5216b36793348909c48f49efe36b8ca

Observation 29f827db-996a-4237-a769-64a90c6b7694 · outbound

This paper cites Why does music source separation benefit from cacophony?.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Why does music source separation benefit from cacophony?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.293933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T11:26:10.917974Z digest=sha256:b4c07428d494b4f95daa5eaa52b7e72ad4240027f41285c3749a4c92cf1af5b2

Pith citing papers

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · inbound

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds cites this paper.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation caee2fc7-fefb-4d55-bb43-363e3ab0548b · inbound

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) cites this paper.

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:29.070570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-07T08:16:56.671345Z digest=sha256:121eae3195b364242fd301165d2824618330bd3f529a93c567ab1a2c49ff94c7