Pith. sign in

Paper Citation Record · LEDGER

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 2 inbound Pith citation observations for arXiv:2506.02499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02499 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:10.917974Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-07T06:34:17.273281+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:08.307554Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T10:01:29.067292Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · outbound

This paper cites DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation 477113ed-8c3a-4570-9f59-53ac54726830 · outbound

This paper cites an unresolved cited work.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:26:13.691594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.397233Z digest=sha256:a9884b5ae2fb310ed363de14543a35f26cbf7129741629d5a998c801f3121edd

Observation 240f02a4-00a8-467e-85ed-205cb400a7a6 · outbound

This paper cites Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.633367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.494147Z digest=sha256:5f6b0123706449dcbcc2f83bcda2088a2be64257d355ca8921f3d97ea26cbfcc

Observation 863e12a4-1327-4f5c-9934-ab957207a5b1 · outbound

This paper cites Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.612512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.559016Z digest=sha256:f1c39f747f55dd3104b6372530dc3264113e762f3809166257536390f937e17c

Observation 3b355aea-c3fd-4d79-9704-b58a086f8e51 · outbound

This paper cites To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.596545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.675340Z digest=sha256:d958d65e2c960363baac6df975e977613b9bf89aad1467ecf17427eaeb5ec6c1

Observation 5260f1eb-e6ad-44bd-9c9d-4ddd636f6c02 · outbound

This paper cites The sound demixing challenge 2023-cinematic demixing track,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The sound demixing challenge 2023-cinematic demixing track,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.579211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.754633Z digest=sha256:26906865589ed129c76046ce6ef3c36088d1c6db6f4e0469862571647b3f91bb

Observation 853dc9ba-605b-4674-a8fc-b52f466a530d · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Supervised speech separation based on deep learning: An overview,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.561698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.834619Z digest=sha256:7d9441f0f361436b8d6f01c0c1f6293a39665278dceedc83321193e8d5a982cf

Observation a6eadb68-fd61-42e2-9745-bdf983223694 · outbound

This paper cites Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.542876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:08.966726Z digest=sha256:88f870d31567b4b1daf6c51052580a9eb49058c59715d74ad7afe3b9ad42da26

Observation bbb4d0b9-de4e-40fc-88e8-c020c634f3c7 · outbound

This paper cites TF-GridNet: Integrating full- and sub-band modeling for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds TF-GridNet: Integrating full- and sub-band modeling for speech separation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.054261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.054261Z digest=sha256:7d024604d781476be354863fb762789b1282a4b3838775607b05a768dbb2ceca

Observation c314b9b3-8fe4-4a57-b994-a22666d5099c · outbound

This paper cites The 2018 signal separation eval- uation campaign,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The 2018 signal separation eval- uation campaign,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.513919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.130709Z digest=sha256:c03bb7445559772efbd05db07cfe399e403cc74a0675e5a6513a40fb1638753b

Observation 08e9fa2b-41cb-4588-8611-2d27d872c9a3 · outbound

This paper cites Open- Unmix - a reference implementation for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Open- Unmix - a reference implementation for music source separation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.492320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.206282Z digest=sha256:58e9b3b6ce0644507d47a1d07238077ed609a111bea2e03aeb4cb56069f7ad35

Observation 94ae7cea-4d47-4906-bd38-e3e254de2160 · outbound

This paper cites Hybrid transformers for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hybrid transformers for music source separation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.470963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.309694Z digest=sha256:e33782e69f7c593cd65c46cf9814ef1a4491f0925094053e26503fe0840c62f4

Observation 452d209e-b91f-4382-aa50-02423d4e3b70 · outbound

This paper cites Universal sound separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Universal sound separation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.451787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.391215Z digest=sha256:e415e9bebf0688ca4de8da98a76abebd2982d9a6f3bb66e061478500cddf676c

Observation 8841dde0-c76c-43ad-a949-2bfb315154a6 · outbound

This paper cites The cocktail fork problem: Three-stem audio separation for real- world soundtracks,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The cocktail fork problem: Three-stem audio separation for real- world soundtracks,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.432325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.462321Z digest=sha256:0fdc7e7283a9a9c2d797dc12a4e4ce6caedc829ccdb0233981fa660b1da198f6

Observation 21b927c6-b541-4d7e-a111-79d74a476add · outbound

This paper cites A general- ized bandsplit neural network for cinematic audio source separa- tion,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds A general- ized bandsplit neural network for cinematic audio source separa- tion,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.413048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.546046Z digest=sha256:0adf1bec0298fc5832133640c9b5f53a331a1abe0b6fa4a9a8fcb9e725aff1ce

Observation b136c3a8-04ec-4a30-ac97-fbd73ea42424 · outbound

This paper cites Music source separation with band-split RNN,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Music source separation with band-split RNN,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.392521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.655749Z digest=sha256:e3a965992d56f1b4d9a749f475b35542bc7d2eda7777e6d1aa7604569ebee323

Observation ec275346-37c0-4174-9872-8a9592af11af · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.369763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.733574Z digest=sha256:160892c08b86a9b43dd5e1475348bb6529d48f22f06855c6bbda746535b52ed2

Observation 9b6ec578-7f61-4172-9148-91bed59b0c0b · outbound

This paper cites FMA: A dataset for music analysis,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FMA: A dataset for music analysis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.322807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:09.819652Z digest=sha256:420e06f3f4add05c78eebd8c36d39b74894f2ad2788a5e2212e41022cf5c32fd

Observation 67de3525-a519-43b2-88af-a073a98efd70 · outbound

This paper cites FSD50K: An Open Dataset of Human-Labeled Sound Events.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.913604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.913604Z digest=sha256:6af201b0abce6b21551ca1f4da8489825325bb8de8cfd1927b9708d0045fc913

Observation 188cfe9e-2344-4090-b6b5-6b75da1901db · outbound

This paper cites Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.168035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.009245Z digest=sha256:7e78b976a1f510f7976f47dc2320a15a3f518879a4770a9419222e109f020c16

Observation 7fff145e-f462-4997-a7c8-7e4cb70e19e4 · outbound

This paper cites Hy- perbolic audio source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hy- perbolic audio source separation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.918502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.138699Z digest=sha256:e04953e29b8ee0eef6e6982de7ee44a9752d4c6b0a18dd079e39a1fd286075b0

Observation e4b34b3d-ccec-44a3-a532-c59dc6312802 · outbound

This paper cites PodcastMix: A dataset for separating music and speech in podcasts,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds PodcastMix: A dataset for separating music and speech in podcasts,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.830047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.209784Z digest=sha256:fb6ef214bd7904a884c6726a6ea252b650d148ecdb6ef3ff413ef7067a4c0597

Observation bc6ca3bc-edb9-4db9-81d4-183c49c8bdc7 · outbound

This paper cites Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:26:11.116952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.280208Z digest=sha256:6a67794f1dc08465f209ea6ba8e6fc2e1948042c02c804ccfcfaae5612b8db06

Observation cfd91ca2-5a4d-4c17-a42d-6061da0cf752 · outbound

This paper cites CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.496001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.372449Z digest=sha256:2d47a70e2a4eebc5da1d1ab27b6e4e6018a5b503ad7cdbea1dfb3c910266a895

Observation ba438656-0a9a-4d98-b8b3-3e3e0131593e · outbound

This paper cites AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.158601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.443563Z digest=sha256:cbdff81049444e3b8dbb52fc805ebe58d7fa913859dd9026d7879c9b607e32ca

Observation da66c525-b791-4f70-b319-f46780d29c71 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Audio set: An ontology and human-labeled dataset for audio events,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.749914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.537215Z digest=sha256:8a0d5429b12a4a7418cfbf93fb9a05c99159e274a3cac1792801045a84d9fd3f

Observation 2f38cb89-90c3-4efc-a684-0e9f230f34ef · outbound

This paper cites GPT-4o System Card.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.608433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.608433Z digest=sha256:0b197733c69dfdb59fc3b452bceff0b5d311876f26ac506e0d6f9f9410e50c07

Observation 2b12572a-98b7-40e1-982f-7e4a3623e13e · outbound

This paper cites Toward a recommendation for a European standard of peak and LKFS loudness levels,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Toward a recommendation for a European standard of peak and LKFS loudness levels,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.515817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.684202Z digest=sha256:764bd3347dd3278cab351eb50c954a8e9b6cbad70fbd64cb6e13fd23cf150144

Observation 59bec6dc-e9bc-4d7a-9212-5e66caa09e6b · outbound

This paper cites Long short-term memory,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Long short-term memory,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.770747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.770747Z digest=sha256:30759cb96cacb7ecdd3a4dd65070c5ea64b086e591f23281394a41bf9fae590b

Observation b1910e17-cd90-41d9-b1cc-604b71c8a438 · outbound

This paper cites Adam: A method for stochastic opti- mization.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Adam: A method for stochastic opti- mization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.421775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.865536Z digest=sha256:9ef331be7249eabaa74c80c573147aceb0ca0af85fe082d559cf06d25b7017c5

Observation 29f827db-996a-4237-a769-64a90c6b7694 · outbound

This paper cites Why does music source separation benefit from cacophony?.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Why does music source separation benefit from cacophony?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.293933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-08-07T11:26:10.917974Z digest=sha256:af61140d87565578c64eb295555880eb4e0813ecf85e9abc514f28f70fd364b3

Pith citing papers

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · inbound

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds cites this paper.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation caee2fc7-fefb-4d55-bb43-363e3ab0548b · inbound

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) cites this paper.

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:29.070570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-07T06:34:17.273281+00:00.

source=pdf_text observed=2026-05-07T08:16:56.671345Z digest=sha256:e07498badc449e530bec0546fe21306b43e5220677d9d6e80e51b5122d3b84d4