Pith. sign in

Paper Citation Record · LEDGER

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

As of 8 August 2026, this Paper Citation Record lists 31 of 31 outbound references and 2 inbound Pith citation observations for arXiv:2506.02499.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2506.02499 v2

Coverage vector

measured 31 of 31 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:10.917974Z

measured 33 of 33 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T11:26:08.307554Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T10:01:29.067292Z

Reference resolution

31 of 31 outbound references displayed

  • verified exact1
  • verified fuzzy24
  • unresolved6
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · outbound

This paper cites DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation 477113ed-8c3a-4570-9f59-53ac54726830 · outbound

This paper cites an unresolved cited work.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T11:26:13.691594Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.397233Z digest=sha256:ff5361b4fdb6546a6f71dda878c08fffaa16822f6401d62375f72aee8f9223d6

Observation 240f02a4-00a8-467e-85ed-205cb400a7a6 · outbound

This paper cites Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Motivation In the actual movie audio, we can decomposex s as follows: xs =x v +x n,(2) wherex v andx n correspond to the waveforms of verbal and non-verbal sounds, respectively

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.633367Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.494147Z digest=sha256:29574a91117df096f469ee1ecb2657a0a15deffc82b2d73e834bae6af2a51d0b

Observation 863e12a4-1327-4f5c-9934-ab957207a5b1 · outbound

This paper cites Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Settings To evaluate the effectiveness of the proposed dataset, we con- ducted CASS experiments

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.612512Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.559016Z digest=sha256:5ffa8f041740d15bbfa9e50eef6cc00c09cc19ab52f3d9360cde68bbeee23a9f

Observation 3b355aea-c3fd-4d79-9704-b58a086f8e51 · outbound

This paper cites To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds To address this issue, we built a new dataset containing non-verbal sounds namedDnR- nonverbal

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.596545Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.675340Z digest=sha256:ed3cdcf07f6b907e1640ea1c2221b2fa3764fb83e6f8a27b50b297dcc2c47768

Observation 5260f1eb-e6ad-44bd-9c9d-4ddd636f6c02 · outbound

This paper cites The sound demixing challenge 2023-cinematic demixing track,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The sound demixing challenge 2023-cinematic demixing track,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.579211Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.754633Z digest=sha256:37b8676df0646bec82d641d3b7297db9cfd86fd11d189138e68a34ca878b06c7

Observation 853dc9ba-605b-4674-a8fc-b52f466a530d · outbound

This paper cites Supervised speech separation based on deep learning: An overview,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Supervised speech separation based on deep learning: An overview,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.561698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.834619Z digest=sha256:8ce9457fdaf742bb5b1876e4f54456d1d14dd1a3d13c8bfcd6f91efa1a5bf045

Observation a6eadb68-fd61-42e2-9745-bdf983223694 · outbound

This paper cites Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Conv-TasNet: Surpassing ideal time– frequency magnitude masking for speech separation,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.542876Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:08.966726Z digest=sha256:f08c258466cc50396faf533ee7b250608cf2bfe74fa1aa97205200994950a039

Observation bbb4d0b9-de4e-40fc-88e8-c020c634f3c7 · outbound

This paper cites TF-GridNet: Integrating full- and sub-band modeling for speech separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds TF-GridNet: Integrating full- and sub-band modeling for speech separation,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.054261Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.054261Z digest=sha256:7d024604d781476be354863fb762789b1282a4b3838775607b05a768dbb2ceca

Observation c314b9b3-8fe4-4a57-b994-a22666d5099c · outbound

This paper cites The 2018 signal separation eval- uation campaign,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The 2018 signal separation eval- uation campaign,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.513919Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.130709Z digest=sha256:fdbe94a1fca8f5ce7156026de1bcf6940df925cbe609603be078a475390684a5

Observation 08e9fa2b-41cb-4588-8611-2d27d872c9a3 · outbound

This paper cites Open- Unmix - a reference implementation for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Open- Unmix - a reference implementation for music source separation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.492320Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.206282Z digest=sha256:c215fbd759ec42af5b6ec3b4a7244ff30686720b4e7fe1bcd5a5091ede6dcf48

Observation 94ae7cea-4d47-4906-bd38-e3e254de2160 · outbound

This paper cites Hybrid transformers for music source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hybrid transformers for music source separation,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.470963Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.309694Z digest=sha256:cbf73e0b68f4bbbcf70b56438803583a2bd02647d89d39a4a45ab2f2a9bb807d

Observation 452d209e-b91f-4382-aa50-02423d4e3b70 · outbound

This paper cites Universal sound separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Universal sound separation,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.451787Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.391215Z digest=sha256:8f8be28cee2ed926be949d3fa4c1c07c3f73bf3c38bcf763657da6eb0701fd50

Observation 8841dde0-c76c-43ad-a949-2bfb315154a6 · outbound

This paper cites The cocktail fork problem: Three-stem audio separation for real- world soundtracks,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds The cocktail fork problem: Three-stem audio separation for real- world soundtracks,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.432325Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.462321Z digest=sha256:4ade7e7984e4d2e131ad2f6c6bcce97f40f6e6f1cadc65c35d4ff8c523e7c981

Observation 21b927c6-b541-4d7e-a111-79d74a476add · outbound

This paper cites A general- ized bandsplit neural network for cinematic audio source separa- tion,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds A general- ized bandsplit neural network for cinematic audio source separa- tion,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.413048Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.546046Z digest=sha256:ac9bce280f67667fb6566b7d4d0877a4eafb5822d4d584add797aab3b2ca2340

Observation b136c3a8-04ec-4a30-ac97-fbd73ea42424 · outbound

This paper cites Music source separation with band-split RNN,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Music source separation with band-split RNN,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.392521Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.655749Z digest=sha256:0920097987340d65dc8764db54771623ffa91d97918aefddfb35522d78e4e066

Observation ec275346-37c0-4174-9872-8a9592af11af · outbound

This paper cites Lib- riSpeech: an ASR corpus based on public domain audio books,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Lib- riSpeech: an ASR corpus based on public domain audio books,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.369763Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.733574Z digest=sha256:2c4829f01efa4c065da702f6427116878b7cf12108e1337a79390955f7f84f2b

Observation 9b6ec578-7f61-4172-9148-91bed59b0c0b · outbound

This paper cites FMA: A dataset for music analysis,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FMA: A dataset for music analysis,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.322807Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:09.819652Z digest=sha256:a1cd55c91de96f1668ff69f21568126f6d4efdc0ccea00a65ad0290d381c8e32

Observation 67de3525-a519-43b2-88af-a073a98efd70 · outbound

This paper cites FSD50K: An Open Dataset of Human-Labeled Sound Events.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds FSD50K: An Open Dataset of Human-Labeled Sound Events

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:09.913604Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:09.913604Z digest=sha256:6af201b0abce6b21551ca1f4da8489825325bb8de8cfd1927b9708d0045fc913

Observation 188cfe9e-2344-4090-b6b5-6b75da1901db · outbound

This paper cites Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Remastering di- vide and remaster: A cinematic audio source separation dataset with multilingual support,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:13.168035Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.009245Z digest=sha256:80f034f4c7aad454ea49ab675ce77f85f12a0a6ddf56ca4b4bec1625e4c4f366

Observation 7fff145e-f462-4997-a7c8-7e4cb70e19e4 · outbound

This paper cites Hy- perbolic audio source separation,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Hy- perbolic audio source separation,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.918502Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.138699Z digest=sha256:f6111716eb56babcec7f1d5d69c1175c53de253a87fb17a0665e7fbc8491d4e7

Observation e4b34b3d-ccec-44a3-a532-c59dc6312802 · outbound

This paper cites PodcastMix: A dataset for separating music and speech in podcasts,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds PodcastMix: A dataset for separating music and speech in podcasts,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.830047Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.209784Z digest=sha256:19edb6d9257202095d8e96cf8a8ded954723812f2732af08f9dfbe286ef7fefe

Observation bc6ca3bc-edb9-4db9-81d4-183c49c8bdc7 · outbound

This paper cites Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Jointly Recognizing Speech and Singing Voices Based on Multi-Task Audio Source Separation

Reference 23

Resolution
verified exact
local_arxiv, observed 2026-08-07T11:26:11.116952Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.280208Z digest=sha256:4269ed5ef062ab0edb9f82edae0c1844254dfe2247ee3444924952f21b61b1a7

Observation cfd91ca2-5a4d-4c17-a42d-6061da0cf752 · outbound

This paper cites CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds CSTR VCTK corpus: English multi-speaker corpus for CSTR voice cloning toolkit,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.496001Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.372449Z digest=sha256:3f908afc6fe3ef24248da8a3954e85cda39f86655914e97c9f63c9765704b168

Observation ba438656-0a9a-4d98-b8b3-3e3e0131593e · outbound

This paper cites AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds AISHELL-1: An open-source Mandarin speech corpus and a speech recognition baseline,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:12.158601Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.443563Z digest=sha256:6b868b9a82d1daef60ea7c8d60eaaaa0e561372c108c006686bc7cd1c3219e85

Observation da66c525-b791-4f70-b319-f46780d29c71 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Audio set: An ontology and human-labeled dataset for audio events,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.749914Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.537215Z digest=sha256:8e8efa356b12c973f34344cf2fcb54308e60eea5f17e1585b93891ef6d99e029

Observation 2f38cb89-90c3-4efc-a684-0e9f230f34ef · outbound

This paper cites GPT-4o System Card.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds GPT-4o System Card

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.608433Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.608433Z digest=sha256:0b197733c69dfdb59fc3b452bceff0b5d311876f26ac506e0d6f9f9410e50c07

Observation 2b12572a-98b7-40e1-982f-7e4a3623e13e · outbound

This paper cites Toward a recommendation for a European standard of peak and LKFS loudness levels,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Toward a recommendation for a European standard of peak and LKFS loudness levels,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.515817Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.684202Z digest=sha256:15a0a321bccdf09c40e372bbe77cce068bbc27f4602a3d854a93435772fb7cde

Observation 59bec6dc-e9bc-4d7a-9212-5e66caa09e6b · outbound

This paper cites Long short-term memory,.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Long short-term memory,

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:10.770747Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:10.770747Z digest=sha256:30759cb96cacb7ecdd3a4dd65070c5ea64b086e591f23281394a41bf9fae590b

Observation b1910e17-cd90-41d9-b1cc-604b71c8a438 · outbound

This paper cites Adam: A method for stochastic opti- mization.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Adam: A method for stochastic opti- mization

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.421775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.865536Z digest=sha256:d7cbde301dd659c20f6c09d6cec2b2ef9b7fe3445073fbf17d3934de85f7cbbd

Observation 29f827db-996a-4237-a769-64a90c6b7694 · outbound

This paper cites Why does music source separation benefit from cacophony?.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds Why does music source separation benefit from cacophony?

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T11:26:11.293933Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-08-07T11:26:10.917974Z digest=sha256:c99a9c763a9d40796fb0f24e51b3bef87e703f4fbb603c4291772453b37a28fd

Pith citing papers

Observation 4bc3b4e2-93ef-4f46-ba8b-e6b2a40ac92e · inbound

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds cites this paper.

DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-07T11:26:08.307554Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T11:26:08.307554Z digest=sha256:a8cb1c54c728df5616306d70f95f30ebca5651b82705bb9dadb477bea6635843

Observation caee2fc7-fefb-4d55-bb43-363e3ab0548b · inbound

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) cites this paper.

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) DnR-nonverbal: Cinematic Audio Source Separation Dataset Containing Non-Verbal Sounds

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:29.070570Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.

source=pdf_text observed=2026-05-07T08:16:56.671345Z digest=sha256:04d0217d645de2714f93a173a2314203307950d38aec83b0794236a9bd8a37b6