Pith. sign in

Paper Citation Record · LEDGER

Separate Anything You Describe

As of 20 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 10 inbound Pith citation observations for arXiv:2308.05037.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2308.05037 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 10 of 10 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-20T06:33:59.587034+00:00

measured 10 of 10 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T10:37:33.254290Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T01:06:23.917534Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation 78d0d2b7-7e5e-410d-935a-c9cac5453cd8 · inbound

VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation cites this paper.

VinTAGe: Joint Video and Text Conditioning for Holistic Audio Generation Separate Anything You Describe

Reference 29

Resolution
unresolved
no resolver link, observed 2026-08-11T15:43:00.970707Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T15:43:00.970707Z digest=sha256:c83e53871c89096db492495b0a3f470ea269702626b8850f124ba44bbb63dc07

Observation 790022a0-24b9-470c-a242-b5139b912fd3 · inbound

Beyond Speaker Identity: Text Guided Target Speech Extraction cites this paper.

Beyond Speaker Identity: Text Guided Target Speech Extraction Separate Anything You Describe

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-10T20:14:00.636550Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T20:14:00.636550Z digest=sha256:1308aecd443ffe1cc69e8ca1eac11d3d77e1de72753f6a7c16e1973c93e50f68

Observation 58d365df-2a64-498a-9c5d-1197f4deb979 · inbound

30+ Years of Source Separation Research: Achievements and Future Challenges cites this paper.

30+ Years of Source Separation Research: Achievements and Future Challenges Separate Anything You Describe

Reference 57

Resolution
unresolved
no resolver link, observed 2026-08-10T17:51:52.922921Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T17:51:52.922921Z digest=sha256:b9289973b62d3a14b8fb31cf3a57f8f81b168a9d562b37dfe94368518b89fcec

Observation 253a674d-a8d9-481b-af07-8ceb909f0402 · inbound

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey cites this paper.

Audio-Language Models for Audio-Centric Tasks: A Systematic Survey Separate Anything You Describe

Reference 95

Resolution
unresolved
no resolver link, observed 2026-08-10T14:36:19.569685Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:36:19.569685Z digest=sha256:b9c6afad1c4b5371e57830acb214032d3f048711d312d022dddf5f04e624cec0

Observation cab5566a-be12-42c0-aea6-03406f441159 · inbound

Unleashing the Power of Natural Audio Featuring Multiple Sound Sources cites this paper.

Unleashing the Power of Natural Audio Featuring Multiple Sound Sources Separate Anything You Describe

Reference 2022

Resolution
unresolved
no resolver link, observed 2026-08-16T10:37:33.254290Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T10:37:33.254290Z digest=sha256:207caa01e0d9c0a08c53952ef69e43fdfebb48b5c9672259549411ca39746ed6

Observation 9df64849-d012-4c6f-9e7d-8dfbed5d46c9 · inbound

ZeroSep: Separate Anything in Audio with Zero Training cites this paper.

ZeroSep: Separate Anything in Audio with Zero Training Separate Anything You Describe

Reference 2021

Resolution
unresolved
no resolver link, observed 2026-08-07T12:46:43.184878Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:46:43.184878Z digest=sha256:0a63afc4e5d48d419c5056fabe2712ac3bea3167cd59f065aa4dffadd84ed8e9

Observation ab0e47be-03e9-43c6-a6ce-4cc77204d38a · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents Separate Anything You Describe

Reference 24

Resolution
verified exact
arxiv_id, observed 2026-05-18T16:46:37.627949Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-05-18T16:44:32.104114Z digest=sha256:be9fde7d56f21f71a1dafb95b05c822740d81336842f0483b04f837f0b063130

Observation b97b780c-df94-479f-a1e5-0a55c2911d6e · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents Separate Anything You Describe

Reference 24

Resolution
unresolved
no resolver link, observed 2026-08-04T16:48:33.087836Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:48:33.087836Z digest=sha256:42f3c5aa066f4850a04e08ffa4f6b1101fbceffc0d3e31b2dce5c8de4a70e801

Observation 60f5033c-3ae9-472c-b698-9112a0626cb0 · inbound

MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation cites this paper.

MMAudioSep: Taming Video-to-Audio Generative Model Towards Video/Text-Queried Sound Separation Separate Anything You Describe

Reference 11

Resolution
metadata mismatch
arxiv_id, observed 2026-05-18T08:21:06.840865Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=pdf_text observed=2026-05-18T08:20:02.986562Z digest=sha256:171f3af83f4138e6107ec02d58d48aa3792e3725849f2300ae08d0eb6a4c06e8

Observation c38663f4-4bb6-4f2f-b7a2-6b9917adee3f · inbound

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing cites this paper.

SpeechEditBench: A Bilingual Multi-Attribute Benchmark for Instruction-Guided Speech Editing Separate Anything You Describe

Reference 49

Resolution
metadata mismatch
arxiv_id, observed 2026-07-02T01:06:23.918980Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-20T06:33:59.587034+00:00.

source=arxiv_source observed=2026-06-28T12:54:20.815371Z digest=sha256:ba585cdf2abc3e073f6f844f1964d6fb0cb2f625cb6eb97202d1b3a09706cbaf