Pith. sign in

Paper Citation Record · LEDGER

FMA: A Dataset For Music Analysis

As of 21 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 57 inbound Pith citation observations for arXiv:1612.01840.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
1612.01840 v3

Coverage vector

measured 0 of 0 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links

measured 57 of 57 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 57 of 57 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-16T12:11:21.676866Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-07-04T14:39:57.208412Z

Reference resolution

0 of 0 outbound references displayed

  • verified exact0
  • verified fuzzy0
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

No outbound reference observations are available for this paper version.

Pith citing papers

Observation d29f0321-b96e-4b07-a2b3-64d988468731 · inbound

DALI: a large Dataset of synchronized Audio, LyrIcs and notes, automatically created using teacher-student machine learning paradigm cites this paper.

DALI: a large Dataset of synchronized Audio, LyrIcs and notes, automatically created using teacher-student machine learning paradigm FMA: A Dataset For Music Analysis

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-25T15:40:59.495490Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-25T15:40:56.024804Z digest=sha256:f5502ecf2ffe2c2917528b646e3a4aa91222cc6eab5c33faa2280418c1f9fd44

Observation 85d258bd-0eff-4a6e-8db1-99233692669d · inbound

From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview cites this paper.

From Audio Deepfake Detection to AI-Generated Music Detection -- A Pathway and Overview FMA: A Dataset For Music Analysis

Reference 67

Resolution
unresolved
no resolver link, observed 2026-08-12T05:15:07.422391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:15:07.422391Z digest=sha256:afa46a84556ae77a7f6f035b17421c202990425bac9b881297250d6131e2877d

Observation a161cdce-a852-4f28-8fb6-2f076727a843 · inbound

Audio Atlas: Visualizing and Exploring Audio Datasets cites this paper.

Audio Atlas: Visualizing and Exploring Audio Datasets FMA: A Dataset For Music Analysis

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-12T05:15:01.203137Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T05:15:01.203137Z digest=sha256:87fc6bfde78054c126ef6e3c36cb45ca7e56840229bee0e3f16575c93914b05f

Observation 928bce3c-4967-43fe-92ec-0b483a988bb8 · inbound

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? cites this paper.

AV-Odyssey Bench: Can Your Multimodal LLMs Really Understand Audio-Visual Information? FMA: A Dataset For Music Analysis

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-11T23:19:11.485731Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T23:19:11.485731Z digest=sha256:12b3d53719342b19000e9434c1d29537f7d717acb247a19093dfa9bdfafad1a8

Observation 7f1a0eae-bc95-49e0-94b6-db766ed7b782 · inbound

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey cites this paper.

Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey FMA: A Dataset For Music Analysis

Reference 80

Resolution
unresolved
no resolver link, observed 2026-08-11T14:59:01.460484Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-11T14:59:01.460484Z digest=sha256:a01164f00657abf9e1a7fb7720ebfc8f52eba92502390bed3a671d8344af13fc

Observation dec191d5-519a-476b-bde9-c7a80e7615cd · inbound

A2SB: Audio-to-Audio Schrodinger Bridges cites this paper.

A2SB: Audio-to-Audio Schrodinger Bridges FMA: A Dataset For Music Analysis

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T18:29:47.747204Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-10T18:29:47.747204Z digest=sha256:b37ae1c741a7c4dd6175ca2ed3fd0b75d99a5d1744a9b2b0f8f17296d5f73290

Observation f443ba02-0d14-4007-91a0-ad1005d83722 · inbound

The ICME 2025 Audio Encoder Capability Challenge cites this paper.

The ICME 2025 Audio Encoder Capability Challenge FMA: A Dataset For Music Analysis

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-10T14:27:09.170151Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T14:27:09.170151Z digest=sha256:4ff5931e1ebeede30f962425d132911d5f1d3d1b141d06b9b2737b089cb054db

Observation f4bb41ba-fa7e-4fe0-bfea-840defa3ec1b · inbound

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention cites this paper.

XAttnMark: Learning Robust Audio Watermarking with Cross-Attention FMA: A Dataset For Music Analysis

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-25T08:30:31.610067Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-25T08:30:15.011210Z digest=sha256:151d53d52f5e3168eb8b6837571fa420e9fd7c11580add31bdbc86c2df143ee3

Observation 5d51f611-753a-4d92-91d2-078b0943e5cb · inbound

MusFlow: Multimodal Music Generation via Conditional Flow Matching cites this paper.

MusFlow: Multimodal Music Generation via Conditional Flow Matching FMA: A Dataset For Music Analysis

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-16T12:11:21.676866Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-16T12:11:21.676866Z digest=sha256:089ed8012deceb9258aed90f2c33427086f1e96c74817c3199a2d72f81d349c1

Observation d7da1f91-be9f-48ea-af98-cd3353e98970 · inbound

Is MixIT Really Unsuitable for Correlated Sources? Exploring MixIT for Unsupervised Pre-training in Music Source Separation cites this paper.

Is MixIT Really Unsuitable for Correlated Sources? Exploring MixIT for Unsupervised Pre-training in Music Source Separation FMA: A Dataset For Music Analysis

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-15T22:15:59.239081Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T22:15:59.239081Z digest=sha256:4b6be1e51fc252fe572eaf9c293abcbaa56bc803dd504c553523c06aa44f47bb

Observation ed313b6f-be7c-49e9-9de2-3c4aff4b2c79 · inbound

X-ARES: A Comprehensive Framework for Assessing Audio Encoder Performance cites this paper.

X-ARES: A Comprehensive Framework for Assessing Audio Encoder Performance FMA: A Dataset For Music Analysis

Reference 40

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:15.124215Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:15.124215Z digest=sha256:04e1a1f7fc055d47ef7ba400586c74566e7712250f09ad7cde0eb594b7f99002

Observation 338d2442-9110-4de6-ae1c-2e93bd18061e · inbound

Interspeech 2025 URGENT Speech Enhancement Challenge cites this paper.

Interspeech 2025 URGENT Speech Enhancement Challenge FMA: A Dataset For Music Analysis

Reference 46

Resolution
unresolved
no resolver link, observed 2026-08-07T12:55:40.415886Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:55:40.415886Z digest=sha256:11c7b22fed3e962a6c6c7b9f7141df80ec7d068f2567dcc905392867d097a446

Observation b1cc5572-49ff-43f5-936e-dbc5cc69d855 · inbound

Learning Normal Patterns in Musical Loops cites this paper.

Learning Normal Patterns in Musical Loops FMA: A Dataset For Music Analysis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-07T14:52:45.610128Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T14:52:45.610128Z digest=sha256:75c600f8becebae3a3cd79817b1624bce93607126cdc00f7c01e5bd540ecea09

Observation 49ad595e-1a96-485a-8c8a-0b34744200d1 · inbound

The iNaturalist Sounds Dataset cites this paper.

The iNaturalist Sounds Dataset FMA: A Dataset For Music Analysis

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-07T12:12:05.264644Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T12:12:05.264644Z digest=sha256:51359001c80244d2b75123af0699e38f75ba770ca67dfed2f12013b6da471f20

Observation 474c5d6f-6068-49f9-8e64-fb3fadfd2693 · inbound

IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling cites this paper.

IMPACT: Iterative Mask-based Parallel Decoding for Text-to-Audio Generation with Diffusion Modeling FMA: A Dataset For Music Analysis

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-07T12:03:23.806821Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-07T12:03:23.806821Z digest=sha256:5705e5a3fa968e51cfe27a72ab7d7de7757d75840de9193d6fe28c8e381d34a5

Observation 20b1d2e6-f710-464a-b6f8-146cc13a3fbc · inbound

WAKE: Watermarking Audio with Key Enrichment cites this paper.

WAKE: Watermarking Audio with Key Enrichment FMA: A Dataset For Music Analysis

Reference 38

Resolution
unresolved
no resolver link, observed 2026-08-07T10:19:28.877334Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T10:19:28.877334Z digest=sha256:661a6b9fe60f2cef781ae0269c054c931577459ec9764583207905b744b6c4f7

Observation 0b91da87-8775-4547-ab6f-04ce1d44213c · inbound

MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners cites this paper.

MuseControlLite: Multifunctional Music Generation with Lightweight Conditioners FMA: A Dataset For Music Analysis

Reference 2009

Resolution
unresolved
no resolver link, observed 2026-08-06T23:19:50.376859Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T23:19:50.376859Z digest=sha256:2e9b28e079435e2572377f27328500555ccd595a0e38da2d18b3b3faf1707ce0

Observation 9aae6ff8-95b4-4c42-a807-87e05aeef59f · inbound

Benchmarking Music Generation Models and Metrics via Human Preference Studies cites this paper.

Benchmarking Music Generation Models and Metrics via Human Preference Studies FMA: A Dataset For Music Analysis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-15T18:41:08.308210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:41:08.308210Z digest=sha256:452c85c60da7d9f0a0af6360e9512312e28a22aeb816425a00657304041c1286

Observation 10d41b4f-b297-4e30-975a-a6e679643230 · inbound

A Fourier Explanation of AI-music Artifacts cites this paper.

A Fourier Explanation of AI-music Artifacts FMA: A Dataset For Music Analysis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-15T18:41:54.612030Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T18:41:54.612030Z digest=sha256:0c796bafa5dc7f881d706788c0da102e0d34dc2e65d1e4d09870c351633f57d1

Observation 9226004b-7c4c-4480-ba4e-5aa60fa31cbb · inbound

Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning cites this paper.

Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning FMA: A Dataset For Music Analysis

Reference 30

Resolution
unresolved
no resolver link, observed 2026-08-06T22:56:46.311880Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T22:56:46.311880Z digest=sha256:c1abaeb6f3cbe4cb26f699e0cda809105bf0e715e39431658ee2cfb26b9287da

Observation 3209eada-4973-45a4-aeda-701aaf59990a · inbound

Spatial and Semantic Embedding Integration for Stereo Sound Event Localization and Detection in Regular Videos cites this paper.

Spatial and Semantic Embedding Integration for Stereo Sound Event Localization and Detection in Regular Videos FMA: A Dataset For Music Analysis

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T19:43:00.151931Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T19:43:00.151931Z digest=sha256:16d5012202115b664deb8d601e2c25665f12ecaf38b57f6592bbbe010e89545f

Observation 943366a9-e9d0-45f8-8378-6edcf31fe4c1 · inbound

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models cites this paper.

Audio Flamingo 3: Advancing Audio Intelligence with Fully Open Large Audio Language Models FMA: A Dataset For Music Analysis

Reference 25

Resolution
verified exact
local_arxiv, observed 2026-05-15T03:42:44.813536Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-15T03:42:44.523919Z digest=sha256:dce7bc67a4bacbff6e56179545d5dd32e3c9a7e5bd84a5efd9938a609c36cb10

Observation 41ee785f-e499-41d4-b0bd-f9e515a3ff6c · inbound

FasTUSS: Faster Task-Aware Unified Source Separation cites this paper.

FasTUSS: Faster Task-Aware Unified Source Separation FMA: A Dataset For Music Analysis

Reference 51

Resolution
unresolved
no resolver link, observed 2026-08-06T17:12:56.174530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T17:12:56.174530Z digest=sha256:378b6e62f38aa17a6b05ce8e2488f9905f4aaea650bddbe2018822d9315d0acb

Observation 4fee4e4e-2041-4248-86d3-983a29494207 · inbound

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries cites this paper.

Cross-Modal Watermarking for Authentic Audio Recovery and Tamper Localization in Synthesized Audiovisual Forgeries FMA: A Dataset For Music Analysis

Reference 2017

Resolution
unresolved
no resolver link, observed 2026-08-06T16:45:54.536804Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T16:45:54.536804Z digest=sha256:200a901fb0a19ebd15a33b1c28e340e2c48677236b87b01310d82c2e4ee574fc

Observation b3f266ff-b6c0-487d-94c1-046df61f1605 · inbound

ASAudio: A Survey of Advanced Spatial Audio Research cites this paper.

ASAudio: A Survey of Advanced Spatial Audio Research FMA: A Dataset For Music Analysis

Reference 34

Resolution
unresolved
no resolver link, observed 2026-08-05T22:54:54.895720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-05T22:54:54.895720Z digest=sha256:3e5ed8862340d297873e9133861416d3bd2740a2769a647c8da511b2b1ef0f39

Observation 9a63cdae-413a-41d1-aadf-f6fb1e82b8a5 · inbound

Integrating Spatial and Semantic Embeddings for Stereo Sound Event Localization in Videos cites this paper.

Integrating Spatial and Semantic Embeddings for Stereo Sound Event Localization in Videos FMA: A Dataset For Music Analysis

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-15T16:20:26.209072Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T16:20:26.209072Z digest=sha256:f0256d295a72faa8ecf35188104268aad07744cbe303bd3ce4745796bea9a2e0

Observation 39cdb5e6-4992-48e7-bb37-2ceba1c75ed8 · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents FMA: A Dataset For Music Analysis

Reference 6

Resolution
verified exact
local_arxiv, observed 2026-05-18T16:46:37.638591Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-18T16:44:32.104114Z digest=sha256:04d814b932b36f1efeabb144733d36b69e9ba4ce0d54421717f47e614bb84ed1

Observation e3a7abb3-fade-4cbf-8b95-e9ab6154ff86 · inbound

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents cites this paper.

CodecSep: Prompt-Driven Universal Sound Separation on Neural Audio Codec Latents FMA: A Dataset For Music Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T16:48:31.235518Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T16:48:31.235518Z digest=sha256:dc52a30f80acf4d525c49125777345ae895baf750f2801bd97af6917a9db10e7

Observation e95830f5-79d3-4cc2-845b-f90fd8f009c5 · inbound

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance cites this paper.

AudioMoG: Guiding Audio Generation with Mixture-of-Guidance FMA: A Dataset For Music Analysis

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-05-18T13:11:23.846476Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-18T13:10:18.700497Z digest=sha256:466d64383869a33e27d2b16e0aa8686b0497205011d47e648f7784e9169d980b

Observation a52013d4-8ba8-4e2b-b8a6-ee0c8a60b53f · inbound

Assessing Factual Music Comprehension in Large Audio Language Models cites this paper.

Assessing Factual Music Comprehension in Large Audio Language Models FMA: A Dataset For Music Analysis

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-04T00:29:46.062032Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=arxiv_source observed=2026-08-04T00:29:46.062032Z digest=sha256:70fbc1e5a1f294ba355fa00a0fd8e498901f31301ef64d577837c397abcbd468

Observation 494ba8dc-1fbc-4e1c-b183-7dafc254f609 · inbound

HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal cites this paper.

HarmonicAttack: An Adaptive Cross-Domain Audio Watermark Removal FMA: A Dataset For Music Analysis

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-21T19:15:30.943833Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-21T19:15:08.392087Z digest=sha256:d5e4d848be654c6a7b4bdacee3462f29d55f0c921a2790d644bb919a83b1b16c

Observation 69ec74ed-753e-4143-b899-91fed4f73e0f · inbound

Two-Dimensional Quantization for Geometry-Aware Audio Coding cites this paper.

Two-Dimensional Quantization for Geometry-Aware Audio Coding FMA: A Dataset For Music Analysis

Reference 18

Resolution
verified exact
local_arxiv, observed 2026-05-21T18:20:29.146259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-05-21T18:16:51.486807Z digest=sha256:6c30a5edfd5c648539fb242d5fa20c6aab317dadc24fdf815a40f79cc48ae517

Observation 88013aec-da27-4554-9288-e3fce48237b8 · inbound

Enhancing Automatic Chord Recognition via Pseudo-Labeling and Knowledge Distillation cites this paper.

Enhancing Automatic Chord Recognition via Pseudo-Labeling and Knowledge Distillation FMA: A Dataset For Music Analysis

Reference 33

Resolution
unresolved
no resolver link, observed 2026-08-02T21:35:00.996581Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T21:35:00.996581Z digest=sha256:c53a0e3cee8f1b91c7fe1bdb8f9653644227aadabaa231c92394127e472689f0

Observation 3349f34a-a4ba-4c31-b76e-bfc992d08c79 · inbound

Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training cites this paper.

Making Separation-First Multi-Stream Audio Watermarking Feasible via Joint Training FMA: A Dataset For Music Analysis

Reference 27

Resolution
unresolved
no resolver link, observed 2026-08-02T18:03:42.342050Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T18:03:42.342050Z digest=sha256:52d95ae817c9d8b0a569a82e4396377665991f3fd3565a7617922e34caf6b3dd

Observation 7fdb7cf7-d862-4054-a787-dff6e4e959ba · inbound

HAFM: Hierarchical Autoregressive Foundation Model for Music Accompaniment Generation cites this paper.

HAFM: Hierarchical Autoregressive Foundation Model for Music Accompaniment Generation FMA: A Dataset For Music Analysis

Reference 22

Resolution
verified exact
arxiv_id, observed 2026-05-11T06:46:22.298444Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T17:27:56.766226Z digest=sha256:c082d6fd2ce0e2272c2f3fb65108d9d0177151ba2f082e0703fe74213decc788

Observation 3ce85a7c-0505-455c-b0c7-f98cb5f9bf46 · inbound

Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music cites this paper.

Cross-Cultural Bias in Mel-Scale Representations: Evidence and Alternatives from Speech and Music FMA: A Dataset For Music Analysis

Reference 23

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:05:58.786562Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T16:16:47.220966Z digest=sha256:d061f4be8db95081cd7948508dae5ecf329345b581a569d740b07bbb5fd8aecc

Observation 0d810de9-39cc-40f3-8ee7-abee38624506 · inbound

Multimodal Dataset Normalization and Perceptual Validation for Music-Taste Correspondences cites this paper.

Multimodal Dataset Normalization and Perceptual Validation for Music-Taste Correspondences FMA: A Dataset For Music Analysis

Reference 19

Resolution
verified exact
arxiv_id, observed 2026-05-11T09:56:03.679770Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T15:44:39.317308Z digest=sha256:2aa3da3d0fa8827364302359ef087df6a517fcc352ffdfeabc27d8be60028c08

Observation 251e5677-3c24-4d60-814a-ad7f567a4f2d · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models FMA: A Dataset For Music Analysis

Reference 42

Resolution
verified exact
arxiv_id, observed 2026-05-10T14:10:28.253227Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T14:10:03.707886Z digest=sha256:287c9b1ffa5f68180b54d3db9f41c9db4e2ee4ae4509927a56ae6e8c71a6675d

Observation be112453-7ba3-4166-8d39-9a3d124d424d · inbound

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models cites this paper.

Audio-Cogito: Towards Deep Audio Reasoning in Large Audio Language Models FMA: A Dataset For Music Analysis

Reference 42

Resolution
unresolved
no resolver link, observed 2026-07-12T21:18:46.566338Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-07-12T21:18:46.566338Z digest=sha256:8ce159eb18d2df2b8db3f9c8302c80ac7cd04de087176af6a7645bdea739b3e4

Observation 05d35713-f041-4f62-862d-267e30ee1a94 · inbound

UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations cites this paper.

UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations FMA: A Dataset For Music Analysis

Reference 47

Resolution
verified exact
arxiv_id, observed 2026-05-10T10:19:20.002930Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-10T10:18:16.972414Z digest=sha256:63baf5798dd5a81272d2a65762d28c50d95571b703880b104a1f0e22569718de

Observation c112a0c5-3b9a-40a7-97a0-58d2355bc977 · inbound

UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations cites this paper.

UniPASE: A Generative Model for Universal Speech Enhancement with High Fidelity and Low Hallucinations FMA: A Dataset For Music Analysis

Reference 47

Resolution
unresolved
no resolver link, observed 2026-08-02T16:16:37.705672Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-02T16:16:37.705672Z digest=sha256:12081fc5a6cfe8455974518e8c12586a757df23242392e1e1b5333cca466e2cc

Observation e8743e60-7d02-4ea0-b800-95276f28e966 · inbound

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) cites this paper.

A Knowledge-Driven Approach to Target Speech Extraction in the Presence of Background Sound Effects for Cinematic Audio Source Separation (CASS) FMA: A Dataset For Music Analysis

Reference 40

Resolution
verified exact
arxiv_id, observed 2026-05-12T10:01:29.076056Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-07T08:16:56.671345Z digest=sha256:b3078b0342cb09768e7463b9a3844d7f266eb13e80252498178b4effdca639c0

Observation 9ad26cf3-fc44-4db2-88c2-495bf41a42b2 · inbound

Fast Text-to-Audio Generation with One-Step Sampling via Energy-Scoring and Auxiliary Contextual Representation Distillation cites this paper.

Fast Text-to-Audio Generation with One-Step Sampling via Energy-Scoring and Auxiliary Contextual Representation Distillation FMA: A Dataset For Music Analysis

Reference 6

Resolution
verified exact
arxiv_id, observed 2026-05-11T15:46:47.351380Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-09T19:17:09.247932Z digest=sha256:dda8e2acf90e25b32318d3a1a55e26dfeee8de7ade5ed673b2fbd4a193a158bb

Observation 7f9b8988-f1de-4e2a-9f24-f87077090665 · inbound

Reducing Linguistic Hallucination in LM-Based Speech Enhancement via Noise-Invariant Acoustic-Semantic Distillation cites this paper.

Reducing Linguistic Hallucination in LM-Based Speech Enhancement via Noise-Invariant Acoustic-Semantic Distillation FMA: A Dataset For Music Analysis

Reference 5

Resolution
verified exact
arxiv_id, observed 2026-05-12T08:26:23.244993Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-12T01:11:50.733585Z digest=sha256:7b5e1ea1ad7a2839c0b778e51c6df6d0b19dd541ef31bb7c3fc519a241705965

Observation d9ea8f60-f6eb-4455-96c8-4fd1fe8a4d83 · inbound

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music cites this paper.

Persian MusicGen: A Large-Scale Dataset and Culturally-Aware Generative Model for Persian Music FMA: A Dataset For Music Analysis

Reference 7

Resolution
verified exact
local_arxiv, observed 2026-06-30T20:25:02.398590Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=arxiv_source observed=2026-06-30T20:24:22.058025Z digest=sha256:46eb0d40b8baaec9d716d5a40f9dc0455431543d9a7c9c4951f86642e4cf65f7

Observation d9ff1b60-93f9-44c7-b838-71f856ccfb64 · inbound

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation cites this paper.

Modeling Music as a Time-Frequency Image: A 2D Tokenizer for Music Generation FMA: A Dataset For Music Analysis

Reference 22

Resolution
verified exact
local_arxiv, observed 2026-05-19T18:47:43.054207Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T18:47:32.003545Z digest=sha256:cc5b776d9403329d361ca5833943a9006e145628acbe6302f9a8e8c1aefeea76

Observation 8da5c278-5617-46a8-afc3-3239764fd6f3 · inbound

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models cites this paper.

A Survey of Advancing Audio Super-Resolution and Bandwidth Extension from Discriminative to Generative Models FMA: A Dataset For Music Analysis

Reference 10

Resolution
verified exact
local_arxiv, observed 2026-05-19T20:27:53.784656Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-19T20:26:59.049472Z digest=sha256:6e62213731f1e04e9f8b28c257bcf2a83c57183e2d441e9f8ba7fc1a7678e80b

Observation 29cbbcfe-d71d-4816-9b23-f22f62f9c68d · inbound

MusicDET: Zero-Shot AI-Generated Music Detection cites this paper.

MusicDET: Zero-Shot AI-Generated Music Detection FMA: A Dataset For Music Analysis

Reference 6

Resolution
metadata mismatch
local_arxiv, observed 2026-05-20T00:32:54.121055Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-05-20T00:31:12.146288Z digest=sha256:044c404dc36490e8f680d247133d4de6bfbc0e915fdd5e2089b4d11ead107aa9

Observation 2bc2b6aa-d75f-49bd-8af0-7a76ee96a0b6 · inbound

Auditing Training Data in Generative Music Models via Black-Box Membership Inference cites this paper.

Auditing Training Data in Generative Music Models via Black-Box Membership Inference FMA: A Dataset For Music Analysis

Reference 29

Resolution
verified exact
local_arxiv, observed 2026-06-29T08:53:16.270304Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T08:45:23.523801Z digest=sha256:1f805b779863211c932c1e66073233c16918e691e50768252e1ea6696b8b7a0a

Observation 4ed9f10c-7157-40f8-affa-bd5ef7f8be97 · inbound

Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors cites this paper.

Audio Pirates: Black-box Audio Watermark Removal via Diffusion Priors FMA: A Dataset For Music Analysis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-06-29T06:23:09.215783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-29T06:19:36.116979Z digest=sha256:933c0a024ea38f5900a1df1b556813ad9d8a9b41fa5efddbd3d9a73cbdf016c8

Observation 708dd38c-60df-4604-a8a5-a31ad1aea4b6 · inbound

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement cites this paper.

PhASE-Flow: Phonetic-Conditioned Acoustic Flow Matching in SSL Representation Domain for Speech Enhancement FMA: A Dataset For Music Analysis

Reference 36

Resolution
verified exact
local_arxiv, observed 2026-07-03T23:09:00.928538Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T23:05:08.894347Z digest=sha256:8176c8779692ff78ffa7199424717245cf6cb1ad93cd1e329d0d81a06e63cf1a

Observation 2c5c59a4-2ba3-4920-9de0-b51067863434 · inbound

STAR-VAE: Structured Topology-Aware Regularization for Audio Reconstruction and Generation cites this paper.

STAR-VAE: Structured Topology-Aware Regularization for Audio Reconstruction and Generation FMA: A Dataset For Music Analysis

Reference 2

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:59:50.428486Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T07:26:28.527338Z digest=sha256:a35829f7c3141cbeeaa2f327b252217d82474a2cef2e19c16fe43f90fc9a98f1

Observation 0505500b-268c-4acf-98c8-38a9c552e141 · inbound

AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation cites this paper.

AudioCALM: Continuous Autoregressive Language Modeling for Universal Audio Generation FMA: A Dataset For Music Analysis

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-07-04T11:59:50.931976Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T07:24:01.244733Z digest=sha256:407a207e1c16b0b126d50834f4a04c1a06db9d80ec7e3e994123660c2e47d7a5

Observation 78b4ed9e-e27b-4e43-9da3-7bd9fd7db88c · inbound

WQ-Fusion: Dynamic Gated Attention for Cross-Domain Audio Representation cites this paper.

WQ-Fusion: Dynamic Gated Attention for Cross-Domain Audio Representation FMA: A Dataset For Music Analysis

Reference 46

Resolution
verified exact
local_arxiv, observed 2026-07-04T14:39:57.209854Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-06-26T03:21:39.842959Z digest=sha256:b00ff75f2a421fc8e57acc4f9add32548dc30b7b29d4549454a41556199faf59

Observation ba313292-db27-4225-80af-cf544961c1cc · inbound

StemFX: Learning Mixing Style Representations via Autoregressive FX Chain Prediction on Source-Separated Stems cites this paper.

StemFX: Learning Mixing Style Representations via Autoregressive FX Chain Prediction on Source-Separated Stems FMA: A Dataset For Music Analysis

Reference 41

Resolution
unresolved
no resolver link, observed 2026-08-01T22:48:43.761176Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T22:48:43.761176Z digest=sha256:267e2d011dd05f48a88cf769b4a4fab0b6e4b5d76c16b1bc830394f13f0e93ad

Observation b313b453-97a4-47ac-8292-f3097e872976 · inbound

Finding the noise: Zero-shot AI Music Detection cites this paper.

Finding the noise: Zero-shot AI Music Detection FMA: A Dataset For Music Analysis

Reference 61

Resolution
unresolved
no resolver link, observed 2026-08-01T02:12:19.819862Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T02:12:19.819862Z digest=sha256:0caa32d235c4ed0d3968c69c20e03086884c2acb30692dc940353a45043d814a

Observation 91e6567f-f575-4ffb-b887-0865b1916e49 · inbound

InvFlowFD: Reference-Free and Background-Set-Free Perceptual Music Quality Metric with Flow Matching Inversion cites this paper.

InvFlowFD: Reference-Free and Background-Set-Free Perceptual Music Quality Metric with Flow Matching Inversion FMA: A Dataset For Music Analysis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-15T14:46:11.748708Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-15T14:46:11.748708Z digest=sha256:2e9b280ae71bf84070c3715d9fb6a830d2ee38cf9b2f5ee0d4beef490cf7aca6