Pith. sign in

Paper Citation Record · LEDGER

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

As of 12 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2501.05068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05068 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:26:24.687450Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:54:36.797305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:54:36.849922Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a2afbe7-8d78-4bff-a1a2-5cb037c3d58b · outbound

This paper cites Automatic piano transcription with hierarchical frequency-time transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Automatic piano transcription with hierarchical frequency-time transformer,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.094203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.575205Z digest=sha256:87bc3d9f3ea6c7ece0a72bf2c1ba42ef67d57aaeb86d4db60be13a26e61665ab

Observation af5aa3e3-d890-4098-a970-af7f2975a60c · outbound

This paper cites Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.581126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.581126Z digest=sha256:69e610ed1c4f1ae20675427abadb333c40c02a28bd8a5401500343da64566971

Observation be6584f6-29c1-489a-bdd6-604c04934039 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.587527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.587527Z digest=sha256:7d2d598db0fd25d042aad18afc7730b8b23b3cc8e2fbcb417068b25f8dea2610

Observation 1b042cbc-17f4-43c7-9503-6237a24e4774 · outbound

This paper cites Analog bits: Generating discrete data using diffusion models with self-conditioning,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Analog bits: Generating discrete data using diffusion models with self-conditioning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.076125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.593978Z digest=sha256:bcdddd19c05b56a317b458b118efcb6a825e1035f754a4596d746f3760fc7645

Observation 77390dac-c929-401d-bc69-9513a9ae5537 · outbound

This paper cites DFormer: Diffusion-guided Transformer for Universal Image Segmentation.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DFormer: Diffusion-guided Transformer for Universal Image Segmentation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.599754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.599754Z digest=sha256:2be85d03dea7026b0a7786cd5902d1a5e954e1b4a1c9ae793b7412e67b7739be

Observation 3868c522-8ab1-436a-b23e-0315db7a688c · outbound

This paper cites Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.061433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.605412Z digest=sha256:ecf9bc3dd56054f0f213e5e0087f75dd1cdb25ced522d6e1ee81bb6d5d28205a

Observation f536c73e-5988-4d12-a8e5-ae44f208059e · outbound

This paper cites Neighborhood attention transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Neighborhood attention transformer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.045803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.610956Z digest=sha256:a5898303a35f77466f01c7897fe62895114f77080c8c8187cadf39650dd3ac26

Observation 26e84d95-6ffa-4905-8dce-71c6eb53f8f3 · outbound

This paper cites Onsets and frames: Dual-objective piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Onsets and frames: Dual-objective piano transcription,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.029369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.615497Z digest=sha256:64c228314ee54aa5817e20cbf278726193dbbb8a8cae4561722e13808a4bcedf

Observation 82ac0a44-825a-49e3-a425-67540e659301 · outbound

This paper cites High-resolution piano transcription with pedals by regressing onset and offset times,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription High-resolution piano transcription with pedals by regressing onset and offset times,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.620428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.620428Z digest=sha256:b0846e803c22819c2b2a5f90bcf97183eb2bd306baabf7f6e12b956fbeb214bf

Observation 9f487184-351b-4703-9a2a-de7fda1cda97 · outbound

This paper cites Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.001794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.625519Z digest=sha256:cb5a300049ee75d1e3788c04bc6ac2df681b59cef568cc3ed89b56c015a02451

Observation e1d870c3-7170-4f85-94a6-6643574d77a7 · outbound

This paper cites Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:26:24.775461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.630524Z digest=sha256:e6125c4c4945b699b59c9c1403247a47df389e5b5d330d74056645f60f7be907

Observation b05360f1-83a2-4d93-9a9f-39389e105218 · outbound

This paper cites Sequence-to-sequence piano transcription with transformers,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Sequence-to-sequence piano transcription with transformers,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.981643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.635317Z digest=sha256:5d657f9a077933e1bed8cfd5f44d28cceed5f96d6269f6527eb92df0140d1a3c

Observation 405d9cae-5b9c-4182-a1f4-8f4a6b0bb21f · outbound

This paper cites Piano transcription with harmonic attention,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Piano transcription with harmonic attention,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.964893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.639520Z digest=sha256:b3d185d29c8470b40660a27f62fd9f0a82bff364b12cc4022518b4a0994c00a6

Observation 1b0bbf07-2237-4a49-9669-c17c9a5fdf95 · outbound

This paper cites Polyphonic piano transcription using autoregressive multi-state note model,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Polyphonic piano transcription using autoregressive multi-state note model,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.947851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.644113Z digest=sha256:daa8309800512c65ed3b7f9b9e4ef3318e2ad8ebf71495e253ba3d9e5e7ee97b

Observation 3f6e7a02-632e-4459-8331-9483b64c8334 · outbound

This paper cites DiffWave: A Versatile Diffusion Model for Audio Synthesis.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DiffWave: A Versatile Diffusion Model for Audio Synthesis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.648689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.648689Z digest=sha256:8e4f4afb8a5eaf4ae65353b56f242e7f91d911a794720edee6cc68456cd250f0

Observation 8b6a760d-85ad-4ff9-82f1-5b7fb3d8b4b9 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.653080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.653080Z digest=sha256:9f5dc3a14440da0e35095e87108f9639efeda9ed761178dea76349300e6bf130

Observation 034f8e8f-3f0a-4631-a576-1e6ec998fbcf · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.657811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.657811Z digest=sha256:ffb20b9b417688409670c2e6763c2126e4da5d218d713839534e63a273a5b7ae

Observation 1c12bec7-848a-48b2-9688-16d7947e1419 · outbound

This paper cites Argmax flows and multinomial diffusion: Learning categorical distributions,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Argmax flows and multinomial diffusion: Learning categorical distributions,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.916246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.662781Z digest=sha256:37d3b29a778c0b447045c30396d33b4386810aaefa86afe64425d5aabc63f002

Observation 44f808ef-c083-4712-a0a1-88bc43b6ab56 · outbound

This paper cites Structured denoising diffusion models in discrete state-spaces,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Structured denoising diffusion models in discrete state-spaces,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.667199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.667199Z digest=sha256:3e97c59b2622d6ddac8600c686e8cb0fd105458e430e465792d50a41bb648923

Observation b51ebf07-f4ed-46b9-a81c-cb481104d6a2 · outbound

This paper cites Vector quantized diffusion model for text-to-image synthesis,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Vector quantized diffusion model for text-to-image synthesis,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.671668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.671668Z digest=sha256:45715f12e0b6558f9e17123db9ea5c96feb34e7d4a56af39428f9bfae383c96a

Observation 0700504f-a762-4558-bc46-3c00a7b318ee · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffsound: Discrete diffusion model for text-to-sound generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.676720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.676720Z digest=sha256:b6e4bd1f22e04dfd09c5d9292c8d4fa9c74b9117ef005624afeddb574ee316d9

Observation d8b673e9-7164-4ef0-8a62-2943caa4e3c1 · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAESTRO dataset,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Enabling factorized piano music modeling and generation with the MAESTRO dataset,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.867467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.682422Z digest=sha256:88c2d704408ea93edcdfef42cb84f06403a0683ac738c7c56e80257eeb20bbc1

Observation 8641f574-4eb7-4755-98af-4ae0b7f0390e · outbound

This paper cites mir eval: A transparent implementation of common MIR metrics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription mir eval: A transparent implementation of common MIR metrics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.845242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.687450Z digest=sha256:5468cce740aa2336df7d2e404caa3f6ba3d76ff6c1cc704ff9e51a3e61764709

Pith citing papers

Observation ac742e93-fbcf-4d93-9d8e-b3d11105d932 · inbound

Quantum Algorithm Software for Condensed Matter Physics cites this paper.

Quantum Algorithm Software for Condensed Matter Physics D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

Reference 172

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:54:36.856769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T04:54:36.797305Z digest=sha256:14bb06fb6229dbc16ade04f061bae7010e89c295709a37376b2c19f383d1f4b5