Pith. sign in

Paper Citation Record · LEDGER

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

As of 13 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2501.05068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05068 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:26:24.687450Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-12T06:34:41.77262+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:54:36.797305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:54:36.849922Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a2afbe7-8d78-4bff-a1a2-5cb037c3d58b · outbound

This paper cites Automatic piano transcription with hierarchical frequency-time transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Automatic piano transcription with hierarchical frequency-time transformer,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.094203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.575205Z digest=sha256:4303fe4d030c571085976bbf13bf92e3344bc0645b42c6e6223a71c0fbb33fe7

Observation af5aa3e3-d890-4098-a970-af7f2975a60c · outbound

This paper cites Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.581126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.581126Z digest=sha256:db3ca433fefff68843fab9457becf8994bec815926b0c92c7c2c113ef66491b2

Observation be6584f6-29c1-489a-bdd6-604c04934039 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.587527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.587527Z digest=sha256:caf6808fc1061a0aa18e13a4d147212b8b0afb9e2b19a26abe068bef594b0356

Observation 1b042cbc-17f4-43c7-9503-6237a24e4774 · outbound

This paper cites Analog bits: Generating discrete data using diffusion models with self-conditioning,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Analog bits: Generating discrete data using diffusion models with self-conditioning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.076125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.593978Z digest=sha256:9d4d10aa625bdf9d0d55e5b319ef0f575a619f87b05cc4381651206acda3f7ce

Observation 77390dac-c929-401d-bc69-9513a9ae5537 · outbound

This paper cites DFormer: Diffusion-guided Transformer for Universal Image Segmentation.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DFormer: Diffusion-guided Transformer for Universal Image Segmentation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.599754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.599754Z digest=sha256:a2c97cbd6e35556d3fe2a69ca251ca6efa54753a777575847138989f800fae79

Observation 3868c522-8ab1-436a-b23e-0315db7a688c · outbound

This paper cites Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.061433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.605412Z digest=sha256:227604617f1c2b9352c7d54c0dd2340e8a84edd91a72ed0bcc6374f0d97fa375

Observation f536c73e-5988-4d12-a8e5-ae44f208059e · outbound

This paper cites Neighborhood attention transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Neighborhood attention transformer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.045803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.610956Z digest=sha256:ae5c2311f576e82ca618ac84cd41a700976962689c17cf066252d4345e723371

Observation 26e84d95-6ffa-4905-8dce-71c6eb53f8f3 · outbound

This paper cites Onsets and frames: Dual-objective piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Onsets and frames: Dual-objective piano transcription,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.029369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.615497Z digest=sha256:bebb6c701aa983af80cb0d84623a37dcf7da0f647f54030cca957a4bfdbdbb5f

Observation 82ac0a44-825a-49e3-a425-67540e659301 · outbound

This paper cites High-resolution piano transcription with pedals by regressing onset and offset times,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription High-resolution piano transcription with pedals by regressing onset and offset times,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.620428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.620428Z digest=sha256:2467f2f3cfccd07d053eb85ee5ede3ac561e8cf4a0d5ef318080ea2060e06426

Observation 9f487184-351b-4703-9a2a-de7fda1cda97 · outbound

This paper cites Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.001794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.625519Z digest=sha256:30b974dd250a4b10abc54b29803f818f70712a20603c7009e6564ddf936d4d27

Observation e1d870c3-7170-4f85-94a6-6643574d77a7 · outbound

This paper cites Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:26:24.775461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.630524Z digest=sha256:5ff2bb466deb229b277c11f4f413b3df4e39684a7b378571b20a016c8d63e58a

Observation b05360f1-83a2-4d93-9a9f-39389e105218 · outbound

This paper cites Sequence-to-sequence piano transcription with transformers,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Sequence-to-sequence piano transcription with transformers,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.981643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.635317Z digest=sha256:575293e4b374dd984354a4eabdbff0a82041d252f55ab0a42297baa568b27bd4

Observation 405d9cae-5b9c-4182-a1f4-8f4a6b0bb21f · outbound

This paper cites Piano transcription with harmonic attention,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Piano transcription with harmonic attention,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.964893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.639520Z digest=sha256:775f511692d62347090af91480a0a7f77c97fb5aefe723e215a87ea5b7787856

Observation 1b0bbf07-2237-4a49-9669-c17c9a5fdf95 · outbound

This paper cites Polyphonic piano transcription using autoregressive multi-state note model,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Polyphonic piano transcription using autoregressive multi-state note model,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.947851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.644113Z digest=sha256:09a933bc48526fae7785a9d218a59ec8ddc469fe1606b1d355e0e5c2e9c40e8f

Observation 3f6e7a02-632e-4459-8331-9483b64c8334 · outbound

This paper cites DiffWave: A Versatile Diffusion Model for Audio Synthesis.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DiffWave: A Versatile Diffusion Model for Audio Synthesis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.648689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.648689Z digest=sha256:8458be175019b9f7336fa4166da8fdb82d5b1d785cf473afcddcfdb27d4a23db

Observation 8b6a760d-85ad-4ff9-82f1-5b7fb3d8b4b9 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.653080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.653080Z digest=sha256:d632e1bf1c34bc0d69c97175a16b895b3d6390b77f12c4c4537586776a15602c

Observation 034f8e8f-3f0a-4631-a576-1e6ec998fbcf · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.657811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.657811Z digest=sha256:b9bf80f974d96eded1b225b5e7b5f4633975312d142eea56029f97e58220443a

Observation 1c12bec7-848a-48b2-9688-16d7947e1419 · outbound

This paper cites Argmax flows and multinomial diffusion: Learning categorical distributions,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Argmax flows and multinomial diffusion: Learning categorical distributions,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.916246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.662781Z digest=sha256:0473e5cb06d0338f4d749cc8851878ba0a1577470765a5231ee80c5bb3a0980c

Observation 44f808ef-c083-4712-a0a1-88bc43b6ab56 · outbound

This paper cites Structured denoising diffusion models in discrete state-spaces,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Structured denoising diffusion models in discrete state-spaces,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.667199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.667199Z digest=sha256:96d8ce17d7564b52a62894a969d8af82205020e35b2ecaa9b2317ccc572a3628

Observation b51ebf07-f4ed-46b9-a81c-cb481104d6a2 · outbound

This paper cites Vector quantized diffusion model for text-to-image synthesis,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Vector quantized diffusion model for text-to-image synthesis,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.671668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.671668Z digest=sha256:467314765e2588373b564b317aff4c631c825765c00ea4b1661a1e5bec644b33

Observation 0700504f-a762-4558-bc46-3c00a7b318ee · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffsound: Discrete diffusion model for text-to-sound generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.676720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.676720Z digest=sha256:cf3b3d81a8feac0d138559b3c7fca088d704d7a136f16d42e25921b31aca35b5

Observation d8b673e9-7164-4ef0-8a62-2943caa4e3c1 · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAESTRO dataset,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Enabling factorized piano music modeling and generation with the MAESTRO dataset,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.867467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.682422Z digest=sha256:2840bcaaa5adacdb8731b39131841fdaea11be3c8f624b0bb97f7fe18711c460

Observation 8641f574-4eb7-4755-98af-4ae0b7f0390e · outbound

This paper cites mir eval: A transparent implementation of common MIR metrics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription mir eval: A transparent implementation of common MIR metrics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.845242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-10T21:26:24.687450Z digest=sha256:c7a851e12da91abf0f847481e8a1c4ece9e3eb6d69972d9a1c7e32c438ccb05e

Pith citing papers

Observation ac742e93-fbcf-4d93-9d8e-b3d11105d932 · inbound

Quantum Algorithm Software for Condensed Matter Physics cites this paper.

Quantum Algorithm Software for Condensed Matter Physics D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

Reference 172

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:54:36.856769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-12T06:34:41.77262+00:00.

source=pdf_text observed=2026-08-07T04:54:36.797305Z digest=sha256:74e0e666a4980b48c5422a624405b61f584fca2aefded5f1c069615574dfdf8e