Pith. sign in

Paper Citation Record · LEDGER

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

As of 17 August 2026, this Paper Citation Record lists 23 of 23 outbound references and 1 inbound Pith citation observation for arXiv:2501.05068.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2501.05068 v2

Coverage vector

measured 23 of 23 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-10T21:26:24.687450Z

measured 24 of 24 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-17T06:30:58.91139+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T04:54:36.797305Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T04:54:36.849922Z

Reference resolution

23 of 23 outbound references displayed

  • verified exact1
  • verified fuzzy12
  • unresolved10
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 1a2afbe7-8d78-4bff-a1a2-5cb037c3d58b · outbound

This paper cites Automatic piano transcription with hierarchical frequency-time transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Automatic piano transcription with hierarchical frequency-time transformer,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.094203Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.575205Z digest=sha256:cfb6903281910245907edaf5bc6b20b218c4a5735c3679e1d850b400dcab5235

Observation af5aa3e3-d890-4098-a970-af7f2975a60c · outbound

This paper cites Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Scoring Time Intervals using Non-Hierarchical Transformer For Automatic Piano Transcription

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.581126Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.581126Z digest=sha256:c6f17cc910c6033033d3e9a2e029a76d2740adffc8619e5933361761bf64c184

Observation be6584f6-29c1-489a-bdd6-604c04934039 · outbound

This paper cites NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription NaturalSpeech 3: Zero-Shot Speech Synthesis with Factorized Codec and Diffusion Models

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.587527Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.587527Z digest=sha256:8f7459ed80066ee0ce2e61162089c40ad81d5c2a10a70484d88b1daaa95f3ecd

Observation 1b042cbc-17f4-43c7-9503-6237a24e4774 · outbound

This paper cites Analog bits: Generating discrete data using diffusion models with self-conditioning,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Analog bits: Generating discrete data using diffusion models with self-conditioning,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.076125Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.593978Z digest=sha256:969f09855c1544c8cd3e3fb14820849e4cfe1e7f9ce1013003531a18ad67032c

Observation 77390dac-c929-401d-bc69-9513a9ae5537 · outbound

This paper cites DFormer: Diffusion-guided Transformer for Universal Image Segmentation.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DFormer: Diffusion-guided Transformer for Universal Image Segmentation

Reference 5

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.599754Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.599754Z digest=sha256:0627ba291a3eb42f5cd5a4f59832d8f2232585a040b019e38bd955a993ced92d

Observation 3868c522-8ab1-436a-b23e-0315db7a688c · outbound

This paper cites Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffroll: Diffusion-based gen- erative music transcription with unsupervised pretraining capability,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.061433Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.605412Z digest=sha256:10798f6cfd14af071116cfe4b78fa66ee92a1576c889bab4542157525ee932c5

Observation f536c73e-5988-4d12-a8e5-ae44f208059e · outbound

This paper cites Neighborhood attention transformer,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Neighborhood attention transformer,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.045803Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.610956Z digest=sha256:e3d6e1de204e21f7b4c23935f9c8be543b8b850f02f67b83831251a822b0fd9e

Observation 26e84d95-6ffa-4905-8dce-71c6eb53f8f3 · outbound

This paper cites Onsets and frames: Dual-objective piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Onsets and frames: Dual-objective piano transcription,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.029369Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.615497Z digest=sha256:e4c2ebb0c23a96c6cf77a4797690624a67cab63821c5608b815470d51356b5de

Observation 82ac0a44-825a-49e3-a425-67540e659301 · outbound

This paper cites High-resolution piano transcription with pedals by regressing onset and offset times,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription High-resolution piano transcription with pedals by regressing onset and offset times,

Reference 9

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.620428Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.620428Z digest=sha256:7f3980aa557b8c4ab561cd90691c6c1b04b88364b58a94ee673152d9d1d4cc86

Observation 9f487184-351b-4703-9a2a-de7fda1cda97 · outbound

This paper cites Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Hppnet: Modeling the harmonic struc- ture and pitch invariance in piano transcription,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:25.001794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.625519Z digest=sha256:0c76dcad365cc423ac21f07110c2528eb5ccd46b6b308df35d1ab04fb4ea1541

Observation e1d870c3-7170-4f85-94a6-6643574d77a7 · outbound

This paper cites Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Towards Efficient and Real-Time Piano Transcription Using Neural Autoregressive Models

Reference 11

Resolution
verified exact
local_arxiv, observed 2026-08-10T21:26:24.775461Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.630524Z digest=sha256:73c45c1350ef92af6e69fb9cfc2d3c6c3dc4d3256f72403b5519f3a7f6948c01

Observation b05360f1-83a2-4d93-9a9f-39389e105218 · outbound

This paper cites Sequence-to-sequence piano transcription with transformers,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Sequence-to-sequence piano transcription with transformers,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.981643Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.635317Z digest=sha256:dc82319d11d579158bb8827d7b5956ed0306b1d5265ce04d9087c09e86329860

Observation 405d9cae-5b9c-4182-a1f4-8f4a6b0bb21f · outbound

This paper cites Piano transcription with harmonic attention,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Piano transcription with harmonic attention,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.964893Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.639520Z digest=sha256:5b4c6a68751ec711d7faf5d52dd9c2ca7207c320e5efad21008a68af0e941e89

Observation 1b0bbf07-2237-4a49-9669-c17c9a5fdf95 · outbound

This paper cites Polyphonic piano transcription using autoregressive multi-state note model,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Polyphonic piano transcription using autoregressive multi-state note model,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.947851Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.644113Z digest=sha256:ef986edadaabee811eae0e9e9072d4167e9e4337607e9569d70b8de4eaf0a1d1

Observation 3f6e7a02-632e-4459-8331-9483b64c8334 · outbound

This paper cites DiffWave: A Versatile Diffusion Model for Audio Synthesis.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription DiffWave: A Versatile Diffusion Model for Audio Synthesis

Reference 15

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.648689Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.648689Z digest=sha256:92b4374765a075d71e72cb29e4e7661663c6bad0473a9bc03a1e34175bacb483

Observation 8b6a760d-85ad-4ff9-82f1-5b7fb3d8b4b9 · outbound

This paper cites WaveNet: A Generative Model for Raw Audio.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription WaveNet: A Generative Model for Raw Audio

Reference 16

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.653080Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.653080Z digest=sha256:fe1bdcf5e8f257a57b9a2eeb2ee7a5f8e45f12d899ddf2b97e6fbd85a0afc411

Observation 034f8e8f-3f0a-4631-a576-1e6ec998fbcf · outbound

This paper cites Deep unsupervised learning using nonequilibrium thermodynamics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Deep unsupervised learning using nonequilibrium thermodynamics,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.657811Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.657811Z digest=sha256:f0b40b6a4ae3c7737599bd3813d42303a53b875130b9e6498a8a56280703c3c9

Observation 1c12bec7-848a-48b2-9688-16d7947e1419 · outbound

This paper cites Argmax flows and multinomial diffusion: Learning categorical distributions,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Argmax flows and multinomial diffusion: Learning categorical distributions,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.916246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.662781Z digest=sha256:35daea501831e6f001654d788621ff2b3bb147df3989a7545861b422dfb3aa53

Observation 44f808ef-c083-4712-a0a1-88bc43b6ab56 · outbound

This paper cites Structured denoising diffusion models in discrete state-spaces,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Structured denoising diffusion models in discrete state-spaces,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.667199Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.667199Z digest=sha256:df2ba794a17c01fb2c6ee81eeae0b4e894a019126e5764988b753d3f9b34fc1d

Observation b51ebf07-f4ed-46b9-a81c-cb481104d6a2 · outbound

This paper cites Vector quantized diffusion model for text-to-image synthesis,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Vector quantized diffusion model for text-to-image synthesis,

Reference 20

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.671668Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.671668Z digest=sha256:41d309a775add773f1a8bab70d4319836febdb83403142a1a26a21bd7eb96560

Observation 0700504f-a762-4558-bc46-3c00a7b318ee · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Diffsound: Discrete diffusion model for text-to-sound generation,

Reference 21

Resolution
unresolved
no resolver link, observed 2026-08-10T21:26:24.676720Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-10T21:26:24.676720Z digest=sha256:8521f384e0da30220e74988c2cd1a19293737611e519a5b6140764d0be158412

Observation d8b673e9-7164-4ef0-8a62-2943caa4e3c1 · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAESTRO dataset,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription Enabling factorized piano music modeling and generation with the MAESTRO dataset,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.867467Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.682422Z digest=sha256:6d1dcc7d9c271a5b3ac5a237330f778fa6de9d805fc6114a653c7ad29c0dd365

Observation 8641f574-4eb7-4755-98af-4ae0b7f0390e · outbound

This paper cites mir eval: A transparent implementation of common MIR metrics,.

D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription mir eval: A transparent implementation of common MIR metrics,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-10T21:26:24.845242Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-10T21:26:24.687450Z digest=sha256:7a00d330558ccfaa6bf9d8ac8ed20da00cedc6beb0a5051ac22eac02230cee24

Pith citing papers

Observation ac742e93-fbcf-4d93-9d8e-b3d11105d932 · inbound

Quantum Algorithm Software for Condensed Matter Physics cites this paper.

Quantum Algorithm Software for Condensed Matter Physics D3RM: A Discrete Denoising Diffusion Refinement Model for Piano Transcription

Reference 172

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T04:54:36.856769Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-17T06:30:58.91139+00:00.

source=pdf_text observed=2026-08-07T04:54:36.797305Z digest=sha256:9935b7e4d195a90b8c9fa15b2700656a7d62169e890dc659a1ac194d0446d5a8