Pith. sign in

Paper Citation Record · LEDGER

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

As of 9 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2507.08530.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08530 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:21:15.102583Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:21:14.926845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T03:16:19.153331Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4dc0ed6d-f6a3-4563-a897-3a0ddb9912d9 · outbound

This paper cites MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.926845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.926845Z digest=sha256:b23a94c6cb007f6861d789ad7c62850631e343faa59698722fe50876780f7f72

Observation adc50779-db5a-4d13-b79b-f53a95f4dfe7 · outbound

This paper cites These TTS-inspired models typically process piano performance MIDIs as pi- ano rolls for audio synthesis.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling These TTS-inspired models typically process piano performance MIDIs as pi- ano rolls for audio synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.839116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.932050Z digest=sha256:21e4893f5e458dee6d979e088c0321f4f5a57435a526f598275f93ca064eafe5

Observation cc66a967-88e0-41f2-af75-4355aec2f94d · outbound

This paper cites an unresolved cited work.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:21:15.822756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.936140Z digest=sha256:76ce57db19f97d2478344672dbc876e5617d0d1f2f945702662d53e5daf33e39

Observation eb9e0395-746b-45ca-afa9-b6fd87f4a6a7 · outbound

This paper cites A total of 8,825 perfor- mance recordings were selected and split into training, val- idation, and test sets in an 8:1:1 ratio.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling A total of 8,825 perfor- mance recordings were selected and split into training, val- idation, and test sets in an 8:1:1 ratio

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.808049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.940657Z digest=sha256:42c0009de69bc2a01e297eceaa84a0fb840dde01f7208fa599737b38e7faf986

Observation 27534c76-e234-4e62-84f5-2a234bdac3b1 · outbound

This paper cites FAD measures the percep- tual quality and realism of generated audio by comparing it to reference performances using embeddings extracted from Piano-Encodec.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling FAD measures the percep- tual quality and realism of generated audio by comparing it to reference performances using embeddings extracted from Piano-Encodec

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.792345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.945077Z digest=sha256:feab746678d5b29d223cece4f4b5d9b87f3e2c07e8ceafdd9ddbba774b8a75e7

Observation c2412f64-3be6-4e97-a5c9-07d5832c08d8 · outbound

This paper cites In addition, Piano-Encodec achieves high-fidelity reconstruction of human performances, with much lower FAD, spectrogram, and chroma distortions than generative models.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling In addition, Piano-Encodec achieves high-fidelity reconstruction of human performances, with much lower FAD, spectrogram, and chroma distortions than generative models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.777406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.949035Z digest=sha256:1f5325798a0df66f1d7de58a8ecb263e1b61b732f7bb4684097759bd9cd9c617

Observation 2d1f7a27-5078-427c-8ab1-ea8b2983983e · outbound

This paper cites an unresolved cited work.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:21:15.762884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.953434Z digest=sha256:18aca95c0f871cd05a3efc056ca95b52905d1619ce875cc6add5c03c65b0c58d

Observation fda4af9b-ca63-4fbf-a324-b04038c80ea8 · outbound

This paper cites Onderzoeksprogramma Artificiële Intelli- gentie (AI) Vlaanderen.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Onderzoeksprogramma Artificiële Intelli- gentie (AI) Vlaanderen

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.747605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.958039Z digest=sha256:90e3d031f353117a7a2dd54a13a3468c159c8c810b1519895da6cdac186064b2

Observation d2c57a8b-2e56-4540-9ab4-adeafd1b3669 · outbound

This paper cites The datasets used in this study — ATEPP [8], Mae- stro [6], and Pijama [31] — contain audio recordings and corresponding MIDI annotations of piano performances.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling The datasets used in this study — ATEPP [8], Mae- stro [6], and Pijama [31] — contain audio recordings and corresponding MIDI annotations of piano performances

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.732182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.961825Z digest=sha256:8c881e270e57142b3af75efc20e1f174bbe3ae7b3a0dc7f7498b8fa82e95d20b

Observation 4aa43205-f3d7-4862-b2a6-7e7580c5704a · outbound

This paper cites MIDI-DDSP: Detailed control of musical per- formance via hierarchical modeling,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-DDSP: Detailed control of musical per- formance via hierarchical modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.716742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.965754Z digest=sha256:aacac99175816a4950c3dcf33f0399dcb4d85b8cd8a21fa78140822bc91c042c

Observation f4f5dddb-1725-4c0c-bf4d-77f39716a44f · outbound

This paper cites Deep performer: Score-to-audio music performance synthesis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Deep performer: Score-to-audio music performance synthesis,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.702118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.969627Z digest=sha256:a8ff3a5f127c0d26175f85b64a0bd80a5547d48d53a0996bc945d4763d955ff0

Observation 27764fd2-095c-4163-bfad-95a642a9755e · outbound

This paper cites Towards an integrated approach for expressive piano performance synthesis from music scores,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Towards an integrated approach for expressive piano performance synthesis from music scores,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.687524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.973662Z digest=sha256:917026f25cd415d7643b0346b14b47e6401576b628b8571ccf75d887fb946cde

Observation 62644193-c50b-4dc6-9a8c-5342284a8c09 · outbound

This paper cites Text-to- speech synthesis techniques for midi-to-audio synthe- sis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Text-to- speech synthesis techniques for midi-to-audio synthe- sis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.671093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.978451Z digest=sha256:fc705021ffc3f21fa0778593845fb8cc21608ede043a638dd6eb2f3a75983f20

Observation 4dd5afd7-4896-4cf9-b881-257f344aa412 · outbound

This paper cites Can knowledge of end-to-end text-to- speech models improve neural midi-to-audio synthesis systems?.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Can knowledge of end-to-end text-to- speech models improve neural midi-to-audio synthesis systems?

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.656402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.982858Z digest=sha256:3350d398758a1a08bd72a33fd507f4d34d2150a541a4b2073903fe5413626480

Observation 7572fad2-59eb-4d17-88fb-5a09824af55f · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAE- STRO dataset,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Enabling factorized piano music modeling and generation with the MAE- STRO dataset,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.642388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.987761Z digest=sha256:00ce3c3e27cd8452c2ccc14af312b7cd7ece08caaf512b36c07eb75c86c98400

Observation 2cd6a077-228d-4003-bff8-099abee8182e · outbound

This paper cites Neural codec language models are zero-shot text to speech synthesizers,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Neural codec language models are zero-shot text to speech synthesizers,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.627133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.991502Z digest=sha256:61bc462344419a6a555fe5951e8806ef1da52c8f5922f9206333e9600ec0bd73

Observation 5b9a38f2-2168-48cc-91bb-043c0c868a2b · outbound

This paper cites ATEPP: A Dataset of Auto- matically Transcribed Expressive Piano Performance,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling ATEPP: A Dataset of Auto- matically Transcribed Expressive Piano Performance,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.995391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.995391Z digest=sha256:8edac17ae81ac8196f84c36779d4449d2b8930b8d4678cce529cb92f49ceabc4

Observation ecfb31d1-da3c-4ae6-afca-123faa57f43d · outbound

This paper cites MusicBERT: Symbolic music understanding with large-scale pre-training,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MusicBERT: Symbolic music understanding with large-scale pre-training,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.612528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:14.999201Z digest=sha256:6bf505f1d3d62bc1d50284a373310f4d3372a08835f346341ef977307ec7d119

Observation 50f4c473-84bf-4fe9-8a1c-2c05c206074e · outbound

This paper cites High fidelity neural audio compression,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling High fidelity neural audio compression,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.005210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.005210Z digest=sha256:dea7dc61f6c7beb4d8040e85cedfa7e442dbe3facdf3e000b0027cdea2bab7e0

Observation 8ba431e2-887b-4c82-8bad-af5912692511 · outbound

This paper cites DDSP-Piano: a Neural Sound Synthesizer Informed by Instrument Knowledge,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling DDSP-Piano: a Neural Sound Synthesizer Informed by Instrument Knowledge,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.587666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.010760Z digest=sha256:70c51fbb3fb058727e4d94ad7e4190d810f226d6b457d38ac0800bdfa7276ec2

Observation c66a0b4f-0232-427d-ab43-b2d351839119 · outbound

This paper cites Neural speech synthesis with transformer network,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Neural speech synthesis with transformer network,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.573259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.016055Z digest=sha256:c48054e3c52b6ba461b998b2f4eb52509654bddd32f99822a80721a23053bfb2

Observation c57e03a9-6f1b-475d-a133-1114f1828a73 · outbound

This paper cites Fastspeech: fast, robust and control- lable text to speech,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Fastspeech: fast, robust and control- lable text to speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.559542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.021278Z digest=sha256:22f9e146bbf24c1fb259b14dcc55193408d3ed9b647423d4f463fc0f13586bde

Observation a50a33c3-6277-42bf-b1dc-9523d6db5521 · outbound

This paper cites Hifi-gan: Generative ad- versarial networks for efficient and high fidelity speech synthesis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Hifi-gan: Generative ad- versarial networks for efficient and high fidelity speech synthesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.545553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.025672Z digest=sha256:03380aa585ceb97c0899b597f7b19fe417f98647a6ea2054b99f2e3e0096d248

Observation 6697cc0b-7c68-44f0-b26c-20a6682421d6 · outbound

This paper cites Reconstructing human expressiveness in piano performances with a transformer network,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Reconstructing human expressiveness in piano performances with a transformer network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.530918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.030143Z digest=sha256:84c1abf6b0b0a523b2cc6034983696125d9f48a5fbde7fc417822ae8cdaccc7c

Observation 67f07ad6-6521-427f-9cdf-d152cf24c606 · outbound

This paper cites Scoreperformer: Expressive piano performance rendering with fine-grained con- trol.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Scoreperformer: Expressive piano performance rendering with fine-grained con- trol

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.516875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.034566Z digest=sha256:170d6edbb75557b57d983ca5b088d69527956eb5bc9924cfbc2682badc323e08

Observation 76354753-295a-401e-9f3e-82ceace3621d · outbound

This paper cites Expressive Piano Performance Rendering from Unpaired Data,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Expressive Piano Performance Rendering from Unpaired Data,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.503199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.039254Z digest=sha256:1faceef49e76b723c49afd53eeb36b335c787ad18d3946f636b4f07f02dd0566

Observation 7271a73b-0ec9-4782-9e04-ae7946876da8 · outbound

This paper cites Vir- tuosonet: A hierarchical rnn-based system for model- ing expressive piano performance,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vir- tuosonet: A hierarchical rnn-based system for model- ing expressive piano performance,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.488767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.044403Z digest=sha256:1599a18c365649a0dd67648cd008ba60eec0052ea0a769b3cf660c97b95e419a

Observation 0618346e-3744-4a64-ad42-61e1367d43eb · outbound

This paper cites Dexter: Learning and controlling performance expression with diffusion models,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Dexter: Learning and controlling performance expression with diffusion models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.473673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.049249Z digest=sha256:3cf80720915b2082659e9718319311ef2740e6bc462ef209a06f5c3c17d54e26

Observation 8165f67f-55ba-4f53-b50c-38912f552028 · outbound

This paper cites Audiolm: a language modeling approach to audio generation,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Audiolm: a language modeling approach to audio generation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.459335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.053958Z digest=sha256:50b176fbb26e0a6487ff1a94cf58dd70b72bc6a14a486dbd92da28c0ef887e9a

Observation e65589ce-34d2-4df4-bb2d-70ad276aa15d · outbound

This paper cites Simple and controllable music generation,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Simple and controllable music generation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.445673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.058705Z digest=sha256:fc510e504409de63bd0dc1b732d0857c4dc0da23bb10346cd07964c32ec4242c

Observation 74fb4272-7b1a-4ead-ae2b-a82c507d54d6 · outbound

This paper cites MusicLM: Generating Music From Text.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MusicLM: Generating Music From Text

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.063353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.063353Z digest=sha256:6f2fd10257fea89797ea61a91389b93db687ea5de23bc1f72b0a8c1564a58b24

Observation 9366f862-42ee-4345-91de-59f34ec595fd · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Soundstream: An end-to-end neural audio codec,

Reference 32

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:21:15.067823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.067823Z digest=sha256:32e12287e8bd48667db25c44559062e28bdd8b3369081d9eadb4721cd9a58ea8

Observation 17d6f87b-6fab-462b-b978-5440b6a6b8f0 · outbound

This paper cites Vector quantization,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vector quantization,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.430826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.072429Z digest=sha256:8ec893406527964f814188fb0485a0240dd88a3b59f4ae5fc7ae873aa058b7fb

Observation 9cd49da2-a945-49a8-a346-574c74059506 · outbound

This paper cites Vall-e: A neural codec language model,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vall-e: A neural codec language model,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.417057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.076464Z digest=sha256:a2bd1a0ed0080b5e2340af202e43a3a0c18a296aa62c55b830133e052c362ddf

Observation eb238fd6-9adb-4e49-abec-7ae058f79ec2 · outbound

This paper cites Compound word transformer: Learning to compose full-song music over dynamic directed hypergraphs,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Compound word transformer: Learning to compose full-song music over dynamic directed hypergraphs,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.080822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.080822Z digest=sha256:a9853a61e5e16383e70f693edbe53fe60cb34ba8506c97623ef467738e8a0e6c

Observation 831b8157-ed7a-4379-aff7-c48e5905d790 · outbound

This paper cites Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.085266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.085266Z digest=sha256:b9978e62881da5b93e1af059e04aeeeb28d09c7ea20744d224fbc89de9ea1984

Observation 4b62de7a-2490-43d3-b983-0e8cd0858188 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Zipformer: A faster and better encoder for automatic speech recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.392896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.090118Z digest=sha256:8a0ca1809582ec79a96703dcbf418869ff04ce2458195245a88c3739405f062b

Observation de143774-171d-48c6-b390-dc9b22105721 · outbound

This paper cites Fréchet audio distance: A reference-free metric for evaluating music enhancement algorithms,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Fréchet audio distance: A reference-free metric for evaluating music enhancement algorithms,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.378393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.094347Z digest=sha256:42126137cc34086de44e2cf1b8c6f2bb2a739c92c537ddbcd78ccf11e8b953c3

Observation 53affdc1-ad3a-4d62-94e5-28260e76e17f · outbound

This paper cites Adapting Frechet Audio Distance for Generative Music Evaluation.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Adapting Frechet Audio Distance for Generative Music Evaluation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.098354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.098354Z digest=sha256:d23d47f7f2d714b63e4493634008aa14a5fcd58253d04b755aa77d45b8549113

Observation 13f5afd0-517e-48dc-a994-faca780ef80e · outbound

This paper cites Pijama: Pi- ano jazz with automatic midi annotations,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Pijama: Pi- ano jazz with automatic midi annotations,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.362987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-06T18:21:15.102583Z digest=sha256:989b18e2a68a33a8ceb008fb30fbae1b96d0fe5978b2bb96580d44064952f71d

Pith citing papers

Observation 4dc0ed6d-f6a3-4563-a897-3a0ddb9912d9 · inbound

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling cites this paper.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.926845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.926845Z digest=sha256:b23a94c6cb007f6861d789ad7c62850631e343faa59698722fe50876780f7f72

Observation 3e9631ad-6b7c-4667-a14e-9888b00353a3 · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:19.155184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:b68379ed418dd30de0c1dba80d045a9a77481bdf9d62dfadd03119bf9e217e53