Pith. sign in

Paper Citation Record · LEDGER

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

As of 7 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 2 inbound Pith citation observations for arXiv:2507.08530.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2507.08530 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:21:15.102583Z

measured 42 of 42 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-06T06:34:29.942622+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-06T18:21:14.926845Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-05-12T03:16:19.153331Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact0
  • verified fuzzy30
  • unresolved9
  • parse uncertain0
  • malformed identifier1
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 4dc0ed6d-f6a3-4563-a897-3a0ddb9912d9 · outbound

This paper cites MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.926845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.926845Z digest=sha256:b23a94c6cb007f6861d789ad7c62850631e343faa59698722fe50876780f7f72

Observation adc50779-db5a-4d13-b79b-f53a95f4dfe7 · outbound

This paper cites These TTS-inspired models typically process piano performance MIDIs as pi- ano rolls for audio synthesis.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling These TTS-inspired models typically process piano performance MIDIs as pi- ano rolls for audio synthesis

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.839116Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.932050Z digest=sha256:943d67e694d05cb146ffcc1deeb8d926f543a4ca08e2d683367c1e9fdc5a85bd

Observation cc66a967-88e0-41f2-af75-4355aec2f94d · outbound

This paper cites an unresolved cited work.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:21:15.822756Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.936140Z digest=sha256:2092a08528ed613fffd24dca51c804c2765e7dc47fdfba02a1cf7ebd25e48f47

Observation eb9e0395-746b-45ca-afa9-b6fd87f4a6a7 · outbound

This paper cites A total of 8,825 perfor- mance recordings were selected and split into training, val- idation, and test sets in an 8:1:1 ratio.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling A total of 8,825 perfor- mance recordings were selected and split into training, val- idation, and test sets in an 8:1:1 ratio

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.808049Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.940657Z digest=sha256:d4c0d13853ba4a0775fe83b7415a38616acbd6f0dda54cc54c42366bdb57868d

Observation 27534c76-e234-4e62-84f5-2a234bdac3b1 · outbound

This paper cites FAD measures the percep- tual quality and realism of generated audio by comparing it to reference performances using embeddings extracted from Piano-Encodec.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling FAD measures the percep- tual quality and realism of generated audio by comparing it to reference performances using embeddings extracted from Piano-Encodec

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.792345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.945077Z digest=sha256:433449e4222cb12bb88e9fbad6d5d4a3f02e1fd09d0ef457038df62e1b4f8684

Observation c2412f64-3be6-4e97-a5c9-07d5832c08d8 · outbound

This paper cites In addition, Piano-Encodec achieves high-fidelity reconstruction of human performances, with much lower FAD, spectrogram, and chroma distortions than generative models.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling In addition, Piano-Encodec achieves high-fidelity reconstruction of human performances, with much lower FAD, spectrogram, and chroma distortions than generative models

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.777406Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.949035Z digest=sha256:4c1c33c9d8d27f6d47ebe1cbb23c78d69fca7a73d0b4a68e1e597afd164ae2d5

Observation 2d1f7a27-5078-427c-8ab1-ea8b2983983e · outbound

This paper cites an unresolved cited work.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Unresolved cited work

Reference 7

Resolution
unresolved
raw_fallback, observed 2026-08-06T18:21:15.762884Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.953434Z digest=sha256:23df643a3b764d38c48bd1690bf9bf3071091903e42134a3cc858053ea32d7a8

Observation fda4af9b-ca63-4fbf-a324-b04038c80ea8 · outbound

This paper cites Onderzoeksprogramma Artificiële Intelli- gentie (AI) Vlaanderen.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Onderzoeksprogramma Artificiële Intelli- gentie (AI) Vlaanderen

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.747605Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.958039Z digest=sha256:c31aeac40b755ee6c4c16945ef1a85b88a1018353edd071e8b7bb05b22a44861

Observation d2c57a8b-2e56-4540-9ab4-adeafd1b3669 · outbound

This paper cites The datasets used in this study — ATEPP [8], Mae- stro [6], and Pijama [31] — contain audio recordings and corresponding MIDI annotations of piano performances.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling The datasets used in this study — ATEPP [8], Mae- stro [6], and Pijama [31] — contain audio recordings and corresponding MIDI annotations of piano performances

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.732182Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.961825Z digest=sha256:aff2918c3f945031ba3280d9abf213b897e2b9d546e6ae2d3a0d1db0d832b876

Observation 4aa43205-f3d7-4862-b2a6-7e7580c5704a · outbound

This paper cites MIDI-DDSP: Detailed control of musical per- formance via hierarchical modeling,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-DDSP: Detailed control of musical per- formance via hierarchical modeling,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.716742Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.965754Z digest=sha256:65a1d4bb252a1d5ed304e3fdddc644b82bd73f5d72097aee02111df57953edec

Observation f4f5dddb-1725-4c0c-bf4d-77f39716a44f · outbound

This paper cites Deep performer: Score-to-audio music performance synthesis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Deep performer: Score-to-audio music performance synthesis,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.702118Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.969627Z digest=sha256:10ab3f42e5ab5553488d2aae90e518987951b1aab527da37e7ebe1d14c6ed47c

Observation 27764fd2-095c-4163-bfad-95a642a9755e · outbound

This paper cites Towards an integrated approach for expressive piano performance synthesis from music scores,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Towards an integrated approach for expressive piano performance synthesis from music scores,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.687524Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.973662Z digest=sha256:69cfb63debf1526e75829eeb878ee9d5a1e8f1aba0d9c552ac1fb8e6938ebba2

Observation 62644193-c50b-4dc6-9a8c-5342284a8c09 · outbound

This paper cites Text-to- speech synthesis techniques for midi-to-audio synthe- sis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Text-to- speech synthesis techniques for midi-to-audio synthe- sis,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.671093Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.978451Z digest=sha256:d2911caa4b56c372b41c37cfecf62fdf1341b2f34606823263d331c5421b4be7

Observation 4dd5afd7-4896-4cf9-b881-257f344aa412 · outbound

This paper cites Can knowledge of end-to-end text-to- speech models improve neural midi-to-audio synthesis systems?.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Can knowledge of end-to-end text-to- speech models improve neural midi-to-audio synthesis systems?

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.656402Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.982858Z digest=sha256:18a9fd08b65a55320d7c582d8a915fdd6d18af5410900d2955b0db7417b106f6

Observation 7572fad2-59eb-4d17-88fb-5a09824af55f · outbound

This paper cites Enabling factorized piano music modeling and generation with the MAE- STRO dataset,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Enabling factorized piano music modeling and generation with the MAE- STRO dataset,

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.642388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.987761Z digest=sha256:d1b11ab8c1c628f400ef1a7c2fa8ce81a7eb27f5604db19ddcd82c39ed740082

Observation 2cd6a077-228d-4003-bff8-099abee8182e · outbound

This paper cites Neural codec language models are zero-shot text to speech synthesizers,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Neural codec language models are zero-shot text to speech synthesizers,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.627133Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.991502Z digest=sha256:f11f186e4a4db5135af98a888f2735ddc9c0e7eefefd0ba495f3e02ab7461e82

Observation 5b9a38f2-2168-48cc-91bb-043c0c868a2b · outbound

This paper cites ATEPP: A Dataset of Auto- matically Transcribed Expressive Piano Performance,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling ATEPP: A Dataset of Auto- matically Transcribed Expressive Piano Performance,

Reference 17

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.995391Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.995391Z digest=sha256:8edac17ae81ac8196f84c36779d4449d2b8930b8d4678cce529cb92f49ceabc4

Observation ecfb31d1-da3c-4ae6-afca-123faa57f43d · outbound

This paper cites MusicBERT: Symbolic music understanding with large-scale pre-training,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MusicBERT: Symbolic music understanding with large-scale pre-training,

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.612528Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:14.999201Z digest=sha256:447438d07048e918f8dff29de30b97e33907cafed3b826943cadc1297fe6e7a3

Observation 50f4c473-84bf-4fe9-8a1c-2c05c206074e · outbound

This paper cites High fidelity neural audio compression,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling High fidelity neural audio compression,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.005210Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.005210Z digest=sha256:dea7dc61f6c7beb4d8040e85cedfa7e442dbe3facdf3e000b0027cdea2bab7e0

Observation 8ba431e2-887b-4c82-8bad-af5912692511 · outbound

This paper cites DDSP-Piano: a Neural Sound Synthesizer Informed by Instrument Knowledge,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling DDSP-Piano: a Neural Sound Synthesizer Informed by Instrument Knowledge,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.587666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.010760Z digest=sha256:f7e41a3570acdc0f22257b9fd27f2d28d703b10b193d7c3aea4c7d239b6df477

Observation c66a0b4f-0232-427d-ab43-b2d351839119 · outbound

This paper cites Neural speech synthesis with transformer network,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Neural speech synthesis with transformer network,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.573259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.016055Z digest=sha256:4732a7446a4ee4c9bb06f39e11a2283bb5cb130e2a8414bc517b8c4333dfd692

Observation c57e03a9-6f1b-475d-a133-1114f1828a73 · outbound

This paper cites Fastspeech: fast, robust and control- lable text to speech,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Fastspeech: fast, robust and control- lable text to speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.559542Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.021278Z digest=sha256:791a00690b393749acb71ef096a78056b54a0504f11e4673085412dacb2d6dd9

Observation a50a33c3-6277-42bf-b1dc-9523d6db5521 · outbound

This paper cites Hifi-gan: Generative ad- versarial networks for efficient and high fidelity speech synthesis,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Hifi-gan: Generative ad- versarial networks for efficient and high fidelity speech synthesis,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.545553Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.025672Z digest=sha256:a342399e96dd7cf5da69e0b97b4d74cc88af71399376d2c1b680be85448d10b2

Observation 6697cc0b-7c68-44f0-b26c-20a6682421d6 · outbound

This paper cites Reconstructing human expressiveness in piano performances with a transformer network,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Reconstructing human expressiveness in piano performances with a transformer network,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.530918Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.030143Z digest=sha256:927c3c0aba5c360f076d46a00b924caaf5bdf054420e2ba6751778cb4ed0b983

Observation 67f07ad6-6521-427f-9cdf-d152cf24c606 · outbound

This paper cites Scoreperformer: Expressive piano performance rendering with fine-grained con- trol.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Scoreperformer: Expressive piano performance rendering with fine-grained con- trol

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.516875Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.034566Z digest=sha256:f2a7c3770205ca327597ccfa65817543b3817a5148a4004da85b388246af99db

Observation 76354753-295a-401e-9f3e-82ceace3621d · outbound

This paper cites Expressive Piano Performance Rendering from Unpaired Data,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Expressive Piano Performance Rendering from Unpaired Data,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.503199Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.039254Z digest=sha256:b03bc819f478f66df8c4a885c9858193c472bd684c2dc63c11c5d521cd57517b

Observation 7271a73b-0ec9-4782-9e04-ae7946876da8 · outbound

This paper cites Vir- tuosonet: A hierarchical rnn-based system for model- ing expressive piano performance,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vir- tuosonet: A hierarchical rnn-based system for model- ing expressive piano performance,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.488767Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.044403Z digest=sha256:966a6bb6d214f1c9425fba5475c7fe8c15884422ceded3861d419dd5ac747632

Observation 0618346e-3744-4a64-ad42-61e1367d43eb · outbound

This paper cites Dexter: Learning and controlling performance expression with diffusion models,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Dexter: Learning and controlling performance expression with diffusion models,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.473673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.049249Z digest=sha256:0c48f892f052bbd8ee8279be11bb2205148b6c20729288e26317f484a0606140

Observation 8165f67f-55ba-4f53-b50c-38912f552028 · outbound

This paper cites Audiolm: a language modeling approach to audio generation,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Audiolm: a language modeling approach to audio generation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.459335Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.053958Z digest=sha256:77cf2403e110d163abbc7faa3c1fc83f479f6636d4d0496106beb27da46c2c91

Observation e65589ce-34d2-4df4-bb2d-70ad276aa15d · outbound

This paper cites Simple and controllable music generation,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Simple and controllable music generation,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.445673Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.058705Z digest=sha256:ef08f068f3bc92a0b6891d6c072d13dbfeb0286c9193c7760d7e744a7eca8658

Observation 74fb4272-7b1a-4ead-ae2b-a82c507d54d6 · outbound

This paper cites MusicLM: Generating Music From Text.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MusicLM: Generating Music From Text

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.063353Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.063353Z digest=sha256:6f2fd10257fea89797ea61a91389b93db687ea5de23bc1f72b0a8c1564a58b24

Observation 9366f862-42ee-4345-91de-59f34ec595fd · outbound

This paper cites Soundstream: An end-to-end neural audio codec,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Soundstream: An end-to-end neural audio codec,

Reference 32

Resolution
malformed identifier
no resolver link, observed 2026-08-06T18:21:15.067823Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.067823Z digest=sha256:32e12287e8bd48667db25c44559062e28bdd8b3369081d9eadb4721cd9a58ea8

Observation 17d6f87b-6fab-462b-b978-5440b6a6b8f0 · outbound

This paper cites Vector quantization,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vector quantization,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.430826Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.072429Z digest=sha256:19579a88f9d0e4292c7b215a5fee004646f9602fc325ab90e9630d540afda2ff

Observation 9cd49da2-a945-49a8-a346-574c74059506 · outbound

This paper cites Vall-e: A neural codec language model,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Vall-e: A neural codec language model,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.417057Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.076464Z digest=sha256:f419e5d5b7c73a0cf9e1c1a06a94336751cfe0f8978d0326bb4dc8908114649d

Observation eb238fd6-9adb-4e49-abec-7ae058f79ec2 · outbound

This paper cites Compound word transformer: Learning to compose full-song music over dynamic directed hypergraphs,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Compound word transformer: Learning to compose full-song music over dynamic directed hypergraphs,

Reference 35

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.080822Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.080822Z digest=sha256:a9853a61e5e16383e70f693edbe53fe60cb34ba8506c97623ef467738e8a0e6c

Observation 831b8157-ed7a-4379-aff7-c48e5905d790 · outbound

This paper cites Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Pop music transformer: Beat-based modeling and generation of expressive pop piano compositions,

Reference 36

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.085266Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.085266Z digest=sha256:b9978e62881da5b93e1af059e04aeeeb28d09c7ea20744d224fbc89de9ea1984

Observation 4b62de7a-2490-43d3-b983-0e8cd0858188 · outbound

This paper cites Zipformer: A faster and better encoder for automatic speech recognition,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Zipformer: A faster and better encoder for automatic speech recognition,

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.392896Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.090118Z digest=sha256:939a82740d6575bc8d0dafaaed15a55d7d029c858c80c411e24b820029617b74

Observation de143774-171d-48c6-b390-dc9b22105721 · outbound

This paper cites Fréchet audio distance: A reference-free metric for evaluating music enhancement algorithms,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Fréchet audio distance: A reference-free metric for evaluating music enhancement algorithms,

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.378393Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.094347Z digest=sha256:f26e699c6deca371507db9261ed72af085b29fda2b0b27096ee31a2fb9098399

Observation 53affdc1-ad3a-4d62-94e5-28260e76e17f · outbound

This paper cites Adapting Frechet Audio Distance for Generative Music Evaluation.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Adapting Frechet Audio Distance for Generative Music Evaluation

Reference 39

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:15.098354Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:15.098354Z digest=sha256:168fe524e84ed12e6734df2cd69c08d26a70c5d021e93242fd50ca1a8ca41e0c

Observation 13f5afd0-517e-48dc-a994-faca780ef80e · outbound

This paper cites Pijama: Pi- ano jazz with automatic midi annotations,.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling Pijama: Pi- ano jazz with automatic midi annotations,

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-06T18:21:15.362987Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-08-06T18:21:15.102583Z digest=sha256:c4d61019aa6b24a2a110b1f2a37404c3215e38a7dc32ce28a20c9c2fa256776b

Pith citing papers

Observation 4dc0ed6d-f6a3-4563-a897-3a0ddb9912d9 · inbound

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling cites this paper.

MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 1

Resolution
unresolved
no resolver link, observed 2026-08-06T18:21:14.926845Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-06T18:21:14.926845Z digest=sha256:b23a94c6cb007f6861d789ad7c62850631e343faa59698722fe50876780f7f72

Observation 3e9631ad-6b7c-4667-a14e-9888b00353a3 · inbound

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs cites this paper.

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs MIDI-VALLE: Improving Expressive Piano Performance Synthesis Through Neural Codec Language Modelling

Reference 12

Resolution
verified exact
arxiv_id, observed 2026-05-12T03:16:19.155184Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-06T06:34:29.942622+00:00.

source=pdf_text observed=2026-05-12T03:13:47.971429Z digest=sha256:c01ae4a65a773aea5bed0d42e80438d520795b2683a56f67c2bafdeb14210a13