Pith. sign in

Paper Citation Record · LEDGER

Generative AI for Music and Audio

As of 21 August 2026, this Paper Citation Record lists 13 of 13 outbound references and 0 inbound Pith citation observations for arXiv:2411.14627.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2411.14627 v1

Coverage vector

measured 13 of 13 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-12T15:10:14.185025Z

measured 13 of 13 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-21T06:32:19.484+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

13 of 13 outbound references displayed

  • verified exact0
  • verified fuzzy5
  • unresolved8
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation c857bac9-c5f9-4aee-a014-c1460a8aa07d · outbound

This paper cites TensorFlow: A system for large- scale machine learning.

Generative AI for Music and Audio TensorFlow: A system for large- scale machine learning

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.286039Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T15:10:14.148529Z digest=sha256:9443b264e9f5729f8e37634434de967c73170bb20097f59b337aaab5978b1bd7

Observation 7c08aaaf-e2c3-4480-aca2-dcde314eee68 · outbound

This paper cites Noise2Music: Text-conditioned Music Generation with Diffusion Models.

Generative AI for Music and Audio Noise2Music: Text-conditioned Music Generation with Diffusion Models

Reference 2

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.151790Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.151790Z digest=sha256:d120adeade132a86de2fbfea5b7d6dc01365cbbefeadab3c656f74fcd7c96ff2

Observation e7d6b493-1a91-4c74-a6a9-089b179f23bf · outbound

This paper cites BERT-like Pre-training for Symbolic Piano Music Classification Tasks.

Generative AI for Music and Audio BERT-like Pre-training for Symbolic Piano Music Classification Tasks

Reference 3

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.155450Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.155450Z digest=sha256:6b5399eb7147cd2b1eb4074c26e1f4349f33d9c728eb6d1976fc3526aff89113

Observation db998460-ee34-4d73-8e1b-6c7ca5feccfa · outbound

This paper cites MMM : Exploring Conditional Multi-Track Music Generation with the Transformer.

Generative AI for Music and Audio MMM : Exploring Conditional Multi-Track Music Generation with the Transformer

Reference 4

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.158686Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.158686Z digest=sha256:9fa93f28e81363b263238b761aa1314b914ed102ef7e924f56fd417947ac5f17

Observation 340f063d-6c05-4ffb-b852-b180a4adcf97 · outbound

This paper cites AudioGen: Textually Guided Audio Generation.

Generative AI for Music and Audio AudioGen: Textually Guided Audio Generation

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.278188Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T15:10:14.161825Z digest=sha256:076a2737d4cf1a3ef0b7e09d13a735a46b5156ee144d85379ea3150d087c44b5

Observation 5c4b5c12-52d4-4b05-908f-d3cd3198861b · outbound

This paper cites WavJourney: Compositional Audio Creation with Large Language Models.

Generative AI for Music and Audio WavJourney: Compositional Audio Creation with Large Language Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.164467Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.164467Z digest=sha256:8b5cc68fff200c88524ef820042632524670c781185b071e1a7f063de01db128

Observation eeeb6c09-e0c6-4fa0-bd46-be027c669473 · outbound

This paper cites PyTorch: An Imperative Style, High-Performance Deep Learning Library.

Generative AI for Music and Audio PyTorch: An Imperative Style, High-Performance Deep Learning Library

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.271008Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T15:10:14.167567Z digest=sha256:dd734e986ccb37fad6c885a4e6003619dbfb8a79768dcadcf56502140bd8423d

Observation 27d1484d-0f59-4b0f-aabe-58c6b8808f32 · outbound

This paper cites Hierarchical Text-Conditional Image Generation with CLIP Latents.

Generative AI for Music and Audio Hierarchical Text-Conditional Image Generation with CLIP Latents

Reference 8

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.170530Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.170530Z digest=sha256:8a0c2358bff803693fb4c0ecbdd4b4a4998cbaafb6b1e36ab3313274111d04ab

Observation ecc9a73f-d6cc-4680-8026-488964378b0c · outbound

This paper cites DeepSinger: Singing Voice Synthesis with Data Mined From the Web.

Generative AI for Music and Audio DeepSinger: Singing Voice Synthesis with Data Mined From the Web

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.263856Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T15:10:14.173330Z digest=sha256:87799b407c4189880b46dd08fe27e8410e92837c2951641d61cda83f121f69dd

Observation d5362403-a65a-4566-9025-37955489269a · outbound

This paper cites FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control.

Generative AI for Music and Audio FIGARO: Generating Symbolic Music with Fine-Grained Artistic Control

Reference 10

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.176396Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.176396Z digest=sha256:1b78c94a704522aad51b8e0b3ed07e2d762fbd5885fa6956d4ccdbf7444de763

Observation ec8ed454-0267-4a88-b1af-d4c5fffb268a · outbound

This paper cites LAION-5B: An open large-scale dataset for training next generation image-text models.

Generative AI for Music and Audio LAION-5B: An open large-scale dataset for training next generation image-text models

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-12T15:10:14.254659Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-21T06:32:19.484+00:00.

source=pdf_text observed=2026-08-12T15:10:14.179278Z digest=sha256:f0037293e679019a548a66ee4e6a9a1372cb1db5410993d839a897783a2c6753

Observation b32daa7e-6e65-46e0-82b4-2ba872cd9d36 · outbound

This paper cites A Survey on Neural Speech Synthesis.

Generative AI for Music and Audio A Survey on Neural Speech Synthesis

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.182250Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.182250Z digest=sha256:4fa64e6a67f02b0a27b24491b72db8fd3ac4fee3e21020323e64f67668755465

Observation d27ae42c-27ac-42ed-ba55-fc6a84f16ab6 · outbound

This paper cites Diffsound: Discrete Diffusion Model for Text-to-sound Generation.

Generative AI for Music and Audio Diffsound: Discrete Diffusion Model for Text-to-sound Generation

Reference 13

Resolution
unresolved
no resolver link, observed 2026-08-12T15:10:14.185025Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-12T15:10:14.185025Z digest=sha256:54257b09aee297045e32360d992a8e9c88266f869d09d143050436f7015da648

Pith citing papers

No inbound Pith citation observations are available.