Pith. sign in

Paper Citation Record · LEDGER

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms

As of 14 August 2026, this Paper Citation Record lists 18 of 18 outbound references and 0 inbound Pith citation observations for arXiv:2509.26007.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.26007 v2

Coverage vector

measured 18 of 18 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-05-18T12:38:27.978216Z

measured 18 of 18 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 0 of 0 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: cited_works

Reference resolution

18 of 18 outbound references displayed

  • verified exact2
  • verified fuzzy16
  • unresolved0
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 83cdff75-4210-4bf5-bf6d-7b3415d51825 · outbound

This paper cites Non-autoregressive neural text-to-speech.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Non-autoregressive neural text-to-speech

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.057478Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:6b06acd3b0634f9cd67c7b12f906183f5b1ab609e58d915f7fe6f9cfd7d3df1c

Observation d098c0e9-e5fe-49fc-8e91-c54a8e701b39 · outbound

This paper cites Parallel wavegan: A fast waveform generation model based on generative adversarial networks with multi- resolution spectrogram.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Parallel wavegan: A fast waveform generation model based on generative adversarial networks with multi- resolution spectrogram

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.060404Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:b923c2838d185fba253dee4a86505cf05cee04b337c16c284ef6a6c0c4b49f58

Observation 3d79dbeb-3b48-4f38-abcd-cd875583c352 · outbound

This paper cites High fidelity speech syn- thesis with adversarial networks.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms High fidelity speech syn- thesis with adversarial networks

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.054470Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:aaf2ecdfeeb44d6f752c67e367317e435e6e325cc0de22dee4fa015b404ec244

Observation cd9fa6aa-95a7-4973-8b25-06d60de61705 · outbound

This paper cites Diffwave: A versatile diffusion model for audio synthesis.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Diffwave: A versatile diffusion model for audio synthesis

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.063637Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:a2cacb4060a3efd125bb26268b8c1b9adc2456c43baa0c73936eee01a8aeafdf

Observation 72bd4f13-6e81-46fc-a026-bc118c13a70d · outbound

This paper cites MelNet: A Generative Model for Audio in the Frequency Domain.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms MelNet: A Generative Model for Audio in the Frequency Domain

Reference 5

Resolution
verified exact
local_arxiv, observed 2026-05-18T12:41:22.890885Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:0ed862472d608a331098a0c3ba384a5e43b86c650037d19ad107c991357f1d94

Observation 3dc91e90-9d1f-4abc-92d1-791a36552d38 · outbound

This paper cites Gansynth: Adversarial neural audio synthesis.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Gansynth: Adversarial neural audio synthesis

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.067574Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:9c3adf87dbb95639a5e07b22c01b9e945dedf46f136a72c8d45b993dffb58dd2

Observation 53f1043c-1522-481f-a8e8-5928e1b37d74 · outbound

This paper cites EDMSound: Spectrogram Based Diffusion Models for Efficient and High-Quality Audio Synthesis.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms EDMSound: Spectrogram Based Diffusion Models for Efficient and High-Quality Audio Synthesis

Reference 7

Resolution
verified exact
arxiv_id, observed 2026-05-18T12:41:22.884762Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:dc0bb6634d17ada8099a40a41229e45d7b89ec7a8761f8310b10005108fcf79b

Observation 7b974308-6f21-489d-9cfa-ec5b7a9f6b71 · outbound

This paper cites Visual autoregressive modeling: Scalable image generation via next-scale prediction.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Visual autoregressive modeling: Scalable image generation via next-scale prediction

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.083269Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:256507dd64d1e0797ddd3f24412403c70c3e190e386930ec5107b9f285510b83

Observation c4e887dd-f867-4952-9af9-26c744b6e9ca · outbound

This paper cites Imagefolder: Autoregres- sive image generation with folded tokens.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Imagefolder: Autoregres- sive image generation with folded tokens

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.088235Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:ba466303fe4e33afdf24e538593580ab4f28ba39c84dc731e8b8b660757e7343

Observation 81aa1604-05e3-436b-a034-65c76d082ef1 · outbound

This paper cites Neural audio synthesis of musical notes with wavenet autoencoders.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Neural audio synthesis of musical notes with wavenet autoencoders

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.080522Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:85894680b8bd5678efc537c7a087cbf93dacd6e9b6bb46390615ea73fec3a6fb

Observation ee707657-182d-4a61-8901-79c29218b51a · outbound

This paper cites Evaluating gen- erative audio systems and their metrics.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Evaluating gen- erative audio systems and their metrics

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.085974Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:7d97c514c53456f5e206c0127ba658397aa306c5450cc304ca6b6c7129bb0cfe

Observation d9d658cc-394b-4bba-b8e4-09c2b4085d7f · outbound

This paper cites DDSP: differentiable digital signal pro- cessing.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms DDSP: differentiable digital signal pro- cessing

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.071183Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:38d2cd1c935e50502ebf0e4c26ba73281f054a980d43d87484f969aab9440cf6

Observation 2b89528e-07d0-4ec3-8a9b-b0376a0adcc2 · outbound

This paper cites Signal estimation from mod- ified short-time fourier transform.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Signal estimation from mod- ified short-time fourier transform

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.077078Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:b6ace307856661704cb4e4086e2ac99be04baf1be97198087130ea2bf7830c48

Observation 3f869f34-3f93-4e11-9ad5-0d32e29417a7 · outbound

This paper cites Image-to-image translation with conditional ad- versarial networks.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Image-to-image translation with conditional ad- versarial networks

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.051621Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:3f6b5cc541459b4ef2f51b6b9f6f9224698150ef3509fbaad29f1c102cd9d64e

Observation 1dd7580a-31cd-42fd-a90c-b4085828bf57 · outbound

This paper cites On gans and gmms.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms On gans and gmms

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.074345Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:a363ab174851313e0393a041955be7d103c75986b7f4095811abba5701b98519

Observation 58992367-69fa-49ae-8f5f-745b0660273e · outbound

This paper cites Com- paring representations for audio synthesis using gener- ative adversarial networks.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Com- paring representations for audio synthesis using gener- ative adversarial networks

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.046259Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:f14b452b2ddcdcefe0d83b4c45299541b0bdd4b29efe10785c5378dbf15603a6

Observation 8f4acb44-15cf-475b-b199-776a9dfbec36 · outbound

This paper cites Demystifying MMD gans.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Demystifying MMD gans

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.043328Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:6182a07df296c545de5ee28ef79c41b86fb05e97473643635a7c56a7dd563c90

Observation 8ffe8b70-7013-40db-b8dd-e7aa9b3cefb6 · outbound

This paper cites Fr´echet audio distance: A reference- free metric for evaluating music enhancement algo- rithms.

MARS: Sound Generation via Multi-Channel Autoregression on Spectrograms Fr´echet audio distance: A reference- free metric for evaluating music enhancement algo- rithms

Reference 18

Resolution
verified fuzzy
raw_fallback, observed 2026-05-18T12:42:38.048804Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-18T12:38:27.978216Z digest=sha256:51b7f0c0cd67e75b54d85b8cfdb612fd5a220d965d44a8ea25b022e836cf4020

Pith citing papers

No inbound Pith citation observations are available.