Pith. sign in

Paper Citation Record · LEDGER

Recomposer: Event-roll-guided generative audio editing

As of 18 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 2 inbound Pith citation observations for arXiv:2509.05256.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2509.05256 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-05T05:32:33.761298Z

measured 37 of 37 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-18T06:34:40.430872+00:00

measured 2 of 2 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-01T12:18:17.950810Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-02T20:07:21.617859Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact2
  • verified fuzzy29
  • unresolved4
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation a42aa3a9-7821-4b24-a49b-4d6281f28b9c · outbound

This paper cites Listen, think, and understand,.

Recomposer: Event-roll-guided generative audio editing Listen, think, and understand,

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.280732Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.616656Z digest=sha256:35c019b5a67931b043d21731d312011f4f85115bea844ba9ec77dc65b926009a

Observation 7495a408-be9e-4f8e-80ec-244160261353 · outbound

This paper cites Text-driven separation of arbitrary sounds,.

Recomposer: Event-roll-guided generative audio editing Text-driven separation of arbitrary sounds,

Reference 2

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.266795Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.621645Z digest=sha256:7017f72e7a3c6c90f37d37a1de7ed79b6664086057f0936ca2c0df46f6427aff

Observation 9255ccf1-9525-456b-b805-e02abe2fbcd5 · outbound

This paper cites AudioGen: Textually guided audio generation,.

Recomposer: Event-roll-guided generative audio editing AudioGen: Textually guided audio generation,

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.253217Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.625970Z digest=sha256:920a74eb3edbb7c20d5e1faaa536fdd7f164c91b59205cbd6507be76ea6b7989

Observation bee84d90-8800-4441-bb63-4eeda2dd8623 · outbound

This paper cites AudioLDM: Text-to-audio generation with latent diffusion models,.

Recomposer: Event-roll-guided generative audio editing AudioLDM: Text-to-audio generation with latent diffusion models,

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.240263Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.630349Z digest=sha256:19d731e57eaac1c11f5e3ee536cca2901eda07ffcea85e57283805387718bc95

Observation 398ad15c-2169-483a-9805-5a539235d454 · outbound

This paper cites Text-to-audio generation using instruction guided latent diffusion model,.

Recomposer: Event-roll-guided generative audio editing Text-to-audio generation using instruction guided latent diffusion model,

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.226823Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.634609Z digest=sha256:b3a5fa7a28dd99df74f4c9e248acd8c5144f6245f372c575779c40aab4173ee6

Observation ec5a9522-6699-4df5-8004-60e0d25d6686 · outbound

This paper cites AudioLDM 2: Learning holistic audio generation with self-supervised pretraining,.

Recomposer: Event-roll-guided generative audio editing AudioLDM 2: Learning holistic audio generation with self-supervised pretraining,

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.213262Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.639157Z digest=sha256:8fdab604950af3fd614b3effa105903c259fe5c2bff7b6277bf3859d0dacfba9

Observation d17e5bbd-84d1-433a-ac45-e34c305e5a0f · outbound

This paper cites Adding conditional control to text-to-image diffusion models,.

Recomposer: Event-roll-guided generative audio editing Adding conditional control to text-to-image diffusion models,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.199360Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.644021Z digest=sha256:d4b9a88c7ed8654bd42c86a7bef512c48aac97539adb774da6c8e36a6e31d35a

Observation 6daed0e6-09b3-46e8-a3a0-1d9e01af5c72 · outbound

This paper cites Uni-ControlNet: All-in-one control to text-to-image diffusion models,.

Recomposer: Event-roll-guided generative audio editing Uni-ControlNet: All-in-one control to text-to-image diffusion models,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.185460Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.648286Z digest=sha256:64447bd6ed4de3d804c6899e7e0e9eff63961de4fa594f267e9c2c7a87ada3c2

Observation 862b467a-9599-4b14-9139-f5f7873916cc · outbound

This paper cites Emu Edit: Precise image editing via recognition and generation tasks,.

Recomposer: Event-roll-guided generative audio editing Emu Edit: Precise image editing via recognition and generation tasks,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.171318Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.652545Z digest=sha256:a5d6587ee33248e27a91b9b483bab88c66bbee65dd990edbd9281f3d378eeebd

Observation 6220212e-3aa5-4dcc-9e51-24385a1c249b · outbound

This paper cites AUDIT: Audio editing by following instructions with latent diffusion models,.

Recomposer: Event-roll-guided generative audio editing AUDIT: Audio editing by following instructions with latent diffusion models,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.155698Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.656487Z digest=sha256:127931d6ecbce3f23d4337db48f087a3c06e94e4257548c5a74dcbe717720d95

Observation d98e338c-2485-4933-8ce8-45fc17728dad · outbound

This paper cites Music ControlNet: Multiple time-varying controls for music generation,.

Recomposer: Event-roll-guided generative audio editing Music ControlNet: Multiple time-varying controls for music generation,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.141426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.660544Z digest=sha256:607c3a34566f6419c39ead0c7f60fab1f0119a5eb3768a88a71f22035d3961ba

Observation 663fdd3e-1b49-4d44-8146-a8fdd5e7a604 · outbound

This paper cites Instruct-MusicGen: Unlocking Text-to-Music Editing for Music Language Models via Instruction Tuning.

Recomposer: Event-roll-guided generative audio editing Instruct-MusicGen: Unlocking Text-to-Music Editing for Music Language Models via Instruction Tuning

Reference 12

Resolution
unresolved
no resolver link, observed 2026-08-05T05:32:33.664662Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:32:33.664662Z digest=sha256:95b683f2b15f82a1a3415ce2e41867bba18b6fd6678dbe9cda580daace64e1e4

Observation 61f9ceaf-ae69-4965-a99f-93d782e08ae8 · outbound

This paper cites Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations.

Recomposer: Event-roll-guided generative audio editing Sketch2Sound: Controllable Audio Generation via Time-Varying Signals and Sonic Imitations

Reference 13

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:32:33.839237Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.669434Z digest=sha256:8c98e20a9aad5d0dc28daa322ca87ae9990ef0977be7f9c71de84476e3c05a36

Observation 639a913f-959b-48e4-9f29-7c470e1b45fb · outbound

This paper cites Audio- Composer: Towards fine-grained audio generation with natural language descriptions,.

Recomposer: Event-roll-guided generative audio editing Audio- Composer: Towards fine-grained audio generation with natural language descriptions,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.127624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.674616Z digest=sha256:736d472e5b7ce8cd1f6075aa0fe3a1e6fce6c6edd94340123db6b684627459d7

Observation 8aaf6312-616e-433b-9ffb-578811d8e898 · outbound

This paper cites PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation.

Recomposer: Event-roll-guided generative audio editing PicoAudio: Enabling Precise Timestamp and Frequency Controllability of Audio Events in Text-to-audio Generation

Reference 15

Resolution
verified exact
local_arxiv, observed 2026-08-05T05:32:33.818520Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.678675Z digest=sha256:45e73674f3552c469700118c1dca2dff6703308879568f105ae9509a166e1553

Observation e2dda31f-d73b-4480-ae0c-7e94fbec2f39 · outbound

This paper cites AudioLM: a language modeling approach to audio generation,.

Recomposer: Event-roll-guided generative audio editing AudioLM: a language modeling approach to audio generation,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.113720Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.683329Z digest=sha256:09695ac5b9a2721758969b7a85db3a2de1728a27efee0280d4a84b427f04b394

Observation a3520dae-36d6-4ef5-a387-574be102d4b5 · outbound

This paper cites SoundStream: An end-to-end neural audio codec,.

Recomposer: Event-roll-guided generative audio editing SoundStream: An end-to-end neural audio codec,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.099548Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.687573Z digest=sha256:21d4c181809a6d612053a13cb99089f4b1f4b5d05f837928a992954504b2cc53

Observation a261696b-5ff1-4a2b-b920-7ef2cbf2ec38 · outbound

This paper cites Neural machine translation by jointly learning to align and translate,.

Recomposer: Event-roll-guided generative audio editing Neural machine translation by jointly learning to align and translate,

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-05T05:32:33.691930Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:32:33.691930Z digest=sha256:fd4bf751b727daa523c05751c3bf0fe78a9c35e7d944b412b44cdbd312f52aeb

Observation 9488dbe5-9f4d-4fa8-83f3-4f281a055b1f · outbound

This paper cites Attention is all you need,.

Recomposer: Event-roll-guided generative audio editing Attention is all you need,

Reference 19

Resolution
unresolved
no resolver link, observed 2026-08-05T05:32:33.695957Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:32:33.695957Z digest=sha256:0f731c0e1360e1da75785ca45012c9f7f4c8b0e860b922d9e35b7569f1a26b78

Observation 38477326-f681-4adb-95d1-ebcfb9f2948e · outbound

This paper cites Sentence-T5: Scalable sentence encoders from pre-trained text-to-text models,.

Recomposer: Event-roll-guided generative audio editing Sentence-T5: Scalable sentence encoders from pre-trained text-to-text models,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.067426Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.700126Z digest=sha256:6bdc83c89b463a07599ab9e5401b94b4a7628ef4db6374aab9051f3d66a34931

Observation f8271f57-1855-41a7-b370-51dc2b92bc19 · outbound

This paper cites Autoregressive image generation using residual quantization,.

Recomposer: Event-roll-guided generative audio editing Autoregressive image generation using residual quantization,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.053783Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.704230Z digest=sha256:28ccf86eec82af4fd9211dcfe286f88eac528661a304d8722f0ea134d065873f

Observation 9e62e05c-d10f-4c46-a55a-c1ee77179051 · outbound

This paper cites Moshi: a speech-text foundation model for real-time dialogue.

Recomposer: Event-roll-guided generative audio editing Moshi: a speech-text foundation model for real-time dialogue

Reference 22

Resolution
unresolved
no resolver link, observed 2026-08-05T05:32:33.708489Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-05T05:32:33.708489Z digest=sha256:4cb9b53eb2c6e7421cf5d5bfdd4de8dec4f87d28f73dc050263bd79e01e25598

Observation 27671ccd-b904-4269-b7e4-4567e5c9d0a0 · outbound

This paper cites Freesound technical demo,.

Recomposer: Event-roll-guided generative audio editing Freesound technical demo,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.040442Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.712985Z digest=sha256:8b746d10055f63e6a85cdcb17e3e260298a1c0927185aea6b3d528a441997941

Observation afe5d0cb-10e1-4103-a8cf-780bfc7eb7fe · outbound

This paper cites Unsupervised sound separation using mixture invariant training,.

Recomposer: Event-roll-guided generative audio editing Unsupervised sound separation using mixture invariant training,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.026602Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.716937Z digest=sha256:7eb6cc26fe3b651d844791ba8ef2e9cc12f04af68997c783a99615c7f6f98d5e

Observation 6ca11ffc-cb85-47c7-93bf-eb127c6656a4 · outbound

This paper cites Evaluation of algorithms using games: the case of music annotation,.

Recomposer: Event-roll-guided generative audio editing Evaluation of algorithms using games: the case of music annotation,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:34.012561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.720905Z digest=sha256:4870916025565b8d06bcf5f520f927d02e18e8454053ffc14481c427959b4268

Observation 5a0930af-89f9-4956-a492-de499d824c68 · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Recomposer: Event-roll-guided generative audio editing BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.998452Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.724937Z digest=sha256:b762f89b77390141a7c6acac93a056a834d448cba9c9f3a1548a5f2aa9a456d3

Observation 5975b4e4-7fe5-4bbc-8441-0c3551eb273f · outbound

This paper cites The benefit of temporally-strong labels in audio event classification,.

Recomposer: Event-roll-guided generative audio editing The benefit of temporally-strong labels in audio event classification,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.984097Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.728939Z digest=sha256:a0da5a75ba496ad309c6aa837239d445d9f986704aea68379858f263c83ab514

Observation 13031998-eb7d-42ce-b2a4-b408499e0aea · outbound

This paper cites Sound classification with Y AMNet,.

Recomposer: Event-roll-guided generative audio editing Sound classification with Y AMNet,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.969734Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.732879Z digest=sha256:83e81e5ca68abb58c424453fb04034361a1005fba39812e02aeb92819d353e34

Observation a7429c27-a626-4d34-ab3a-e790d6e24de1 · outbound

This paper cites FSD50K: an open dataset of human-labeled sound events,.

Recomposer: Event-roll-guided generative audio editing FSD50K: an open dataset of human-labeled sound events,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.955733Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.736855Z digest=sha256:ebfc8de80690303ed0157ca44ee01542f69f261f6a9dd8b698bf654d7b55cf4f

Observation 1ce18887-7738-4fb6-b549-710f87d82507 · outbound

This paper cites Neural source-filter waveform models for statistical parametric speech synthesis,.

Recomposer: Event-roll-guided generative audio editing Neural source-filter waveform models for statistical parametric speech synthesis,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.941214Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.741081Z digest=sha256:3137bb759e84c941e5663bea8fbd61cc15cf50327826e72a0dba64e8c90705af

Observation 5f96546f-e018-453a-930c-d1edffbce357 · outbound

This paper cites DDSP: Differentiable digital signal processing,.

Recomposer: Event-roll-guided generative audio editing DDSP: Differentiable digital signal processing,

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.925841Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.745141Z digest=sha256:9f5ab23a8451561dec70610972320ac6f1560af488d180a800457a81511c0906

Observation c5e36e78-62c6-4984-983c-777e99696164 · outbound

This paper cites Diffsound: Discrete diffusion model for text-to-sound generation,.

Recomposer: Event-roll-guided generative audio editing Diffsound: Discrete diffusion model for text-to-sound generation,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.911670Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.749190Z digest=sha256:8c28f87af94dda6aa5db7049900895dc2028892ef0bbc553f66901361f3ec98a

Observation 9901077a-0838-4ed8-91b9-99b259d46523 · outbound

This paper cites Fr ´echet audio distance: A metric for evaluating music enhancement algorithms,.

Recomposer: Event-roll-guided generative audio editing Fr ´echet audio distance: A metric for evaluating music enhancement algorithms,

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.896931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.753098Z digest=sha256:575eff267f79d96a6918fbae95957509709bb072ac537c17e4c4469fd93b371f

Observation 3be47c5d-004a-4bc6-9fcb-7b1518a8248e · outbound

This paper cites Adapting Fr ´echet audio distance for generative music evaluation,.

Recomposer: Event-roll-guided generative audio editing Adapting Fr ´echet audio distance for generative music evaluation,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.883079Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.757206Z digest=sha256:5f7a568c3572222fe754a6abe5730491f88a7e21e1c5e49421fe858b02bb544b

Observation b229e865-e96e-4a8d-9ae5-5aee982029ba · outbound

This paper cites Correlation of Fr ´echet audio distance with human perception of environmental audio is embedding dependent,.

Recomposer: Event-roll-guided generative audio editing Correlation of Fr ´echet audio distance with human perception of environmental audio is embedding dependent,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-05T05:32:33.868695Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-08-05T05:32:33.761298Z digest=sha256:13cbdb2ff43419a1ca5e42ad75fd26d3edf142bd66c539493add2194d74911eb

Pith citing papers

Observation e49812c5-c163-4796-881b-ab1cb714c36c · inbound

MMAE: A Massive Multitask Audio Editing Benchmark cites this paper.

MMAE: A Massive Multitask Audio Editing Benchmark Recomposer: Event-roll-guided generative audio editing

Reference 15

Resolution
verified exact
arxiv_id, observed 2026-07-02T20:07:21.619554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-18T06:34:40.430872+00:00.

source=pdf_text observed=2026-06-27T21:02:04.021977Z digest=sha256:03493a3175966513b61dc79ad1de8cee6bcccdb234335e3902c38f144c39b2a5

Observation fc55a138-3b0f-493b-8f2f-56f192e94180 · inbound

RIME: Enabling Large-Scale Agentic Music Post-Production cites this paper.

RIME: Enabling Large-Scale Agentic Music Post-Production Recomposer: Event-roll-guided generative audio editing

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-01T12:18:17.950810Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-01T12:18:17.950810Z digest=sha256:da97f988737fddf61772a193532e85295f712ae315bda1d8dbf414a18a17312e