Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:16:19.338517Z
Paper Citation Record · LEDGER
As of 8 August 2026, this Paper Citation Record lists 64 of 64 outbound references and 36 inbound Pith citation observations for arXiv:2506.00045.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-07T13:16:19.338517Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-08T06:32:00.761636+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:14:43.128810Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-07-03T21:48:59.860115Z
64 of 64 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation a3000579-c655-4ccd-8eeb-6ad66dab2aef · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Yue: Scaling open foundation models for long-form music generation, 2025
Reference 1
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4255b198-fae2-4325-a752-c7ad77d29d30 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Songgen: A single stage auto-regressive transformer for text-to-song generation, 2025
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d98a8a89-41db-4cec-936a-a6fd0dab535f · outbound
ACE-Step: A Step Towards Music Generation Foundation Model DiffRhythm: Blazingly Fast and Embarrassingly Simple End-to-End Full-Length Song Generation with Latent Diffusion
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b6a86041-e35e-4705-83e3-54706ec5eb98 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Sana: Efficient high-resolution image synthesis with linear diffusion transformer, 2024
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 433aee22-c312-42d5-86a4-7cf845f2b03f · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Deep Compression Autoencoder for Efficient High-Resolution Diffusion Models
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2589ed82-51ac-4461-8b3c-f03ef7a4ae49 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Mert: Acoustic music understanding model with large-scale self-supervised training, 2023
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3c5ea20a-81e4-40f1-a753-405ba1c4e262 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model mhubert-147: A compact multilingual hubert model
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cf59c7bb-f0ad-4a2a-9c97-9ebdb1aecd2e · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Representation alignment for generation: Training diffusion transformers is easier than you think
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3aa12378-f07a-4b68-b16b-689f0133f0bd · outbound
ACE-Step: A Step Towards Music Generation Foundation Model High-resolution image synthesis with latent diffusion models, 2021
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a665c522-449b-4a78-9fe9-ab3b64547272 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Unresolved cited work
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2528ee31-d1c1-4366-b2e2-c0c7be8d4db0 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model LLaMA: Open and Efficient Foundation Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5dbff2c0-301d-4a87-acd6-9b7f09410249 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Qwen Technical Report
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d73db496-44ac-42ad-88d2-2a32b812dc3d · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Deepseek-v3 technical report, 2024
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 54a29bf8-ac64-44fc-87d2-774e64440123 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Improving Image Generation with Better Captions
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3db75307-76be-43f4-be64-410f63a112bc · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Unresolved cited work
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 032a81a8-48ef-4875-acc7-b16e737b0078 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Imagen 3
Reference 16
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b3d4face-d1e9-4fe9-8575-ff0f68040a27 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Video generation models as world simulators
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36bfe483-0532-4eb5-a634-1a25ad3b1c5c · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Kling AI Video Generation Large Model
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 0eeb16a1-3e25-4afd-939d-d8ab09c85dcc · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Wan: Open and Advanced Large-Scale Video Generative Models
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c1ab91cc-a288-4e78-94fd-1b9f9c61077f · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Jukebox: A Generative Model for Music
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb8fcf73-5fa6-4dad-94bf-0430639983ac · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Simple and controllable music generation, 2024
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cefd8126-934c-499e-86dd-8758f887d83b · outbound
ACE-Step: A Step Towards Music Generation Foundation Model AudioLDM: Text-to-audio generation with latent diffusion models
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2c80b67f-2374-4816-a3e5-48bee299dd49 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Plumbley
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be19aa28-23b5-4c92-b8ac-b9afdf88f0ea · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Suno: Ai music generation platform
Reference 24
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 77fd87bf-e74c-457d-b26c-12c7b13daf85 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Udio: Ai music creation platform
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation da6d2e8c-483f-4fd1-86da-7988f2c1a9f4 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Riffusion: Stable diffusion for real-time music generation
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 502e9e4b-5afd-432d-8ca9-7055c521d108 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Parker, CJ Carr, Zack Zukowski, Josiah Taylor, and Jordi Pons
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 41b8446b-b645-45db-b479-d1ca6c03f281 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Fast text-to-audio generation with adversarial post-training, 2025
Reference 28
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 5b8c6a90-b403-45be-91b9-c5ad4371ff17 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Flow Straight and Fast: Learning to Generate and Transfer Data with Rectified Flow
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 67bd7ed7-129c-4f63-a0c9-86e7737876e3 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Efficientvit: Multi-scale linear attention for high-resolution dense prediction, 2024
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6d5f9850-9028-4703-a449-c41822ad9d5a · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Hu, Yelong Shen, Phillip Wallis, Zeyuan Allen-Zhu, Yuanzhi Li, Shean Wang, Lu Wang, and Weizhu Chen
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 63d68aac-c2fc-4035-8987-286191a8ca6c · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Adding conditional control to text-to-image diffusion models, 2023
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b9368db9-aa8b-432d-9efc-e8f9e1fadb3a · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Statistical parametric speech synthesis, 2007
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 587953c3-962d-4b96-8ae5-507bfbef7791 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model MegaTTS 3: Sparse Alignment Enhanced Latent Diffusion Transformer for Zero-Shot Speech Synthesis
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2543d06-48ff-44c9-b707-fe77133774a3 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Neural discrete representation learning, 2018
Reference 35
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6dd4fbcc-2fb6-41c2-ae42-37249840015a · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Denk, Zalán Borsos, Jesse Engel, Mauro Verzetti, Antoine Caillon, Qingqing Huang, Aren Jansen, Adam Roberts, Marco Tagliasacchi, Matt Sharifi, Neil Zeghidour, and Christian Frank
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 6e5baced-7c4b-4e37-96f8-8f2d22c21b12 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model High fidelity neural audio compression, 2022
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b38abdfa-eddb-436a-96b3-1811361198a5 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Large- scale contrastive language-audio pretraining with feature fusion and keyword-to-caption augmentation
Reference 38
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a5bbe826-79cd-46da-9e6c-766a694cd5b4 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Unresolved cited work
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a48cfb5-87c1-4360-94a0-4a48302fc962 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Unresolved cited work
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 55040861-a98a-43d6-87d5-39e4c0d1eed6 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Scalable diffusion models with transformers, 2023
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 40dc43ad-6a7c-4351-b11e-9b8bd383284b · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Gomez, Lukasz Kaiser, and Illia Polosukhin
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2465be3-9e95-4ce1-8488-3cc5f589b692 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model U-net: Convolutional networks for biomedical image segmentation, 2015
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6b2dcfad-8129-4fce-96a4-18d49de674f5 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Meta audiobox aesthetics: Unified automatic quality assessment for speech, music, and sound
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 62638e80-0a3d-4f37-8f4c-bc087b8beeee · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Qwen2.5-Omni Technical Report
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a724f43c-4e91-4834-95d8-45eda269d2b9 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Robust speech recognition via large-scale weak supervision, 2022
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation bfc12106-f8f9-4142-96a3-2b41058488de · outbound
ACE-Step: A Step Towards Music Generation Foundation Model ekzhu/datasketch: v1.6.5, May 2024
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 35c29adf-ab98-49a5-b5b2-aa40a98aee3c · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Byt5 model for massively multilingual grapheme-to-phoneme conversion
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b98c5d1f-f36e-4a92-99c6-a50be36cd851 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model All-in-one metrical and functional structure analysis with neighborhood attentions on demixed audio
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 68373a3d-a027-485c-9967-8950194a1811 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Beat this! accurate beat tracking without DBN postpro- cessing
Reference 50
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation cbdea552-a5c0-4afe-81ec-8c671937469c · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Essentia: An audio analysis library for music information retrieval
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 711bf726-379a-44c7-a7b0-2038a34f9ceb · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Xtts: a massively multilingual zero-shot text-to-speech model, 2024
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 876fc305-bd3e-4431-9b59-e4b551bf33c4 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Fish-speech: Leveraging large language models for advanced multilingual text-to-speech synthesis, 2024
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e626520f-3c7f-4b8d-88fa-fefb94a03737 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Pixart-α: Fast training of diffusion transformer for photorealistic text-to-image synthesis, 2023
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9879f042-94ca-47cf-899d-307e3e426a08 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Unimax: Fairer and more effective language sampling for large-scale multilingual pretraining, 2023
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 3f70d972-b2da-4ae0-9ce9-9a8f7eae6228 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Learning diverse features with part-level resolution for person re-identification, 2020
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 4ef08c98-252d-4b9d-8399-48c7620debb5 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Scaling rectified flow transformers for high-resolution image synthesis, 2024
Reference 57
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9e9a9e3c-20c2-44da-b4ca-9d0df6151958 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Adapting frechet audio distance for generative music evaluation
Reference 58
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 837ce3fc-4811-4b2a-81eb-be22b0b2d8ac · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Stable Audio Metrics, 2024
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation a56da72c-0086-41a0-85d5-db3e88f850a4 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f76fba90-df7f-43ae-af5a-a196407df3fe · outbound
ACE-Step: A Step Towards Music Generation Foundation Model stable-ts: Stabilizing Timestamps for Whisper, 2024
Reference 61
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 73391fcb-d9e8-4f78-a451-9e57ef435084 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model SongEval: A Benchmark Dataset for Song Aesthetics Evaluation
Reference 62
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5ccf2bb-3846-43b3-b876-238fa7d2be45 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model Simplifying, stabilizing and scaling continuous-time consistency models, 2025
Reference 63
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1a7b8550-da2e-4817-86e1-545e4bad6098 · outbound
ACE-Step: A Step Towards Music Generation Foundation Model FlowEdit: Inversion-Free Text-Based Editing Using Pre-Trained Flow Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 607794e3-305a-4413-8fa2-5d90040a27f4 · inbound
JAM: A Tiny Flow-based Song Generator with Fine-grained Controllability and Aesthetic Alignment ACE-Step: A Step Towards Music Generation Foundation Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c7bb091b-edf2-421f-808d-1cf2109f21c1 · inbound
AudioGen-Omni: A Unified Multimodal Diffusion Transformer for Video-Synchronized Audio, Speech, and Song Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0a61d2d4-36d4-439c-83c2-634dc4032319 · inbound
Towards Hallucination-Free Music: A Reinforcement Learning Preference Optimization Framework for Reliable Song Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5110b3af-f508-44e5-b08f-763b95cffc20 · inbound
UniVerse-1: Unified Audio-Video Generation via Stitching of Experts ACE-Step: A Step Towards Music Generation Foundation Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af8bd234-643b-4344-aedd-ca8b67093cd6 · inbound
Ovi: Twin Backbone Cross-Modal Fusion for Audio-Video Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation b6b2c2c2-c7a3-481f-8e9c-b03a6dd75d6a · inbound
SongFormer: Scaling Music Structure Analysis with Heterogeneous Supervision ACE-Step: A Step Towards Music Generation Foundation Model
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7ae5821c-dec4-4de1-bb54-c93df909c2ee · inbound
JoVA: Unified Multimodal Learning for Joint Video-Audio Generation and Editing ACE-Step: A Step Towards Music Generation Foundation Model
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9968680e-dce0-4bf3-84ac-c4ce297dd30f · inbound
TADA! Tuning Audio Diffusion Models through Activation Steering ACE-Step: A Step Towards Music Generation Foundation Model
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 7c6cabd7-a8aa-43c8-be3d-6880c211ae87 · inbound
MIDI-Informed Singing Accompaniment Generation in a Compositional Song Pipeline ACE-Step: A Step Towards Music Generation Foundation Model
Reference 12
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 310031d3-f4a0-4f8d-9b1a-2c3d248d3cb6 · inbound
Echoes: A semantically-aligned music deepfake detection dataset ACE-Step: A Step Towards Music Generation Foundation Model
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ecbcf569-4432-4a67-8b22-40bd3d323ec8 · inbound
LaDA-Band: Language Diffusion Models for Vocal-to-Accompaniment Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 18
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 13a0087d-9539-4c4a-8519-843e847b4df9 · inbound
LaDA-Band: Language Diffusion Models for Vocal-to-Accompaniment Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d0263534-865a-4bae-b7db-23f713ff310f · inbound
SongBench: A Fine-Grained Multi-Aspect Benchmark for Song Quality Assessment ACE-Step: A Step Towards Music Generation Foundation Model
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 86ef11ed-3261-4695-a0b5-79949d4777c1 · inbound
TMD-Bench: A Multi-Level Evaluation Paradigm for Music-Dance Co-Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation be05e593-f1f9-4e55-abdc-b9638dcd151c · inbound
APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music ACE-Step: A Step Towards Music Generation Foundation Model
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 36991cea-e9fa-440d-b263-15473293870c · inbound
APEX: Large-scale Multi-task Aesthetic-Informed Popularity Prediction for AI-Generated Music ACE-Step: A Step Towards Music Generation Foundation Model
Reference 47
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 8cef5aeb-d6b0-4e75-aee4-3ee51dfd1d14 · inbound
Cutting rules in strong field QED with application to trident pair production ACE-Step: A Step Towards Music Generation Foundation Model
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 12dd018e-0aad-4d75-9e9d-1619b69a25eb · inbound
OmniNFT: Modality-wise Omni Diffusion Reinforcement for Joint Audio-Video Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 2
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 788b2689-12cb-414e-8b9a-0b20e04cbf89 · inbound
S2Accompanist: A Semantic-Aware and Structure-Guided Diffusion Model for Music Accompaniment Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ebc97580-98ed-49a8-a723-6ca14b67c2c3 · inbound
Instrumental Text-to-Music Generation with Auxiliary Conditioning Branches ACE-Step: A Step Towards Music Generation Foundation Model
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 9d0fd53c-60e0-46d6-8c91-6b165351c0aa · inbound
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment ACE-Step: A Step Towards Music Generation Foundation Model
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation d47d9dbf-9398-49a1-b09c-e73fb9ac7a28 · inbound
Towards Unified Song Generation and Singing Voice Conversion with Accompaniment Co-Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation e3b93717-bbc5-4583-8be1-24859dd65462 · inbound
An Empirical Analysis of AI Slop in Music Streaming ACE-Step: A Step Towards Music Generation Foundation Model
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation ddcc77e6-6770-4132-9653-1005ed10afa5 · inbound
LeVo 2: Stable and Melodious Song Generation via Hierarchical Representation Modeling and Progressive Post-Training ACE-Step: A Step Towards Music Generation Foundation Model
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-08T06:32:00.761636+00:00.
Observation 1db8f3b9-87af-4035-ab09-b0e72081ee04 · inbound
MusicMark: A Robust Generative Watermarking Framework for Music Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f2752145-d739-4f3b-abbc-59bf5cc42406 · inbound
Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance ACE-Step: A Step Towards Music Generation Foundation Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 91482bf6-aca7-4c5f-9ea9-c288237debd6 · inbound
Anysynth:Zero-Shot Instrument Cloning via In-Context Learning and Asymmetric Hierarchical Guidance ACE-Step: A Step Towards Music Generation Foundation Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27e079cd-46e0-4357-a4b9-92661222b514 · inbound
Qwen-Music Technical Report ACE-Step: A Step Towards Music Generation Foundation Model
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fe00cef8-e6c0-4493-ab55-6566a33784d9 · inbound
Qwen-Music Technical Report ACE-Step: A Step Towards Music Generation Foundation Model
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db7d261c-a992-4881-837a-d860a877bf17 · inbound
Genre Bias or Aesthetic Perception? Identifying and Mitigating Shortcut Learning in Music Evaluation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d949e0fb-134b-4f84-8d6c-0bedcbd70441 · inbound
A Diagnostic Evaluation Framework for AI-Generated Cover Songs Using Music-Theoretic and Acoustic Features ACE-Step: A Step Towards Music Generation Foundation Model
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c2a48299-6730-44a6-baaf-3ad500282caf · inbound
Pushing the Frontier of Full-Song Generation: Hierarchical Autoregressive Planning Meets Flow-Matching Rendering ACE-Step: A Step Towards Music Generation Foundation Model
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 17568aa2-93fd-49dc-b420-c0f47a369205 · inbound
MPEcho: A Melody and Phoneme-Aware Generative Framework for Controllable Cover Song Generation ACE-Step: A Step Towards Music Generation Foundation Model
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 37152fb8-a22d-4c0f-b544-e5e048dca679 · inbound
TORUS: A Test of Rendering-Understanding Self-Coherence for Unified Audio Models ACE-Step: A Step Towards Music Generation Foundation Model
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4942d2d7-e8a2-4feb-bb5a-321663b226e2 · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks ACE-Step: A Step Towards Music Generation Foundation Model
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0518d6f5-8d75-48c3-b78a-d633529e7370 · inbound
SwanTale: Unified Multi-Speaker Speech and Audio Generation for Instruct and Zero-Shot Tasks ACE-Step: A Step Towards Music Generation Foundation Model
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.