Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:31:44.023128Z
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 79 of 79 outbound references and 10 inbound Pith citation observations for arXiv:2412.06660.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-11T19:31:44.023128Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-07T00:49:44.099128Z
A source-named dated measurement, never combined with another source.
Source: arxiv_reference, observed 2026-08-05T02:28:24.338817Z
79 of 79 outbound references displayed
External citation measurements
0
arxiv_reference, observed 2026-08-05T02:28:24.338817Z
Observation 52747d69-ecaa-42ae-944b-b6c5af526adc · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models , " * write output.state after.block = add.period write newline
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0ba7e298-946c-4c5b-aee2-046c69434faf · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models write newline
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d359a3da-9657-47eb-ba09-8ad9fedf5f58 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models MusicLM: Generating Music From Text
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2aab8bef-3821-4dfa-9b42-4a40c0798e34 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 4
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 88ba6a51-11ed-41f5-8d46-10efce83322a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 5
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8928c1fc-95d9-4b4e-aba5-526d748473b0 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 6
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e0d86532-eb3e-412c-9e36-7a1f0beff2c6 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 7
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 80f68da0-f51d-40a7-ac0c-b12f9ec1b657 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Stable Video Diffusion: Scaling Latent Video Diffusion Models to Large Datasets
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f525db-bb51-4be1-96d3-848765969541 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 9
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 83817e5c-7fe5-459b-aad0-537c6c10f49e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 3feb965a-9521-4145-99a7-74a4b43be759 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c7093b46-d3e5-406a-a6fc-ec4ee3e93aed · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Simple and Controllable Music Generation
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5b0b685d-f88a-4002-89ce-00f4d21d3269 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 13
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 81f03d1d-2b10-43df-a194-58202a90c243 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 09192207-9f26-44b1-b4c2-c6c38798cefe · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 8a88e08a-9bfa-43b6-9806-b4a8e16df373 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models DreamLLM: Synergistic Multimodal Comprehension and Creation
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f6097280-c2f2-47f8-b9d0-f5c224f2f2cd · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 42cb127e-1952-4855-87e1-c32ff338a4cf · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models LLaMA-Adapter V2: Parameter-Efficient Visual Instruction Model
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c83293ac-b572-44e4-a03e-2a2e4eec0f1a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Planting a SEED of Vision in Large Language Model
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 78898fc8-66cf-4b12-b871-78e0c9f0cdf3 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Making LLaMA SEE and Draw with SEED Tokenizer
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a4d8a4db-3c20-419f-8fc8-2ea0cc59f4af · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models F.; Ellis, D
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cb8d17b4-f7d8-46cf-8426-0e4ef07eaa56 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models V.; Joulin, A.; and Misra, I
Reference 22
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e0842bcb-6311-4948-8dc5-4b55942112be · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 23
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7e376fd6-1a0e-4692-82fe-5598a7617139 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Listen, Think, and Understand
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation baeb32e2-5344-40fc-922a-7c98d4eaf7cb · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 25
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation a6bf126f-83b2-464b-8675-01c2cbbe3763 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 26
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation c601d7a0-ab0d-4afa-92ed-e20d9e70784c · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Point-Bind & Point-LLM: Aligning Point Cloud with Multi-modality for 3D Understanding, Generation, and Instruction Following
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06b2e3f3-6fe6-4a5e-b8af-917de5d3ce31 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models InstructME: An Instruction Guided Music Edit And Remix Framework with Latent Diffusion Models
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1f62d293-4a55-4a37-bd16-b44ea017a713 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 29
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 69052478-f215-4616-8153-e1276cff4bd0 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 30
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 84b756ee-1ead-4558-93c0-3770b665280a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models J.; Shen, Y.; Wallis, P.; Allen-Zhu, Z.; Li, Y.; Wang, S.; Wang, L.; and Chen, W
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 15a81bfd-452f-4c7a-8bcd-67b378a9de1a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4791fb4-4bc1-42b8-8642-add112dd2ead · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 33
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation dc8013f5-4ab4-46a3-994c-6c29a2339efc · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Mistral 7B
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation fb776e67-9be2-4663-aec3-6aca20e9d67e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 689e732f-3887-489d-a2eb-001b8d873cde · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 36
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5ebf0617-3da7-46c6-8701-c814db27bf03 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 37
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4849aeb2-9e27-45af-a0a9-683d9500c0d7 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models BLIP-2: Bootstrapping Language-Image Pre-training with Frozen Image Encoders and Large Language Models
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5be3d1a-7846-4d7f-a791-0cf5174197af · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c725aae5-cdcb-44f2-9cce-da35d0be786b · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Video-LLaVA: Learning United Visual Representation by Alignment Before Projection
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation adbe5ba0-3e8d-4b71-8940-b63ba98a2d1a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7dec7a92-355f-4f77-a4f7-50b6ca51471e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 42
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 55b9226c-4cb1-4cf6-b331-4103cb85e500 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 43
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 06d17157-3251-42da-8043-e289a7c4f7b1 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 44
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d415168c-88d2-4d38-bbdc-e30ffb24f81e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Visual Instruction Tuning
Reference 45
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b236c06c-d809-4d78-85e8-82618aae45c2 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models AudioLDM 2: Learning Holistic Audio Generation with Self-supervised Pretraining
Reference 46
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4edca8fd-f310-411c-aeb6-b66e85c338ce · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Music Understanding LLaMA: Advancing Text-to-Music Generation with Question Answering and Captioning
Reference 47
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 98d8bb8f-6348-421a-afb2-7b281e9f1bfc · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models WavJourney: Compositional Audio Creation with Large Language Models
Reference 48
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ca5a84a2-40c7-4805-9831-6311674156ae · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 49
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 919c5fa8-8e2c-4e6f-bb69-deb37b4425b7 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Macaw-LLM: Multi-Modal Language Modeling with Image, Audio, Video, and Text Integration
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bb54209-c6a5-412d-bf0a-4d788672e223 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 51
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation da59dc1a-1487-4b6a-9753-c1a332670814 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models D.; and Wang, W
Reference 52
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6f88dc4b-655b-4c6d-8a65-a96c73c210bd · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models H.; Lee, H.; Shin, W.; Kim, Y.-H.; and Choi, E
Reference 53
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 39898fb8-8f9c-4568-a9d2-4a67acd28b4f · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Video-ChatGPT: Towards Detailed Video Understanding via Large Vision and Language Models
Reference 54
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c0a99d80-a4bd-4ebc-bd64-0cb2eb7e69b5 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 55
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 08a4e1f3-be52-4eda-ba5b-462b371474f0 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 56
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 24979a9e-48f5-458c-b7a9-0ceb7918634d · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 57
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6a6e6337-2662-435a-a33c-a7994f71352c · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models High-Resolution Image Synthesis with Latent Diffusion Models
Reference 58
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 09797383-3f95-4f8a-8e03-7331d5477f42 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models 3D-GPT: Procedural 3D Modeling with Large Language Models
Reference 59
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3f8a90fa-b2eb-4a95-90ca-708a44c1ea6f · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models SALMONN: Towards Generic Hearing Abilities for Large Language Models
Reference 60
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4c53b14c-d68c-4132-894e-658527a1e44a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Any-to-Any Generation via Composable Diffusion
Reference 61
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 18dc28b7-52c9-47de-8221-1caa30c9d611 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 62
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f26c9e4d-38dd-459d-91ca-0183f3514fe2 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 63
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 47cfa9fa-e367-42fe-80ff-2c5cd8f8067e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Chameleon: Mixed-Modal Early-Fusion Foundation Models
Reference 64
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 82ddab59-89c4-4e42-9fb9-70384bb357aa · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Llama 2: Open Foundation and Fine-Tuned Chat Models
Reference 65
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bdfeabf-4b3b-45b9-bee5-f3c996a5441b · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models N.; Kaiser, .; and Polosukhin, I
Reference 66
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation b5b553fa-f43a-4502-a6c4-debe01289e51 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models AUDIT: Audio Editing by Following Instructions with Latent Diffusion Models
Reference 67
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 326ea139-20b5-4057-994c-1c66eaeec16e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models NExT-GPT: Any-to-Any Multimodal LLM
Reference 68
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23ca5c1a-42a3-4b83-8e68-b5d7ab8094b9 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Unresolved cited work
Reference 69
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fdbf86e9-be1b-4249-ac9a-e77a0b9907c4 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models H.; Fan, L.; and Wu, X.-M
Reference 70
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 87ca801d-b520-4f0f-88bc-98fafc8c6841 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models PointLLM: Empowering Large Language Models to Understand Point Clouds
Reference 71
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 43600cc5-b3bf-4b74-81bd-f52404f90fc6 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models TEAL: Tokenize and Embed ALL for Multi-modal Large Language Models
Reference 72
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cc1d4927-f74e-49f2-9f80-230c6843b3f4 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models A Survey on Multimodal Large Language Models
Reference 73
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation db60093f-5b2c-4ecd-bb64-70d7f5f447ab · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Vis2Mus: Exploring Multimodal Representation Mapping for Controllable Music Generation
Reference 74
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7cc0f05a-2758-4178-a122-9745b03256b9 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Q.; and Artzi, Y
Reference 75
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d38993f5-6468-44a0-bfa4-316f18c18cff · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Q.; and Artzi, Y
Reference 76
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 20c9109e-81e2-4bf3-a976-c96c84423cd9 · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Loop Copilot: Conducting AI Ensembles for Music Generation and Iterative Editing
Reference 77
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 06aeaa55-97f4-41d7-b525-65f38280117e · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models a henb \
Reference 78
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 959a04ba-52b2-48f6-8ac3-de7c8176ae1a · outbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models Transfusion: Predict the Next Token and Diffuse Images with One Multi-Modal Model
Reference 79
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 29b928b4-861d-4832-8757-7750994f7a18 · inbound
Video-Guided Text-to-Music Generation Using Public Domain Movie Collections MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 44
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e0a56444-c904-473d-bc8c-293cddf235e7 · inbound
CoComposer: LLM Multi-agent Collaborative Music Composition MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e22d0eea-1baa-4009-9178-c7d91576540d · inbound
WeaveMuse: An Open Agentic System for Multimodal Music Understanding and Generation MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 15dfa6d9-63a3-407b-b1a1-ef073fbfabcc · inbound
Zero-Effort Image-to-Music Generation: An Interpretable RAG-based VLM Approach MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 7c7d4301-72d4-4c73-8a61-293e4a56bae5 · inbound
Assessing Factual Music Comprehension in Large Audio Language Models MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cb1a608a-ac30-449e-8508-c744c0187524 · inbound
MusTBENCH: Benchmarking and Advancing Temporal Grounding in Music LLMs MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 3
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6f86704a-5c1a-4950-9d1b-daabb4d47bc4 · inbound
JenBridge: Adaptive Long-Form Video Soundtracking across Scene Transitions MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation fb1dc0a5-c76d-40de-b689-6b9ebc421660 · inbound
EntangleCodec: A Unified Discrete Audio Tokenizer via Semantic-Acoustic Entanglement MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 19
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation f9ce5774-f600-457c-93c2-fc45393e91ee · inbound
AudioX-Turbo: A Unified Framework for Efficient Anything-to-Audio Generation MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 4bb19c6a-ee30-4b9a-9400-4353f2b0f2dd · inbound
TORUS: A Test of Rendering-Understanding Self-Coherence for Unified Audio Models MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.