Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
Paper Citation Record · LEDGER
As of 14 August 2026, this Paper Citation Record lists 0 of 0 outbound references and 35 inbound Pith citation observations for arXiv:2306.00107.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-12T18:16:20.551474Z
A source-named dated measurement, never combined with another source.
Source: pith, observed 2026-08-05T02:28:24.338817Z
0 of 0 outbound references displayed
External citation measurements
27
pith, observed 2026-08-05T02:28:24.338817Z
No outbound reference observations are available for this paper version.
Observation 648dad47-c0e2-4a1d-8b0e-2897f07b607f · inbound
SALMONN: Towards Generic Hearing Abilities for Large Language Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 31
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation cbb8af77-3443-4945-bdf7-29efd5eff172 · inbound
Do Captioning Metrics Reflect Music Semantic Alignment? MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b5be3d1a-7846-4d7f-a791-0cf5174197af · inbound
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 27369928-abb5-4a40-998e-8848e82acc7d · inbound
Next Token Prediction Towards Multimodal Intelligence: A Comprehensive Survey MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 243
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6fa4d1ab-61c2-4b4f-b250-73617929a07e · inbound
MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8653f544-404e-42e8-8a25-cb98b346e98b · inbound
Towards Unified Music Emotion Recognition across Dimensional and Categorical Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6c9ae4b0-9bca-41e5-a443-ecfa11e4a866 · inbound
Semantic-Aware Interpretable Multimodal Music Auto-Tagging MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d38b7e93-573d-4b25-b831-06d3394bab70 · inbound
Investigating the Reasonable Effectiveness of Speaker Pre-Trained Models and their Synergistic Power for SingMOS Prediction MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e4069a85-3333-437f-b636-a2c8f78ba737 · inbound
Towards Source Attribution of Singing Voice Deepfake with Multimodal Foundation Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 174d3af3-f042-43d5-b345-337e0e8e60ae · inbound
DEL: Dense Event Localization for Multi-modal Audio-Visual Understanding MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7af0fa36-0256-4ed7-9d7d-a9600c00945e · inbound
Scaling Self-Supervised Representation Learning for Symbolic Piano Performance MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 81
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 515c194d-e8bf-4af2-a06d-42134746d7af · inbound
Workflow-Based Evaluation of Music Generation Systems MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 96
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 094ec0fb-8f30-4326-9ff4-c0b2f93231c3 · inbound
OMAR-RQ: Open Music Audio Representation Model Trained with Multi-Feature Masked Token Prediction MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a77f882-56df-428e-a1ab-501cd04a7761 · inbound
MusiScene: Leveraging MU-LLaMA for Scene Imagination and Enhanced Video Background Music Generation MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d3551d87-3533-466c-8827-07c207307ded · inbound
MixAssist: An Audio-Language Dataset for Co-Creative AI Assistance in Music Mixing MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8bb2fca9-b7d9-4b04-8130-381de20ba017 · inbound
Enhancing In-Domain and Out-Domain EmoFake Detection via Cooperative Multilingual Speech Foundation Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ab597606-0dac-4cf4-9ba7-620af391b0f4 · inbound
Affect-aware Cross-Domain Recommendation for Art Therapy via Music Preference Elicitation MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 3e291965-80a8-4978-8dbd-981f289eabac · inbound
Training a Perceptual Model for Evaluating Auditory Similarity in Music Adversarial Attack MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 95a1f3cb-8e23-4042-a4e2-af70ee365a5e · inbound
UniVerse-1: Unified Audio-Video Generation via Stitching of Experts MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 1aff77a5-2bc7-4da2-9f0f-ec9d7ff429c2 · inbound
Exploring How Audio Effects Alter Emotion with Foundation Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 466a9129-a629-42a0-99e4-52db9dc7a8e6 · inbound
Assessing Factual Music Comprehension in Large Audio Language Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 15
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7a0a1d41-2097-49f4-9dcc-a14f3e366eb4 · inbound
Expectation and Acoustic Neural Network Representations Enhance Music Identification from Brain Activity MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 59
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5fad3ff9-fb68-4ccc-a8bb-054687c776bd · inbound
Unsupervised Evaluation of Deep Audio Embeddings for Music Structure Analysis MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cce0c1e3-c4fe-4243-b464-6a55876b087a · inbound
ArtifactNet: Detecting AI-Generated Music via Forensic Residual Physics MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 10
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 9c5f152f-f46a-4ba5-857a-118d3f5f3fb7 · inbound
Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 17
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 82206991-8955-4217-90b9-8f2ba3d275c4 · inbound
Adopting State-of-the-Art Pretrained Audio Representations for Music Recommender Systems MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 41
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation afb7bdbf-5781-4f19-8bd7-74eddb440120 · inbound
ARIA: A Diagnostic Framework for Music Training Data Attribution MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 32
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e292391c-f19b-4486-a2fb-a0b7211ec3ad · inbound
MERIT: Learning Disentangled Music Representations for Audio Similarity MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 14
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation e2d95631-def2-45f2-b5ff-21ce2edf6bf4 · inbound
ImmersiveTTS: Environment-Aware Text-to-Speech with Multimodal Diffusion Transformer and Domain-Specific Representation Alignment MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 39
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 5fbd5ea8-fb63-4e40-b732-416492efbf74 · inbound
From Objectives to Applications: Aligning Architectural Biases in Audio Self-Supervised Learning MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 83
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation 6dfb6e8a-eceb-49a8-86e2-e512d2929b61 · inbound
MADB: A Large-Scale Music Aesthetics Dataset with Professional and Multi-Dimensional Annotations MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 11
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.
Observation d83fa3d8-300e-4e25-b351-e4031f6ae035 · inbound
Dance to Music Generation leveraging Pre-training with Unpaired data and Contrastive Alignment MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 21
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1f2f3da-f1e6-48bc-88ac-3a7e80071f5a · inbound
StemFX: Learning Mixing Style Representations via Autoregressive FX Chain Prediction on Source-Separated Stems MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c819db37-9690-4c80-8a36-6ab584be0a12 · inbound
Do Music Foundation Models Embed Pitch in Helical Structure? MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5edca3c8-d14f-49dd-a1f7-e91746e43062 · inbound
CustomDance: Customized 3D Dance Generation with Coarse-to-Fine Human-Centered Interactive Control MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.