Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T06:12:47.611011Z
Paper Citation Record · LEDGER
As of 22 August 2026, this Paper Citation Record lists 52 of 52 outbound references and 1 inbound Pith citation observation for arXiv:2607.22000.
A citation records a reference. It does not transfer a finding from one paper to another.
Typed states for the displayed outbound observations.
Source: paper_references, paper_reference_links, observed 2026-08-01T06:12:47.611011Z
One-hop event checks from named stored sources.
Source: scholarly_work_events, retraction_status_cache, observed 2026-08-22T06:32:14.747728+00:00
Pith citing papers itemized under the disclosed page cap.
Source: paper_references, paper_reference_links, observed 2026-08-01T06:12:40.962023Z
A source-named dated measurement, never combined with another source.
Source: cited_works
52 of 52 outbound references displayed
External citation measurements
No source-named external measurement is stored.
Observation af5362bb-95d7-4d1d-b047-3ad8e452d148 · outbound
Music-JEPA: Learning a World Model of Sound from Action Music-JEPA: Learning a World Model of Sound from Action
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 593cc8b4-7f03-43c8-9805-c02377f9c65b · outbound
Music-JEPA: Learning a World Model of Sound from Action Unresolved cited work
Reference 2
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9310074f-1716-43ea-aff3-6dd70e1fd63a · outbound
Music-JEPA: Learning a World Model of Sound from Action Unresolved cited work
Reference 3
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8467e091-8626-4951-83e0-390fdb8bb302 · outbound
Music-JEPA: Learning a World Model of Sound from Action We first describe the dataset, training procedure (Section 4.1), and baselines (Section 4.2)
Reference 4
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f5d89169-d190-4730-a375-c8576c0df9aa · outbound
Music-JEPA: Learning a World Model of Sound from Action We demonstrate three key capabilities
Reference 5
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 462bfbe0-92a8-4f21-8846-349753783729 · outbound
Music-JEPA: Learning a World Model of Sound from Action Unresolved cited work
Reference 6
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 23b4aaf0-2ba9-41e0-b687-3f51ddca48c3 · outbound
Music-JEPA: Learning a World Model of Sound from Action Hohwy,The predictive mind
Reference 7
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation e6f91316-3b83-4dae-886a-c0e6c5b8b914 · outbound
Music-JEPA: Learning a World Model of Sound from Action A path towards autonomous machine in- telligence version 0.9.2, 2022-06-27,
Reference 8
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 32630d7b-c87a-456f-94e3-a1723d7dd23b · outbound
Music-JEPA: Learning a World Model of Sound from Action Self-supervised learning from im- ages with a joint-embedding predictive architec- ture,
Reference 9
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b2cc5bb5-4c14-4045-971a-6203a223747f · outbound
Music-JEPA: Learning a World Model of Sound from Action V-JEPA 2: Self-Supervised Video Models Enable Understanding, Prediction and Planning
Reference 10
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0c49ba69-6036-40ff-885a-caa58fac3cbb · outbound
Music-JEPA: Learning a World Model of Sound from Action A-JEPA: Joint-Embedding Predictive Architecture Can Listen
Reference 11
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 163f1b05-700e-4916-8794-06d0bdd0842c · outbound
Music-JEPA: Learning a World Model of Sound from Action Audio-JEPA: Joint-Embedding Predictive Architecture for Audio Representation Learning
Reference 12
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation cbeb8660-15de-45da-aa3e-bcea2e426376 · outbound
Music-JEPA: Learning a World Model of Sound from Action Investigating design choices in joint-embedding predictive architectures for general audio repre- sentation learning,
Reference 13
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation f0abb9ab-8540-4ab7-83f3-d55642f4f899 · outbound
Music-JEPA: Learning a World Model of Sound from Action Stem-jepa: A joint-embedding predictive architecture for musical stem compatibility estimation,
Reference 14
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ff2fffe6-2cd4-4508-a5fb-a3a954a1c02a · outbound
Music-JEPA: Learning a World Model of Sound from Action Emergent musical properties of a transformer under contrastive self- supervised learning,
Reference 15
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation dfe89b82-c828-4a57-9547-4cf97014b229 · outbound
Music-JEPA: Learning a World Model of Sound from Action Contrastive learning of musical representations,
Reference 16
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation c62c7420-ea10-4613-91c8-b1ac2e417055 · outbound
Music-JEPA: Learning a World Model of Sound from Action Masked autoencoders that listen,
Reference 17
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 156b46a4-3d01-4dba-b295-8bd853c367dc · outbound
Music-JEPA: Learning a World Model of Sound from Action MERT: Acoustic Music Understanding Model with Large-Scale Self-supervised Training,
Reference 18
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 94a03cc4-6a09-411f-ab1a-3ab2fcaa2c62 · outbound
Music-JEPA: Learning a World Model of Sound from Action Exponential moving av- erage normalization for self-supervised and semi- supervised learning,
Reference 19
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 246a3af9-8cc6-44cd-ade5-42df6ef4906a · outbound
Music-JEPA: Learning a World Model of Sound from Action Iterative amortized policy optimization,
Reference 20
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a2cf254a-9ada-4b50-b6b7-880df049d17f · outbound
Music-JEPA: Learning a World Model of Sound from Action High-resolution sustain pedal depth estimation from piano audio across room acoustics,
Reference 21
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation f6f71b0d-2307-4902-b745-7082e314e94b · outbound
Music-JEPA: Learning a World Model of Sound from Action Critique of World Model
Reference 22
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8a428ec8-1cc0-4fc3-8f65-79261938f527 · outbound
Music-JEPA: Learning a World Model of Sound from Action Variance-Covariance Regularization Improves Representation Learning
Reference 23
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 13c6c331-710a-4d62-a806-d0e079823859 · outbound
Music-JEPA: Learning a World Model of Sound from Action PAN: A world model for general, interactable, and long-horizon world simulation,
Reference 24
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d89bca9d-453e-47a2-9976-f60f22d1e218 · outbound
Music-JEPA: Learning a World Model of Sound from Action Pointworld: Scaling 3d world models for in-the-wild robotic manipulation,
Reference 25
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation dc765bb4-dc8b-4fba-b9c3-8637aa5e78b8 · outbound
Music-JEPA: Learning a World Model of Sound from Action A theory of cortical responses,
Reference 26
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6a24290f-0b31-4273-850d-d2c635c0b7bc · outbound
Music-JEPA: Learning a World Model of Sound from Action Unresolved cited work
Reference 27
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9b4f6962-1111-4dbe-85bd-83b6011ff5c9 · outbound
Music-JEPA: Learning a World Model of Sound from Action London,Hearing in Time: Psychological Aspects of Musical Meter
Reference 28
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 4bc32673-fe35-43c8-b288-f6313b164ceb · outbound
Music-JEPA: Learning a World Model of Sound from Action Unsupervised disentanglement of content and style via variance- invariance constraints,
Reference 29
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 6514da8e-e5ab-445e-9b71-f24898f1bb08 · outbound
Music-JEPA: Learning a World Model of Sound from Action An image is worth 16x16 words: Transformers for image recognition at scale,
Reference 30
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d05699c8-aae3-4f8d-a96c-d4ac62eab95b · outbound
Music-JEPA: Learning a World Model of Sound from Action Emerging properties in self-supervised vision transformers,
Reference 31
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation d966f9b0-0e15-4ce9-bcc6-cdcef251e637 · outbound
Music-JEPA: Learning a World Model of Sound from Action LeJEPA: Provable and Scalable Self-Supervised Learning Without the Heuristics
Reference 32
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 9038b9cb-4fe1-40c7-8495-8f074d0d2290 · outbound
Music-JEPA: Learning a World Model of Sound from Action Balancing information preservation and disentangle- ment in self-supervised music representation learn- ing,
Reference 33
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 68352f67-509d-4375-9439-a0ade8d4962e · outbound
Music-JEPA: Learning a World Model of Sound from Action Audio barlow twins: Self-supervised audio represen- tation learning,
Reference 34
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 0750932b-266f-4471-9fc8-6cd496787145 · outbound
Music-JEPA: Learning a World Model of Sound from Action PESTO: pitch estimation with self-supervised transposition-equivariant objective,
Reference 35
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation 6be93974-6482-47b7-829c-3fced9f5bbe1 · outbound
Music-JEPA: Learning a World Model of Sound from Action Adam: A Method for Stochastic Optimization
Reference 36
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5c7452a7-6724-491c-bea4-87ad6124a6fd · outbound
Music-JEPA: Learning a World Model of Sound from Action Bootstrap your own latent - a new approach to self-supervised learning,
Reference 37
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 441b6b8d-9ae3-4d50-a9e1-3af44a8d294c · outbound
Music-JEPA: Learning a World Model of Sound from Action ASAP: a dataset of aligned scores and performances for piano transcription,
Reference 38
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 5271a9ff-aece-468e-b047-8c0cb3beeb9c · outbound
Music-JEPA: Learning a World Model of Sound from Action Closing the train-test gap in world models for gradient-based planning,
Reference 39
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 7c58a97f-5e86-4c97-b493-11653bb3657a · outbound
Music-JEPA: Learning a World Model of Sound from Action Tem- poral straightening for latent planning,
Reference 40
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 2fbcf5b4-4e5f-43c5-a457-408cb5b44bc8 · outbound
Music-JEPA: Learning a World Model of Sound from Action Tutorial on amortized optimization for learning to optimize over continuous domains,
Reference 41
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation a9321afb-314f-4ed5-bea3-b7eb62499eda · outbound
Music-JEPA: Learning a World Model of Sound from Action Latent Geometry Beyond Search: Amortizing Planning in World Models
Reference 42
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation b1c09ccd-d0d4-401b-9d21-d04dafb52111 · outbound
Music-JEPA: Learning a World Model of Sound from Action Enabling factorized piano music modeling and generation with the MAE- STRO dataset,
Reference 43
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ede7d27e-11cb-4756-958c-b7b6e30ccba9 · outbound
Music-JEPA: Learning a World Model of Sound from Action Beat this! accurate beat tracking without DBN postprocessing,
Reference 48
Source-reported events for the cited work
No event found in the named queried sources as of 2026-08-22T06:32:14.747728+00:00.
Observation b3cc2071-3d37-4f1b-8568-716014f3b21d · outbound
Music-JEPA: Learning a World Model of Sound from Action madmom: a new Python Audio and Mu- sic Signal Processing Library,
Reference 49
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation eff797c3-b216-4486-bd87-9cf6ac6e3613 · outbound
Music-JEPA: Learning a World Model of Sound from Action High-resolution piano transcription with pedals by regressing onset and offset times,
Reference 50
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ed5d03ec-d920-41ad-aec3-e60e99a8f1d3 · outbound
Music-JEPA: Learning a World Model of Sound from Action Available: https://doi.org/10.1109/ TASLP.2021.3121991
Reference 51
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8581351a-fa08-430a-872f-cf75d78ed15a · outbound
Music-JEPA: Learning a World Model of Sound from Action Scoring time intervals using non-hierarchical transformer for automatic piano transcription,
Reference 52
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation aa78319c-ff7a-49b9-bd36-49d39fa6d2bc · outbound
Music-JEPA: Learning a World Model of Sound from Action Available: http://archives.ismir.net/ ismir2020/paper/000127.pdf
Reference 541
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 28ab99fc-51d1-4b8f-916b-d5a64efef46c · outbound
Music-JEPA: Learning a World Model of Sound from Action [Online]
Reference 2021
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation ea6ddf7b-cf18-4572-b4dc-7a4905e03cbb · outbound
Music-JEPA: Learning a World Model of Sound from Action Variance-Covariance Regularization Improves Representation Learning
Reference 2023
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation 8c384633-909e-4c97-af7b-3a1a2a9b14f7 · outbound
Music-JEPA: Learning a World Model of Sound from Action Critique of World Model
Reference 2025
Source-reported events for the cited work
Unavailable: canonical work link unavailable.
Observation af5362bb-95d7-4d1d-b047-3ad8e452d148 · inbound
Music-JEPA: Learning a World Model of Sound from Action Music-JEPA: Learning a World Model of Sound from Action
Reference 1
Source-reported events for the cited work
Unavailable: canonical work link unavailable.