Pith. sign in

Paper Citation Record · LEDGER

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

As of 14 August 2026, this Paper Citation Record lists 40 of 40 outbound references and 4 inbound Pith citation observations for arXiv:2505.16306.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.16306 v1

Coverage vector

measured 40 of 40 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:56.973279Z

measured 44 of 44 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-14T06:32:32.682623+00:00

measured 4 of 4 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T15:05:55.280579Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: arxiv_reference, observed 2026-07-04T20:50:12.622669Z

Reference resolution

40 of 40 outbound references displayed

  • verified exact1
  • verified fuzzy32
  • unresolved7
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch0

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 9efc2332-6bf7-45d7-aff9-f81f0e0fd96e · outbound

This paper cites One of the most significant advantages of SSL mod- els is their ability to leverage unlabeled audio data, enabling the possibility of training with large-scale datasets.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models One of the most significant advantages of SSL mod- els is their ability to leverage unlabeled audio data, enabling the possibility of training with large-scale datasets

Reference 1

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.921801Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:54.872351Z digest=sha256:5c10f523b70e895175417c39bef36b2962147fd9d4f97bc645e4815da482efb6

Observation 5432a338-3e0a-41d7-958d-62c4f8c5292e · outbound

This paper cites an unresolved cited work.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:06:01.723689Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:54.927510Z digest=sha256:f17293b9ca129b49984fcc95d07c2ea8bc3827be1b961e453d92b77a74b0c8e9

Observation 4e428a11-788d-4cad-b699-519f384a2c3d · outbound

This paper cites Work performed during an internship at Tencent AI Lab.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Work performed during an internship at Tencent AI Lab

Reference 3

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.505537Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.009689Z digest=sha256:5a3413173efaaa89fbf7cff89a711228526e64c4c86cd54d678d8354bbc02f64

Observation 2257135f-cad6-4523-8a82-668e65d577b7 · outbound

This paper cites However, re- search in this area has been hindered by the limitations in data access [6].

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models However, re- search in this area has been hindered by the limitations in data access [6]

Reference 4

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.329849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.074949Z digest=sha256:3c5b1568db3938a47ce9193630e0808ce5d4d9b0d0444426f4a80e6bf845de21

Observation 46bac54d-5dd4-4305-8af0-d80f5213af3d · outbound

This paper cites Layer-wise analysis provides us with a detailed perspective on the relationship between the performance of each specific layer and each specific task.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise analysis provides us with a detailed perspective on the relationship between the performance of each specific layer and each specific task

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.046361Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.453571Z digest=sha256:ce94fc0056e0573ad30ab9c1dc5df87d451ff15074db75e9cb3d66916898f548

Observation 39a9a0a6-60e4-4052-bea7-1b5c7ae6983e · outbound

This paper cites Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.280579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.280579Z digest=sha256:f3dda005bc91057189a817ffd45cc8367d804349c32615a68c2436b878ca2a2a

Observation 6c7e2d90-dd72-4f85-8010-59217d9f5447 · outbound

This paper cites SSL outperforms low-level features.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models SSL outperforms low-level features

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:01.148931Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.370470Z digest=sha256:fd565f68ff55ff03c182c47e4b937aa78d8951eb73139e7846438be451008ed1

Observation 2a2c04ad-e596-4459-afc8-422aafedc336 · outbound

This paper cites wav2vec 2.0: A framework for self-supervised learning of speech representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models wav2vec 2.0: A framework for self-supervised learning of speech representations

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.112543Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.885908Z digest=sha256:dee6f733e02de3b7f5b7de614a5626bd86a032af5022cf9d1f553c5a8970955d

Observation 85d6cc51-2b55-4f3b-885c-427bfc96bd02 · outbound

This paper cites Mert: Acoustic music un- derstanding model with large-scale self-supervised training, 2023.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Mert: Acoustic music un- derstanding model with large-scale self-supervised training, 2023

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.892246Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.539934Z digest=sha256:65c3f11e312a979a38c1405870b1a66e84f7cbba7fbeee5c624ddf306625f453

Observation d8f84f0a-0a37-42db-a3e9-68fc1698e0a4 · outbound

This paper cites an unresolved cited work.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Unresolved cited work

Reference 10

Resolution
unresolved
raw_fallback, observed 2026-08-07T15:06:01.246021Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.178803Z digest=sha256:219a210d6929f58a66d12da210dfbddffb79fe3474713b603d4ae3f0064c8acb

Observation f1945d7b-1e6b-47e4-bd1b-aac6b4dfd18a · outbound

This paper cites A Foundation Model for Music Informatics.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models A Foundation Model for Music Informatics

Reference 11

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.604317Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.604317Z digest=sha256:6b4fc27f1ed186382fe18e058927377fab103fff714663e73155f933b7f6dd71

Observation d70d69ec-f262-4e6d-9acb-7913b0d923e4 · outbound

This paper cites Contrastive learn- ing of musical representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Contrastive learn- ing of musical representations

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.736849Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.644568Z digest=sha256:54fad3dcb6672abfb5424a3dd37cf84fc7cec0cbd8a6cc607656eeb11218ad92

Observation d9a493ef-32e2-400e-9217-934118d4d34b · outbound

This paper cites Codified audio language modeling learns useful representations for music information retrieval.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Codified audio language modeling learns useful representations for music information retrieval

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.556330Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.674618Z digest=sha256:5d352cf578927e400a784d5cd3be625bef0e4f102269e1596f037efcdb7356e5

Observation b138a273-f183-4747-8d7d-8b2a8f188898 · outbound

This paper cites Supervised and Unsupervised Learning of Audio Representations for Music Understanding.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Supervised and Unsupervised Learning of Audio Representations for Music Understanding

Reference 14

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.715841Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.715841Z digest=sha256:d2e2b439e9c50854bd8606d40e45b898d5227cb5eb44ff07e7464a39f9dae6df

Observation c2bf2af4-3915-4dc3-91b7-ecdb7f6e5993 · outbound

This paper cites Freeman, Jessie Wang, Sherry Cai, and KatherineM.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Freeman, Jessie Wang, Sherry Cai, and KatherineM

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.431173Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.773738Z digest=sha256:0b83d1b64fb050e4d85ca4bb839ba7f8ae05a9a40c14e4eb0190ab1a36a53d24

Observation 315b912d-4b93-4224-aa5c-b09749dee9b4 · outbound

This paper cites wav2vec: Unsupervised pre-training for speech recognition.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models wav2vec: Unsupervised pre-training for speech recognition

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:06:00.230779Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.832610Z digest=sha256:21358ae9f66e11d066828800f5b97ce2ff7ce222d350b6de4f66adfbdd1c3058

Observation 25d1bd04-b69d-4344-a163-75ea9161965a · outbound

This paper cites Hubert: Self-supervised speech representation learning by masked prediction of hidden units.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Hubert: Self-supervised speech representation learning by masked prediction of hidden units

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.928652Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:55.922222Z digest=sha256:c6060db6b459f66ef95ac3763878028b9ecc449ad1b3bc4d477b8391bff26325

Observation 218c925e-a4aa-4497-ba42-15e85c138bfd · outbound

This paper cites MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models MuQ: Self-Supervised Music Representation Learning with Mel Residual Vector Quantization

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.961928Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.961928Z digest=sha256:cdfd7db4b21fbc965a6c2c9cf827ce3dd0b324f98f1db4535b5cc3f5c313c694

Observation 93e2fabd-2865-48a5-bb8b-88f70466276e · outbound

This paper cites Self-supervised learning with random-projection quantizer for speech recognition.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Self-supervised learning with random-projection quantizer for speech recognition

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.685321Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.033041Z digest=sha256:86cdf0d42c2d82be64e40e83f01b97b66c2c200160b275dea38df5fa3e002835

Observation b6d2a11f-30f8-4932-8eb3-a91a6585aefb · outbound

This paper cites Why does Self-Supervised Learning for Speech Recognition Benefit Speaker Recognition?.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Why does Self-Supervised Learning for Speech Recognition Benefit Speaker Recognition?

Reference 20

Resolution
verified exact
local_arxiv, observed 2026-08-07T15:05:57.096033Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.086578Z digest=sha256:405001c83679270e5710a28cf634c7d08b7c690b584acbe375fc6f5f40025e81

Observation ad8b857a-d35a-4139-9563-4b7cb6a970f6 · outbound

This paper cites Layer-wise analysis of a self-supervised speech representation model.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise analysis of a self-supervised speech representation model

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.394216Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.135270Z digest=sha256:2323150cd3487fa15df767cc325d02cd4ff6c82601bc62bb0abf9ca56df001f2

Observation c035fea0-94a9-4f72-bed2-aef1cf739023 · outbound

This paper cites Liu, Cheng- I Lai, Haibin Wu, Jiatong Shi, Xuankai Chang, Hsiang-Sheng Tsai, Wen-Chin Huang, Tzu hsun Feng, Po-Han Chi, Yist Y.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Liu, Cheng- I Lai, Haibin Wu, Jiatong Shi, Xuankai Chang, Hsiang-Sheng Tsai, Wen-Chin Huang, Tzu hsun Feng, Po-Han Chi, Yist Y

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.144765Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.183033Z digest=sha256:ed0c14e986ae56cce0649e62528dc37436365b3d8fa1a4bcaf6a432d42646b2f

Observation 77e055ca-dc0f-400b-b57b-2b65d10e45a9 · outbound

This paper cites MARBLE: Music Audio Representation Benchmark for Universal Evaluation.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models MARBLE: Music Audio Representation Benchmark for Universal Evaluation

Reference 23

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:56.227445Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:56.227445Z digest=sha256:c5af70dd7c45eca3671211bf39f25ca2b5d2518d3bf44543d0ef0ff7b8dbba20

Observation e56e578f-8b3a-4696-b2f1-43f5bb32b03e · outbound

This paper cites Bert: Pre-training of deep bidirectional transformers for language understanding.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Bert: Pre-training of deep bidirectional transformers for language understanding

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:59.057122Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.275271Z digest=sha256:745c490cec429c0a52a6aeef116205b4436cbd3cbfc4be48c6e14192c02d23c5

Observation 377bcae6-c52f-4742-8655-64623ff324cc · outbound

This paper cites Advances in residual vector quantization: A review.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Advances in residual vector quantization: A review

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.877723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.329474Z digest=sha256:384aeaa561ba7700be0abde42088ce3c494a831a7d0c54c6800ee819e072335e

Observation 6cf9e074-cb8e-4a1e-8941-39bdc97b284b · outbound

This paper cites The million song dataset.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models The million song dataset

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.692091Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.373393Z digest=sha256:58f46e762832f6d2af070c2114c9b3bb07699fda90ebe0dc122ca1773166bb91

Observation fa72f0d1-d06c-43f9-92cd-631e3e0d6125 · outbound

This paper cites Relations between two sets of variates.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Relations between two sets of variates

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.557201Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.406095Z digest=sha256:95c4d42cbf88fec7f0a728df51d0f4b049f1397982ace3e63655baf6291ba66d

Observation cd1447f1-e243-4277-b2f7-4ec82ae1aeaa · outbound

This paper cites Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Svcca: Singular vector canonical correlation analysis for deep learning dynamics and interpretability

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.385874Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.464439Z digest=sha256:81fb40defde3354393607312a5c7bbf8f2068628478ba076d22a89fc18f0c488

Observation a980500c-e345-4267-828b-e9528db167d4 · outbound

This paper cites A neural network that finds a naturalistic so- lution for the production of muscle activity.Nature Neuroscience, page 1025–1033, Jul 2015.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models A neural network that finds a naturalistic so- lution for the production of muscle activity.Nature Neuroscience, page 1025–1033, Jul 2015

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.196899Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.520896Z digest=sha256:3fdbe47d895a8a49310c5f6be771ed13b9ab8623109261c7d0ea813446f83dc5

Observation 6c0c51e5-f56c-471a-a48f-0438729b505a · outbound

This paper cites Morcos, Maithra Raghu, and Samy Bengio.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Morcos, Maithra Raghu, and Samy Bengio

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:58.054768Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.566040Z digest=sha256:2515c60df8f4ef0e3d1663be53e12924281165c8f5370fbf1c31dacd7179a7c5

Observation 5e80005a-8637-412d-9ac7-f8a0d9613f3f · outbound

This paper cites Neu- ral audio synthesis of musical notes with wavenet autoencoders.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Neu- ral audio synthesis of musical notes with wavenet autoencoders

Reference 31

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.981388Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.603367Z digest=sha256:0b8d9553a6e904c375ba0ed05498b3dde559b6f5983290713eb97ec3e23e4c2f

Observation 9826e2c9-f844-4b75-9456-a9b9e296db94 · outbound

This paper cites V ocalset: A singing voice dataset.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models V ocalset: A singing voice dataset

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.911137Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.648332Z digest=sha256:eb5f564f09f8d5fac1ebf2a214fb8cdff661c988687c45e599c0cf25e006b9dd

Observation 621ffcf2-1421-48d8-850f-a471c50c7a3e · outbound

This paper cites Tzanetakis and P.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Tzanetakis and P

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.834666Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.691957Z digest=sha256:4ffb916ed8c430aa32f20c8594b4204886f85323ff4713cb870484832cab0552

Observation c6b6f3c0-e917-40c0-9004-3552bfc45adc · outbound

This paper cites Caro, Erik M.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Caro, Erik M

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.754059Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.727428Z digest=sha256:0e61d9bd87f23ca42a13d653112befc348eed30f58729fbbff4a13d9769f3125

Observation ab32d531-7d2a-45b1-b59b-7745c758fc49 · outbound

This paper cites The harmonix set: Beats, downbeats, and functional segment annotations of western popular music.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models The harmonix set: Beats, downbeats, and functional segment annotations of western popular music

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.668921Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.770281Z digest=sha256:0d230b86b3addd0d795c65e0f2187b56ed31b1fe5ae88e51193aefbb582c75fd

Observation 60082b96-8d3c-42eb-a152-06b10dfa428b · outbound

This paper cites Evaluation of algorithms using games: The case of mu- sic tagging.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Evaluation of algorithms using games: The case of mu- sic tagging

Reference 36

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.609063Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.818004Z digest=sha256:b7bcb05c9e52ee66b502794aa41770c72f7976861e00f193f0188ac0fff0d550

Observation 341419e3-4c77-4e32-a151-59477eb26d50 · outbound

This paper cites Two data sets for tempo estimation and key detection in electronic dance music annotated from user corrections.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Two data sets for tempo estimation and key detection in electronic dance music annotated from user corrections

Reference 37

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.512753Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.855205Z digest=sha256:9d9a78eecea357b53f4e61b4ce15455fb729c9046e0f3cb926187f79e4d41f79

Observation 4cc1e258-4270-4c12-bc7a-01dcf23f582b · outbound

This paper cites Mir eval: A transparent implementation of common mir metrics.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Mir eval: A transparent implementation of common mir metrics

Reference 38

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.419662Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.902772Z digest=sha256:f060e6f41114031058c6f82fae135d456b16aaead534d2c6667c1db376e5389d

Observation 30a79f16-94e4-41ef-91c5-4ab9dd9e0b5e · outbound

This paper cites Deep contextualized word representations.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Deep contextualized word representations

Reference 39

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.335343Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.938557Z digest=sha256:7308762f9f55b83848764580872fdcb8ccedc5296d90ae1c7d80b9f0483fa994

Observation 999f0691-12ec-4619-915f-d4b0badc3a51 · outbound

This paper cites Comparative layer- wise analysis of self-supervised speech models.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Comparative layer- wise analysis of self-supervised speech models

Reference 40

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T15:05:57.247631Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-08-07T15:05:56.973279Z digest=sha256:7bc60e5103a94b25733cab6cec47a966b3b70429116f8a4326993ba029bbc161

Pith citing papers

Observation 39a9a0a6-60e4-4052-bea7-1b5c7ae6983e · inbound

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models cites this paper.

Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 6

Resolution
unresolved
no resolver link, observed 2026-08-07T15:05:55.280579Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T15:05:55.280579Z digest=sha256:f3dda005bc91057189a817ffd45cc8367d804349c32615a68c2436b878ca2a2a

Observation 1cd830bc-e233-4431-a554-d7fca4c1f6c8 · inbound

Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models cites this paper.

Revisiting Content-Based Music Recommendation: Efficient Feature Aggregation from Large-Scale Music Models Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 57

Resolution
verified exact
arxiv_id, observed 2026-05-16T02:40:30.828825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-05-16T02:39:46.434831Z digest=sha256:5def0e0197d0cb88f283a23ab3795547760ae53d84b536bb054f097d91653abc

Observation f8a44b3f-8f2f-44f0-a274-b5067cebb103 · inbound

Frequency-Aware Self-Supervised Music Representation Learning cites this paper.

Frequency-Aware Self-Supervised Music Representation Learning Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-07-04T20:50:12.625036Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-25T19:25:54.322923Z digest=sha256:766a5176120af10eb70b4f69e506842090cd3d4ace5ca3cc6fa3f8b88dabb55e

Observation 8830d88e-5624-44f1-ab3c-0209d364f729 · inbound

Frequency-Aware Self-Supervised Music Representation Learning cites this paper.

Frequency-Aware Self-Supervised Music Representation Learning Layer-wise Investigation of Large-Scale Self-Supervised Music Representation Models

Reference 38

Resolution
verified exact
arxiv_id, observed 2026-06-30T10:04:35.631723Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-14T06:32:32.682623+00:00.

source=pdf_text observed=2026-06-30T10:04:11.233115Z digest=sha256:8bf3170f81e5501db37f213578a83aa91cdf6997f4e30986a484007838ab683c