Pith. sign in

Paper Citation Record · LEDGER

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

As of 16 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2505.18984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18984 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:15.094318Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-16T06:30:59.297886+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:13.455801Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:26:15.140016Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · outbound

This paper cites Self-supervised learning method using multiple sampling strategies for general-purpose audio representation.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:164c80f8d62c7cfc1f4034a39d8057106d81ce03d5689ae3ddb33cad18304d76

Observation 1d365e53-61e9-47cd-a079-47686ed747c0 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.473825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.522458Z digest=sha256:a81f5ced22d571ffc0d01992e4da64e0ef604322520277125f19e648c325d12a

Observation ecf4b75a-3130-4d09-9856-f5da78ec815b · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.464221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.596456Z digest=sha256:d8aea885ad65f80b5059ad01201e391b74fb48ea858f9ec392feaf31bc36d9a3

Observation 7d23d873-1395-4f19-95d9-d6c6627e9ec1 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.455145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.717712Z digest=sha256:139815ecc4e18a3bd778b3c89c928edf888d847f9fdf0839f849d406bbcfef46

Observation 1f11e76a-f560-4804-a45b-3c62de208ae4 · outbound

This paper cites Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.856526Z digest=sha256:734ea2153a1c1aa9145528eb48cb83933e523af786868b929670c57c293f4a6e

Observation 6e6c43ae-0e3b-42bf-90e3-daab6052f574 · outbound

This paper cites Zhang, J.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Zhang, J

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.434624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.964205Z digest=sha256:47a2a57b162b823fd085020c8a058b90d2511b2dd5eceb2f921ecb277a76e277

Observation 8cf1eabf-32f1-4f57-87bd-db48d3abf6f8 · outbound

This paper cites V oxCeleb2: Deep Speaker Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation V oxCeleb2: Deep Speaker Recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.424791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.125586Z digest=sha256:f659344e168d4ab47ac507c713f6dbd7547e77372889640141a27f8e0940bbef

Observation b6cc71bd-6a73-4439-9b19-e67a166e88ab · outbound

This paper cites Broadcasted Resid- ual Learning for Efficient Keyword Spotting,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Broadcasted Resid- ual Learning for Efficient Keyword Spotting,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.414577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.205583Z digest=sha256:5da33dc733a44f78d46eb18614742022814724a01c09ba9fb63ef89f7eec46aa

Observation cc1bcaec-6077-4705-9f60-e47444a973e4 · outbound

This paper cites Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.404451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.292479Z digest=sha256:f65a406dfec9af2f41ff514dd2ae467894c3b1090274dd7c1b70b6d71450120c

Observation 1d5e790f-2b66-47be-9af5-b4792fcac3de · outbound

This paper cites Crepe: A Convolutional Representation for Pitch Estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A Convolutional Representation for Pitch Estimation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.394428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.421892Z digest=sha256:55aa35177a003317f2660978c5336bc92327649b2fd79306f8c0b44ede25c4e7

Observation 271aff89-e090-484e-9276-2133fc6a1935 · outbound

This paper cites PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.383661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.452385Z digest=sha256:ce1c300b75ebc4b7514b46567adba372bc8d08c7f2074adb21dd05b7719164e9

Observation eb5f3bcf-9037-43ff-9499-4981d1a64270 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Audio set: An ontology and human-labeled dataset for audio events,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.374236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.463942Z digest=sha256:10a66a85ddfd34f9b9f8c93d59eb96c4e1415bc061a8e89189f578e396c871d8

Observation fa39727f-bf1f-453e-8145-6cec55c245f3 · outbound

This paper cites Language models are few-shot learners,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.364561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.499846Z digest=sha256:cf7477ef6d21cf537125b2dbae363852903c5a1f31e497cb710216933b8ffd03

Observation 2a765ba9-705f-44e3-8df0-30025229487f · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.353992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.538877Z digest=sha256:fa878410ed8182082026463500ca789b081681fa4cdaac868a30ecaf6872f4ed

Observation 9dfdb037-9d11-4739-9a84-b2bfcb93c3d6 · outbound

This paper cites van den Oord, Y.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation van den Oord, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.343929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.576176Z digest=sha256:5f75048a5f927767bc6b12941fda4269e2c1f0c893401999a7f3f6e502452c4f

Observation e785da60-2e2f-4746-a0d1-1ecf54df335f · outbound

This paper cites Spatiotemporal con- trastive video representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Spatiotemporal con- trastive video representation learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.335105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.611759Z digest=sha256:aa13ba27ac3c9ce75180a017745c6bc31c2b1b66949506b9ca4f298e632dfc9a

Observation 875a203f-6db4-4d08-a62a-c8a8a6123d2b · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Momentum contrast for unsupervised visual representation learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.325640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.684273Z digest=sha256:288bb7083ae22cf58d084fc2d5aa2d04dcac34c959ba1cc9fc5e1abfbd9c8f81

Observation b4135544-6834-4e57-aa62-4f5d9e78f94d · outbound

This paper cites CURL: Contrastive Unsupervised Representations for Reinforcement Learning.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation CURL: Contrastive Unsupervised Representations for Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:14.769986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:14.769986Z digest=sha256:8433e00a14c0703007d18cdfadc0a05c89b9b89647b5f2d8f8a5d64bdf292ff4

Observation b692c067-3335-431a-87f7-a97e44c6312b · outbound

This paper cites Data aug- menting contrastive learning of speech representations in the time domain,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Data aug- menting contrastive learning of speech representations in the time domain,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.315626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.842785Z digest=sha256:b40a863aee9f88950f39719f8f1f65a8486f2965b4ce175b92898e4e0f7c3683

Observation 59b0cce2-ceb8-43a8-91c5-3eee57250880 · outbound

This paper cites Un- supervised pretraining transfers well across languages,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Un- supervised pretraining transfers well across languages,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.305950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.922554Z digest=sha256:53e7072b7655443e4acf4ffec0a42af95725370a25ccaf79b22522a80f34ba23

Observation d515b184-3187-4d3d-b367-45ffb719a087 · outbound

This paper cites Vq-wav2vec: Self- supervised learning of discrete speech representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.295581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:14.983591Z digest=sha256:44291ea071eda0916f1af4d31f717eaf9b60ab443449f6fe3fe289987ce2c876

Observation 378785f9-1a3f-429e-9821-9f05fcd1a878 · outbound

This paper cites Towards Learning a Uni- versal Non-Semantic Representation of Speech,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Towards Learning a Uni- versal Non-Semantic Representation of Speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.285790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.019483Z digest=sha256:129ca65120a1bccc7f25e6099862471464bde29d07b19cef424a5630ea9aeecd

Observation 7fdff210-619f-4579-9761-6314246933fe · outbound

This paper cites Unsupervised learn- ing of semantic audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unsupervised learn- ing of semantic audio representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.275794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.032551Z digest=sha256:c1b247f5f7b68711aa538cdc884eb303b05c3dae0fa75df5642c3848aac8efad

Observation 1c684355-4e1b-416c-98a5-c0a4e876d0fc · outbound

This paper cites Contrastive learning of general-purpose audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Contrastive learning of general-purpose audio representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.266366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.038803Z digest=sha256:f0def87ef8865c769f41f162e765db4bd91c4cf4039a143b21fe88ae8840ff80

Observation a1f249f7-88e6-4ee0-ad61-eeb72e973e28 · outbound

This paper cites Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.256554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.045522Z digest=sha256:169f79954509bc6e58468f06c3dacc3df2e546e414d93946830b049145aa5847

Observation 2bde6d46-b343-4d3f-9d5f-97ad3700ed27 · outbound

This paper cites Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.246276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.050539Z digest=sha256:edc8542721093fc4309384990c055c733e640cd7af6e81039dacf466d120f2b6

Observation c18c0708-47a8-4b94-8ee3-6b515452dd27 · outbound

This paper cites Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.236043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.058282Z digest=sha256:8473e5047333f5683def84be3bff44ad6990c7febf07ad27a4122cd433775beb

Observation 85cd9a4a-1503-44f9-929f-a47320bc853d · outbound

This paper cites Neural audio syn- thesis of musical notes with wavenet autoencoders,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Neural audio syn- thesis of musical notes with wavenet autoencoders,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.225323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.064430Z digest=sha256:0cbba1c1c99dcbfe89ffa698f284aded69c95826966e98dc0741ae17aaccebac

Observation 2ab1526b-ed67-48f7-b827-42d9e438671a · outbound

This paper cites Crepe: A convo- lutional representation for pitch estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A convo- lutional representation for pitch estimation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.214614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.071478Z digest=sha256:2f83e3fa9065738b77e1d81f6ba29ee1e96c34e73f87939770090e56ee7cf9d1

Observation fef3c815-c64d-4fda-8226-fe579a862e91 · outbound

This paper cites EfficientNet: Rethinking model scaling for convolutional neural networks,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation EfficientNet: Rethinking model scaling for convolutional neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.204775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.075322Z digest=sha256:f1b5f3277e4b1e7b9cf73489fd737f1c33227c048092ed24f014aad401a9c033

Observation ee623c89-baf4-4df3-afc2-a2cc622df993 · outbound

This paper cites Layer Normalization.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Layer Normalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:15.079404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:15.079404Z digest=sha256:0a148d2ca3c104b9d2c9a61beec85408ce40bd66a96f8da5a59ca7ce4ee3e33e

Observation 1a08c4ff-ea39-4cda-afa8-d243f55b969a · outbound

This paper cites Adam: A method for stochastic op- timization,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Adam: A method for stochastic op- timization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.193151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.083555Z digest=sha256:674ec8cde4d8e63adbf272eb4757ea3834d963643d21d1d6dd27074090f0dd92

Observation d8909262-716a-4fa9-8160-8cc1eeb391d2 · outbound

This paper cites Tagliasacchi, B.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Tagliasacchi, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.181985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.087529Z digest=sha256:cbfa953518edecb991b44d2b3384120d766a53ad0cfe556473a134da2aab82e7

Observation 9ec2a0ce-34a3-4a42-91b9-06c3e843fc90 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation A simple framework for contrastive learning of visual representations,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.170421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.091212Z digest=sha256:c9a4a207e0feaa44bb30bc04cd50a0fb045daa3bdb30be63381f5597ff9fbf51

Observation 0a1819fe-7103-42a9-a291-f07153c04876 · outbound

This paper cites Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.157936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:15.094318Z digest=sha256:7a1f65880cc77b3cb7f6b1079f604afd23f35c59dd1e1af1bec1c19e5beac09e

Pith citing papers

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · inbound

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation cites this paper.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-16T06:30:59.297886+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:164c80f8d62c7cfc1f4034a39d8057106d81ce03d5689ae3ddb33cad18304d76