Pith. sign in

Paper Citation Record · LEDGER

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

As of 9 August 2026, this Paper Citation Record lists 35 of 35 outbound references and 1 inbound Pith citation observation for arXiv:2505.18984.

A citation records a reference. It does not transfer a finding from one paper to another.

pith.paper-citation-record.v1
2505.18984 v1

Coverage vector

measured 35 of 35 reference resolution

Typed states for the displayed outbound observations.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:15.094318Z

measured 36 of 36 standing notices

One-hop event checks from named stored sources.

Source: scholarly_work_events, retraction_status_cache, observed 2026-08-09T06:31:02.800959+00:00

measured 1 of 1 inbound itemization

Pith citing papers itemized under the disclosed page cap.

Source: paper_references, paper_reference_links, observed 2026-08-07T14:26:13.455801Z

measured 0 of 1 external citation measurements

A source-named dated measurement, never combined with another source.

Source: pith, observed 2026-08-07T14:26:15.140016Z

Reference resolution

35 of 35 outbound references displayed

  • verified exact0
  • verified fuzzy29
  • unresolved5
  • parse uncertain0
  • malformed identifier0
  • metadata mismatch1

External citation measurements

No source-named external measurement is stored.

Outbound references

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · outbound

This paper cites Self-supervised learning method using multiple sampling strategies for general-purpose audio representation.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:86ca8455453209dc97c7754473a139642db6035198189128f569be2e82de0a68

Observation 1d365e53-61e9-47cd-a079-47686ed747c0 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 2

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.473825Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.522458Z digest=sha256:d7b5874b7ba0eaae9865e56675aa517668ca3eecc40cbc6856faeee2e6fb3664

Observation ecf4b75a-3130-4d09-9856-f5da78ec815b · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 3

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.464221Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.596456Z digest=sha256:86c0bc5d3e1015c3aeeb92baa8a5673f54462f77bd97e51ac3e4e74bb5b3e533

Observation 7d23d873-1395-4f19-95d9-d6c6627e9ec1 · outbound

This paper cites an unresolved cited work.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unresolved cited work

Reference 4

Resolution
unresolved
raw_fallback, observed 2026-08-07T14:26:15.455145Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.717712Z digest=sha256:a3bfc677e47c1c5883c0b53e246cec51da33fa639cb4f6765bfccefdb196fef4

Observation 1f11e76a-f560-4804-a45b-3c62de208ae4 · outbound

This paper cites Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Our method improves the performance of all tasks compared to existing methods in the experiment of the subset of Audioset

Reference 5

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.444845Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.856526Z digest=sha256:3dfcf774ab4136e88bf2f481a7d6746a624922e32ee4e05923a3fd2d004198d6

Observation 6e6c43ae-0e3b-42bf-90e3-daab6052f574 · outbound

This paper cites Zhang, J.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Zhang, J

Reference 6

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.434624Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.964205Z digest=sha256:5035f85631152d30763a7352a2f8e74cb2b6c9611a9ac524d4aeac65aacc0136

Observation 8cf1eabf-32f1-4f57-87bd-db48d3abf6f8 · outbound

This paper cites V oxCeleb2: Deep Speaker Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation V oxCeleb2: Deep Speaker Recognition,

Reference 7

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.424791Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.125586Z digest=sha256:d8914ab02fc1507a35f7b50f328f0566be043d8e69a062fd9c711e6734491e76

Observation b6cc71bd-6a73-4439-9b19-e67a166e88ab · outbound

This paper cites Broadcasted Resid- ual Learning for Efficient Keyword Spotting,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Broadcasted Resid- ual Learning for Efficient Keyword Spotting,

Reference 8

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.414577Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.205583Z digest=sha256:942ac9bc566473ab703df63a0cc1c990a1883a597b1f7ce000ab4d78bb3f1ac0

Observation cc1bcaec-6077-4705-9f60-e47444a973e4 · outbound

This paper cites Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Acous- tic Event Detection Method Using Semi-Supervised Non- Negative Matrix Factorization with Mixtures of Local Dictio- naries,

Reference 9

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.404451Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.292479Z digest=sha256:e4b7fd2634ecedb27fe69e8c37a6bd7714f1a188777f63e2ed9669a3bb01d173

Observation 1d5e790f-2b66-47be-9af5-b4792fcac3de · outbound

This paper cites Crepe: A Convolutional Representation for Pitch Estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A Convolutional Representation for Pitch Estimation,

Reference 10

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.394428Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.421892Z digest=sha256:a4d22bd5e59cc46197944ccc2f63e80906ffa72dae78bee2e9fa1e437058b155

Observation 271aff89-e090-484e-9276-2133fc6a1935 · outbound

This paper cites PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation PANNs: Large-scale pre- trained audio neural networks for audio pattern recognition,

Reference 11

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.383661Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.452385Z digest=sha256:9153dfeb1c6810499a8a4d2cb4a760a443bb4c943d8fafe3211f10b26198a290

Observation eb5f3bcf-9037-43ff-9499-4981d1a64270 · outbound

This paper cites Audio set: An ontology and human-labeled dataset for audio events,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Audio set: An ontology and human-labeled dataset for audio events,

Reference 12

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.374236Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.463942Z digest=sha256:0e3c8998ecbcf070bf4091e9177d90d99331d95667eb531e39976eb77a5516d2

Observation fa39727f-bf1f-453e-8145-6cec55c245f3 · outbound

This paper cites Language models are few-shot learners,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Language models are few-shot learners,

Reference 13

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.364561Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.499846Z digest=sha256:4a7b99c1dbcd8abacdeb3fb68114bf03a2681a37ffca5d86829aa7e0fb720e66

Observation 2a765ba9-705f-44e3-8df0-30025229487f · outbound

This paper cites BERT: Pre-training of deep bidirectional transformers for language understanding,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation BERT: Pre-training of deep bidirectional transformers for language understanding,

Reference 14

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.353992Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.538877Z digest=sha256:96cb78084c2aaf6bf528b09282cc755ae2f7f25112f43fce455c0cd431456cc8

Observation 9dfdb037-9d11-4739-9a84-b2bfcb93c3d6 · outbound

This paper cites van den Oord, Y.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation van den Oord, Y

Reference 15

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.343929Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.576176Z digest=sha256:12e3f84f5c54be6d035c7b1cc26053671c19e1920a48f0dc76d730f6dd2ffe85

Observation e785da60-2e2f-4746-a0d1-1ecf54df335f · outbound

This paper cites Spatiotemporal con- trastive video representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Spatiotemporal con- trastive video representation learning,

Reference 16

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.335105Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.611759Z digest=sha256:07919a70634f82d63666c92b16ebb8e37860f0797369bdc772af445641c028e1

Observation 875a203f-6db4-4d08-a62a-c8a8a6123d2b · outbound

This paper cites Momentum contrast for unsupervised visual representation learning,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Momentum contrast for unsupervised visual representation learning,

Reference 17

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.325640Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.684273Z digest=sha256:93421e891c2d581a64bf67f149b2efd547beaa9790cbc10cdc2029e35351f14e

Observation b4135544-6834-4e57-aa62-4f5d9e78f94d · outbound

This paper cites CURL: Contrastive Unsupervised Representations for Reinforcement Learning.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation CURL: Contrastive Unsupervised Representations for Reinforcement Learning

Reference 18

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:14.769986Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:14.769986Z digest=sha256:b4b3107f35df0f9e872d6ff6cf47372d13a818fc7670d70d415fb2a2ca3c1db2

Observation b692c067-3335-431a-87f7-a97e44c6312b · outbound

This paper cites Data aug- menting contrastive learning of speech representations in the time domain,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Data aug- menting contrastive learning of speech representations in the time domain,

Reference 19

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.315626Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.842785Z digest=sha256:00a764d629546357b969e599de882bb836d2f51c6667eaf2534c157b1ab0dff9

Observation 59b0cce2-ceb8-43a8-91c5-3eee57250880 · outbound

This paper cites Un- supervised pretraining transfers well across languages,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Un- supervised pretraining transfers well across languages,

Reference 20

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.305950Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.922554Z digest=sha256:95a37eceb236916c9bb05a23133afe4a16e059647d100d2d12ac1c811f5361e5

Observation d515b184-3187-4d3d-b367-45ffb719a087 · outbound

This paper cites Vq-wav2vec: Self- supervised learning of discrete speech representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Vq-wav2vec: Self- supervised learning of discrete speech representations,

Reference 21

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.295581Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:14.983591Z digest=sha256:9ff55a2842600a984286c12deb7407eb9009d881d8c08512e185c8f84dab1d4d

Observation 378785f9-1a3f-429e-9821-9f05fcd1a878 · outbound

This paper cites Towards Learning a Uni- versal Non-Semantic Representation of Speech,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Towards Learning a Uni- versal Non-Semantic Representation of Speech,

Reference 22

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.285790Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.019483Z digest=sha256:3a7d21d340a033c2086576da15a72772148328c30d646f40b86519420e401fe4

Observation 7fdff210-619f-4579-9761-6314246933fe · outbound

This paper cites Unsupervised learn- ing of semantic audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Unsupervised learn- ing of semantic audio representations,

Reference 23

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.275794Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.032551Z digest=sha256:1c4bfe611ac41b8b3171d80629a3b7ea3dcb9a95730715c9f8d2fb882accece5

Observation 1c684355-4e1b-416c-98a5-c0a4e876d0fc · outbound

This paper cites Contrastive learning of general-purpose audio representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Contrastive learning of general-purpose audio representations,

Reference 24

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.266366Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.038803Z digest=sha256:74df87b65e27221c0991719ce024afbdda32bc24e8583dbef1859cc9d93b8f21

Observation a1f249f7-88e6-4ee0-ad61-eeb72e973e28 · outbound

This paper cites Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Speech Commands: A Dataset for Limited- V ocabulary Speech Recognition,

Reference 25

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.256554Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.045522Z digest=sha256:a53001b56c5e499972bd9008b02b0b2c8f1c662f61c985818a251121adef2608

Observation 2bde6d46-b343-4d3f-9d5f-97ad3700ed27 · outbound

This paper cites Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Detection and classification of acoustic scenes and events: Outcome of the DCASE 2016 challenge,

Reference 26

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.246276Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.050539Z digest=sha256:f75d660d90fd9b10a254dfcf60ed27ba265d4487ca87a48334b9be80beb924ec

Observation c18c0708-47a8-4b94-8ee3-6b515452dd27 · outbound

This paper cites Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Sound event de- tection in synthetic audio: Analysis of the DCASE 2016 task results,

Reference 27

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.236043Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.058282Z digest=sha256:b11549c2a2ec283db67704464512bc7cb23e7334bf0aec67e030eba1c634f8b7

Observation 85cd9a4a-1503-44f9-929f-a47320bc853d · outbound

This paper cites Neural audio syn- thesis of musical notes with wavenet autoencoders,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Neural audio syn- thesis of musical notes with wavenet autoencoders,

Reference 28

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.225323Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.064430Z digest=sha256:dba8114f16e5f4bea16909189be9dc1609b99dd885fc106bec7c771f9b67e65a

Observation 2ab1526b-ed67-48f7-b827-42d9e438671a · outbound

This paper cites Crepe: A convo- lutional representation for pitch estimation,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Crepe: A convo- lutional representation for pitch estimation,

Reference 29

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.214614Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.071478Z digest=sha256:aa9af397dfa499fb57761261a2404ea1d08c55a6ea42bab8fb9e16b933dfd621

Observation fef3c815-c64d-4fda-8226-fe579a862e91 · outbound

This paper cites EfficientNet: Rethinking model scaling for convolutional neural networks,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation EfficientNet: Rethinking model scaling for convolutional neural networks,

Reference 30

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.204775Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.075322Z digest=sha256:84df1dae280c761748bebaa3e682f2260ba9534c28057b54f9b383d69bbce1aa

Observation ee623c89-baf4-4df3-afc2-a2cc622df993 · outbound

This paper cites Layer Normalization.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Layer Normalization

Reference 31

Resolution
unresolved
no resolver link, observed 2026-08-07T14:26:15.079404Z

Source-reported events for the cited work

Unavailable: canonical work link unavailable.

source=pdf_text observed=2026-08-07T14:26:15.079404Z digest=sha256:1a7818042edbd65322b929268019519d1f115b2abb1b6952ba2bfb561591e7f6

Observation 1a08c4ff-ea39-4cda-afa8-d243f55b969a · outbound

This paper cites Adam: A method for stochastic op- timization,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Adam: A method for stochastic op- timization,

Reference 32

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.193151Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.083555Z digest=sha256:53375884dec498a254880714310e3023dc10d01f286d52443657b48dc6e340f7

Observation d8909262-716a-4fa9-8160-8cc1eeb391d2 · outbound

This paper cites Tagliasacchi, B.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Tagliasacchi, B

Reference 33

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.181985Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.087529Z digest=sha256:afc6d72bcf041e250bd7ea1ae4c8c7316fe85635744d42aab9666000ec1c6065

Observation 9ec2a0ce-34a3-4a42-91b9-06c3e843fc90 · outbound

This paper cites A simple framework for contrastive learning of visual representations,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation A simple framework for contrastive learning of visual representations,

Reference 34

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.170421Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.091212Z digest=sha256:ea7a1e7e408286a2f0a552877293caedda1f7a0d4371e7e7c231e7cfb05ad99c

Observation 0a1819fe-7103-42a9-a291-f07153c04876 · outbound

This paper cites Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Seeing voices and hearing voices: Learning discriminative embeddings us- ing cross-modal self-supervision,

Reference 35

Resolution
verified fuzzy
raw_fallback, observed 2026-08-07T14:26:15.157936Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:15.094318Z digest=sha256:7841b191425e3540bfd6fa358c4c9bf6828683c87b4bb088cd4a12c0e5a17205

Pith citing papers

Observation 48b46bf4-b2c4-416a-ba41-4483c963d721 · inbound

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation cites this paper.

Self-supervised learning method using multiple sampling strategies for general-purpose audio representation Self-supervised learning method using multiple sampling strategies for general-purpose audio representation

Reference 1

Resolution
metadata mismatch
local_arxiv, observed 2026-08-07T14:26:15.146229Z

Source-reported events for the cited work

No event found in the named queried sources as of 2026-08-09T06:31:02.800959+00:00.

source=pdf_text observed=2026-08-07T14:26:13.455801Z digest=sha256:86ca8455453209dc97c7754473a139642db6035198189128f569be2e82de0a68